datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
finetune_data
tdro-llm/finetune_data
tDRO: Task-level Distributionally Robust Optimization for Large Language Model-based Dense Retrieval. Guangyuan Ma, Yongliang Ma, Xing Wu, Zhenpeng Su, Ming Zhou and Songlin Hu.
This repo contains all fine-tuning data for Large Language Model-based Dense Retrieval. Please refer to this repo for details to reproduce.
A total of 25 heterogeneous retrieval fine-tuning datasets with Hard Negatives and Deduplication (with test sets) are listed as belows.… See the full description on the dataset page: https://huggingface.co/datasets/tdro-llm/finetune_data.fusion-pairwise-evals-finetuned
Automatic pairwise preference evaluations for: Making, not taking, the Best-of-N
Content
This data contains pairwise automatic win-rate evaluations for the m-ArenaHard-v2.0 benchmark and it compares 2 models against gemini-2.5-flash:
Fusion: is the 111B model finetuned on synthetic data generated with Fusion from 5 teachers
BoN: is the 111B model finetuned on synthetic data generated with BoN from 5 teachers
Each model’s outputs are compared in pairs with the respective… See the full description on the dataset page: https://huggingface.co/datasets/CohereLabs/fusion-pairwise-evals-finetuned.2D_finetuned_filtered_DF_Audio_Embeddingsmath-classification-finetuned-resultsSEA-Instruct-2602-fine-tunedfinetune_datafine-tuned-deepseek2-16b-distillation-datasetmt5-small-finetuned-amazon-en-es_tokenized_datasetsZrov2_FineTunedmt5-small-finetuned-amazon-en-es_books_datasetvector_dataset_roberta-fine-tunedfinetune-dataset-testplancognvs_ckpt_test_time_finetunedbart-finetuned-lyrlen-256-tokens_2024-03-22_runbart-finetuned-lyrlen-512-tokens_2024-03-24_runfinetune-dataset-testplanmT5_multilingual_XLSum-finetuned-wiki-linguabart-finetuned-lyrlen-128-tokens_2024-03-22_runfinetuned_mmlu_ml_output_layer_20_results
Dataset Card for Evaluation run of richmondsin/finetuned-gemma-2-2b-output-layer-20-4k-0
Dataset automatically created during the evaluation run of model richmondsin/finetuned-gemma-2-2b-output-layer-20-4k-0
The dataset is composed of 0 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train"… See the full description on the dataset page: https://huggingface.co/datasets/richmondsin/finetuned_mmlu_ml_output_layer_20_results.FineTuned_indian_foodfinetune-data-for-vision-based-llms5fine_tuned_llama2
Dataset Card for Dataset Name
This dataset card aims to be a base template for new datasets. It has been generated using this raw template.
Dataset Details
Dataset Description
Curated by: [More Information Needed]
Funded by [optional]: [More Information Needed]
Shared by [optional]: [More Information Needed]
Language(s) (NLP): [More Information Needed]
License: [More Information Needed]
Dataset Sources [optional]
Repository: [More… See the full description on the dataset page: https://huggingface.co/datasets/and89/fine_tuned_llama2.nvidia-faq-llm-gemma2-fine-tunedRAMP_finetuned_data_Cora
Dataset Details
This dataset contains finetuning data constructed from the Cora citation network for downstream text-rich graph tasks. It is used for finetuning RAMP (Raw-text Anchored Message Passing), which recasts the LLM as a graph-native aggregation operator on text-rich graphs.
The dataset includes the following files:
finetuned_cora_v1.json — Training set
finetuned_cora_val_v1.json — Validation set
eval_cora_v1.json — Test set
This is a release from our paper LLM as Graph… See the full description on the dataset page: https://huggingface.co/datasets/JJYDXFS/RAMP_finetuned_data_Cora.fine_tune_dataset_testSFT-Fine-Tuned
Qwen3 4B LIMA-Style SFT Variants
Two compact supervised fine-tuning datasets prepared for a Qwen3 4B base model.
The goal is quality over volume: a human-written instruction-following anchor plus verified math/code/reasoning examples.
Variants
Config
Max formatted tokens
Train rows
Validation rows
Total rows
Intended use
qwen3_lima_sft_mix_1024
1024
32,697
500
33,197
Conservative first-pass SFT matching the pretraining sequence length.
qwen3_lima_sft_mix_2048… See the full description on the dataset page: https://huggingface.co/datasets/TeamClaude/SFT-Fine-Tuned.tuninig-dataset_pref_20pct_v2_full-sft-finetuned-stage4-iter86000-v2fine_tune_dataset_STER_idsnvidia-faq-EleutherAI-pythia-1b-fine-tunedimdb_Sentimental_Finetuned_20K
