preference-data
gemma-2-2b-it-preference_dataset_mixture2_and_safe_pku-Preferenceopen-image-preferences-v1-flux-dev-loragemma-3-1b-it-preference_dataset_mixture2_and_safe_pku-Preferencegemma-3-1b-it-preference_dataset_mixture2_and_safe_pku-Preference1Llama_3_2_1B_Filler_context_based_training_data_20251105_172125_DPO_preference_finetuned_dec10Llama_3_2_1B_Filler_preference_based_training_data_20251210_080644_DPOorpo_run_yash_lora_44_preference_data_v10_for_orpogemma-2-2b-it-preference_dataset_mixture2_and_safe_pku-Preference
tulu-2.5-preference-data
Tulu 2.5 Preference Data
This dataset contains the preference dataset splits used to train the models described in Unpacking DPO and PPO: Disentangling Best Practices for Learning from Preference Feedback.
We cleaned and formatted all datasets to be in the same format.
This means some splits may differ from their original format.
To see the code used for creating most splits, see here.
If you only wish to download one dataset, each dataset exists in one file under the data/… See the full description on the dataset page: https://huggingface.co/datasets/allenai/tulu-2.5-preference-data.700k_Human_Preference_Dataset_FLUX_SD3_MJ_DALLE3
NOTE: A newer version of this dataset is available Imagen3_Flux1.1_Flux1_SD3_MJ_Dalle_Human_Preference_Dataset
Rapidata Image Generation Preference Dataset
This Dataset is a 1/3 of a 2M+ human annotation dataset that was split into three modalities: Preference, Coherence, Text-to-Image Alignment.
Link to the Coherence dataset: https://huggingface.co/datasets/Rapidata/Flux_SD3_MJ_Dalle_Human_Coherence_Dataset
Link to the Text-2-Image Alignment dataset:… See the full description on the dataset page: https://huggingface.co/datasets/Rapidata/700k_Human_Preference_Dataset_FLUX_SD3_MJ_DALLE3.preference-dataset-qwen3preference-datasets-tulupreference_data_llama_factory_wo_checklist
Dataset Card for "preference_data_llama_factory_wo_checklist"
More Information needed
preference_data_llama_factory_len_8k
