CoolFace
17 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01cstr /ultrafeedback-binarized-preferences-cleaned-deGerman translation from Mixtral (not the best one, and might contain comments etc, despite it prompted not to, but this is mostly for testing purposes atm) of a first part of the dataset as provided by argilla. tabular1K<n<10K0 likes41 downloads3y agoHugging Face02PJMixers /argilla_ultrafeedback-multi-binarized-quality-preferences-cleaned-PreferenceShareGPTtabularreinforcement-learning100K<n<1M1 likes28 downloads2y agoHugging Face03sumya123 /students-subject-preferences Students' Subject Preferences A small survey-style dataset recording which school subjects five students like and dislike. Each row is one student: their ID, the subjects they named as favorites, and the subjects they named as least favorites. Subject names are in Mongolian Cyrillic. Files File Rows Description data/train.jsonl 5 One JSON object per student Schema Column Type Description student_id int Student identifier… See the full description on the dataset page: https://huggingface.co/datasets/sumya123/students-subject-preferences.textn<1K0 likes26 downloads12d agoHugging Face04PJMixers /argilla_distilabel-math-preference-dpo-PreferenceShareGPTtabularreinforcement-learning1K<n<10K0 likes21 downloads2y agoHugging Face05cstr /ultrafeedback-binarized-preferences-cleaned-de-2tabularn<1K0 likes17 downloads3y agoHugging Face06PJMixers /Magpie-Align_Magpie-Pro-DPO-200K-PreferenceShareGPTtabularreinforcement-learning100K<n<1M0 likes16 downloads2y agoHugging Face07PJMixers /argilla_Capybara-Preferences-PreferenceShareGPTtabularreinforcement-learning10K<n<100K0 likes13 downloads2y agoHugging Face08PJMixers /argilla_ultrafeedback-binarized-preferences-cleaned-PreferenceShareGPTtabularreinforcement-learning10K<n<100K1 likes9 downloads2y agoHugging Face09jacobavalanchel /assignment4-pairrm-preferences Assignment 4 Preference Dataset Generated from GAIR/lima instructions with Qwen2.5-7B-Instruct and ranked with PairRM. tabularn<1K0 likes8 downloads5mo agoHugging Face10PJMixers /argilla_Capybara-Preferences-Filtered-PreferenceShareGPTtabularreinforcement-learning10K<n<100K1 likes7 downloads2y agoHugging Face11PJMixers /argilla_ultrafeedback-multi-binarized-preferences-cleaned-PreferenceShareGPTtabularreinforcement-learning100K<n<1M1 likes5 downloads2y agoHugging Face12nancy925 /lima-qwen25-7b-pairrm-preferences LIMA Qwen2.5-7B PairRM Preference Dataset This dataset contains 50 PairRM-ranked preference examples constructed from LIMA instructions. For each instruction, Qwen2.5-7B-Instruct generated 5 candidate responses, and PairRM was used to select chosen and rejected responses for DPO training. Dataset Details Source instructions: LIMA Number of instructions: 50 Base model for response generation: Qwen2.5-7B-Instruct Number of candidate responses per instruction: 5 Preference… See the full description on the dataset page: https://huggingface.co/datasets/nancy925/lima-qwen25-7b-pairrm-preferences.tabularn<1K0 likes3 downloads5mo agoHugging Face13SetonLiang2 /assignment4-pairrm-preferencestabularn<1K0 likes3 downloads5mo agoHugging Face14qducnguyen /ultrafeedback_binarized_preferences_cleanedgatedtabular10K<n<100K0 likes1 downloads2y agoHugging Face15PJMixers /Intel_orca_dpo_pairs-ArmoRM-ReRanked-PreferenceShareGPTgatedtabular1K<n<10K0 likes1 downloads2y agoHugging Face16Jellysillyfish /llm-judge-preferencestabularn<1K0 likes1 downloads1y agoHugging Face17Jellysillyfish /pairm-preferencestabularn<1K0 likes1 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.