CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01Eehan /eval-imdb-drpo-134567-dpo-5000tabular10K<n<100K0 likes42 downloads1y agoHugging Face02Eehan /eval-tldr-dpo-drpo-0.9tmp-sft-1000text1K<n<10K0 likes36 downloads1y agoHugging Face03Kyleyee /train_data_tldr_for_drpo TL;DR Dataset for DRPO Summary The TL;DR dataset is a processed version of Reddit posts Data Structure Columns: "prompt": The unabridged Reddit post. "a1": A summary of the post. "a2": An alternative summary of the post. "rank": The rank of the summary, where 1 indicates the first summary is preferred and 0 indicates the second summary is preferred. This structure enables models to learn the relationship between detailed content and its abbreviated form… See the full description on the dataset page: https://huggingface.co/datasets/Kyleyee/train_data_tldr_for_drpo.text100K<n<1M0 likes34 downloads1y agoHugging Face04michaelvu1207 /narrative-arc-tp-drpotextn<1K0 likes33 downloads2y agoHugging Face05Kyleyee /train_data_hh_for_drpo HH-RLHF-Helpful-Base Dataset Summary The HH-RLHF-Helpful-Base dataset is a processed version of Anthropic's HH-RLHF dataset, specifically curated to train models using the TRL library for preference learning and alignment tasks. It contains pairs of text samples, each labeled as either "chosen" or "rejected," based on human preferences regarding the helpfulness of the responses. This dataset enables models to learn human preferences in generating helpful responses… See the full description on the dataset page: https://huggingface.co/datasets/Kyleyee/train_data_hh_for_drpo.text10K<n<100K0 likes31 downloads1y agoHugging Face06Eehan /eval-tldr-dpo-ppo-drpo-dm-sft-1000-cut2text1K<n<10K0 likes24 downloads1y agoHugging Face07Eehan /eval-tldr-dpo-ppo-drpo-dm-sft-1000text1K<n<10K0 likes21 downloads1y agoHugging Face08ihughes15234 /kp_cfr_drpo_1200_v2text1K<n<10K0 likes20 downloads2y agoHugging Face09drpolygon /semiautomatic-aestheticsimagen<1K0 likes20 downloads10mo agoHugging Face10august66 /drpo_hh_qwen2.5_1.5b_with_ref_btpreftabular10K<n<100K0 likes20 downloads1y agoHugging Face11ihughes15234 /kp_cfr_drpo_12000text10K<n<100K0 likes18 downloads2y agoHugging Face12Eehan /eval-imdb-drpo-3-drpo-4-1000tabular1K<n<10K0 likes17 downloads1y agoHugging Face13Kyleyee /tldr_test_tiny_data_drpo TL;DR Dataset Summary The TL;DR dataset is a processed version of Reddit posts, specifically curated to train models using the TRL library for summarization tasks. It leverages the common practice on Reddit where users append "TL;DR" (Too Long; Didn't Read) summaries to lengthy posts, providing a rich source of paired text data for training summarization models. Data Structure Format: Conversational Type: Preference Columns: "prompt": The user query.… See the full description on the dataset page: https://huggingface.co/datasets/Kyleyee/tldr_test_tiny_data_drpo.textn<1K0 likes15 downloads2y agoHugging Face14august66 /drpo_ultrafeedback_qwen2.5-1.5b_first_iter_20ktext10K<n<100K0 likes14 downloads1y agoHugging Face15august66 /drpo_hh_qwen2.5_1.5b_with_ref_prob_vllm_convtabular10K<n<100K0 likes13 downloads8mo agoHugging Face16Eehan /eval-imdb-drpo-1-3-4-dpo-1000tabular1K<n<10K0 likes12 downloads1y agoHugging Face17Eehan /eval-tldr-dpo-drpo-0.75tmp-sft-ppo-1000text10K<n<100K0 likes11 downloads1y agoHugging Face18Eehan /eval-tldr-dpo-ppo-drpo-dm-sft-1000-cuttext1K<n<10K0 likes11 downloads1y agoHugging Face19august66 /drpo_hh_qwen2.5_1.5b_with_ref_prob_sampledtext10K<n<100K0 likes11 downloads8mo agoHugging Face20ihughes15234 /kp_cfr_drpo_12000_nonadversarialtext10K<n<100K0 likes10 downloads2y agoHugging Face21august66 /drpo_hh_qwen2.5_1.5btext10K<n<100K0 likes10 downloads1y agoHugging Face22drpointbreak /bankingtextn<1K0 likes9 downloads2y agoHugging Face23august66 /DRPO_data_from_ultrafeed_new_templatetext10K<n<100K0 likes8 downloads1y agoHugging Face24august66 /hh_helpfulness_drpo_from_sfttext10K<n<100K0 likes8 downloads7mo agoHugging Face25august66 /DRPO_data_from_ultrafeedtext10K<n<100K0 likes7 downloads1y agoHugging Face26august66 /drpo_ultrafeedback_qwen2.5-1.5b-1text1K<n<10K0 likes7 downloads1y agoHugging Face27Eehan /eval-imdb-drpo-sft-dpo-5000tabular10K<n<100K0 likes6 downloads1y agoHugging Face28august66 /DRPO_first_itertext10K<n<100K0 likes5 downloads1y agoHugging Face29august66 /drpo_ultrafeedback_qwen2.5-1.5b-2text1K<n<10K0 likes5 downloads1y agoHugging Face30august66 /drpo_hh_qwen2.5_1.5b_with_ref_prob_vllmtabular10K<n<100K0 likes5 downloads8mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.