CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01argilla /ultrafeedback-binarized-preferences-cleaned UltraFeedback - Binarized using the Average of Preference Ratings (Cleaned) This dataset represents a new iteration on top of argilla/ultrafeedback-binarized-preferences, and is the recommended and preferred dataset by Argilla to use from now on when fine-tuning on UltraFeedback. Read more about Argilla's approach towards UltraFeedback binarization at argilla/ultrafeedback-binarized-preferences/README.md. Differences with argilla/ultrafeedback-binarized-preferences… See the full description on the dataset page: https://huggingface.co/datasets/argilla/ultrafeedback-binarized-preferences-cleaned.tabulartext-generation10K<n<100K165 likes27k downloads3y agoHugging Face02arcee-ai /distilabel-intel-orca-dpo-pairs-binarizedThis is the binarized version of distilabel Orca Pairs for DPO and ORPO. Reference: https://huggingface.co/datasets/argilla/distilabel-intel-orca-dpo-pairs?row=0 text10K<n<100K1 likes24k downloads2y agoHugging Face03HuggingFaceH4 /ultrafeedback_binarized Dataset Card for UltraFeedback Binarized Dataset Description This is a pre-processed version of the UltraFeedback dataset and was used to train Zephyr-7Β-β, a state of the art chat model at the 7B parameter scale. The original UltraFeedback dataset consists of 64k prompts, where each prompt is accompanied with four model completions from a wide variety of open and proprietary models. GPT-4 is then used to assign a score to each completion, along criteria like… See the full description on the dataset page: https://huggingface.co/datasets/HuggingFaceH4/ultrafeedback_binarized.tabulartext-generation100K<n<1M348 likes23k downloads2y agoHugging Face04argilla /distilabel-capybara-dpo-7k-binarized Capybara-DPO 7K binarized A DPO dataset built with distilabel atop the awesome LDJnr/Capybara This is a preview version to collect feedback from the community. v2 will include the full base dataset and responses from more powerful models. Why? Multi-turn dialogue data is key to fine-tune capable chat models. Multi-turn preference data has been used by the most relevant RLHF works (Anthropic, Meta Llama2, etc.). Unfortunately, there are very few… See the full description on the dataset page: https://huggingface.co/datasets/argilla/distilabel-capybara-dpo-7k-binarized.tabularquestion-answering1K<n<10K184 likes23k downloads2y agoHugging Face05argilla /ultrafeedback-binarized-preferences-cleaned-kto UltraFeedback - Binarized using the Average of Preference Ratings (Cleaned) KTO A KTO signal transformed version of the highly loved UltraFeedback Binarized Preferences Cleaned, the preferred dataset by Argilla to use from now on when fine-tuning on UltraFeedback This dataset represents a new iteration on top of argilla/ultrafeedback-binarized-preferences, and is the recommended and preferred dataset by Argilla to use from now on when fine-tuning on UltraFeedback. Read more about… See the full description on the dataset page: https://huggingface.co/datasets/argilla/ultrafeedback-binarized-preferences-cleaned-kto.texttext-generation100K<n<1M10 likes16k downloads3y agoHugging Face06trl-lib /ultrafeedback_binarizedtabular10K<n<100K29 likes3.7k downloads2y agoHugging Face07Rendra86318 /vision-feedback-mix-binarized Dataset Card for Vision-Feedback-Mix-Binarized Introduction This dataset aims to provide large-scale vision feedback data. It is a combination of the following high-quality vision feedback datasets: zhiqings/LLaVA-Human-Preference-10K: 9,422 samples MMInstruction/VLFeedback: 80,258 samples YiyangAiLab/POVID_preference_data_for_VLLMs: 17,184 samples openbmb/RLHF-V-Dataset: 5,733 samples openbmb/RLAIF-V-Dataset: 83,132 samples We also offer a cleaned version in… See the full description on the dataset page: https://huggingface.co/datasets/Rendra86318/vision-feedback-mix-binarized.image100K<n<1M0 likes1.8k downloads9mo agoHugging Face08allenai /ultrafeedback_binarized_cleaned Dataset Card for "ultrafeedback_binarized_cleaned" Update 1/12/2023: I've removed examples identified as faulty by Argilla - see their awesome work for more details. This is a version of the UltraFeedback binarized dataset but with TruthfulQA prompts removed and source annotations added (so you can filter out samples from different sources yourself if you want!). Please see the binarized dataset card for more information, or the original UltraFeedback dataset card. tabular100K<n<1M72 likes1.7k downloads3y agoHugging Face09euclaise /WritingPrompts_binarizedWritingPrompts_preferences, but processed like SHP text100K<n<1M2 likes551 downloads3y agoHugging Face10Jennny /ultrafeedback_binarized_honesty_prefstabular10K<n<100K0 likes410 downloads2y agoHugging Face11argilla /ultrafeedback-binarized-preferences Ultrafeedback binarized dataset using the mean of preference ratings Introduction This dataset contains the result of curation work performed by Argilla (using Argilla 😃). After visually browsing around some examples using the sort and filter feature of Argilla (sort by highest rating for chosen responses), we noticed a strong mismatch between the overall_score in the original UF dataset (and the Zephyr train_prefs dataset) and the quality of the chosen response. By… See the full description on the dataset page: https://huggingface.co/datasets/argilla/ultrafeedback-binarized-preferences.tabular10K<n<100K84 likes408 downloads3y agoHugging Face12Jennny /ultrafeedback_binarized_truthfulness_prefstabular10K<n<100K0 likes377 downloads2y agoHugging Face13data-is-better-together /open-image-preferences-v1-binarized Open Image Preferences Prompt: Anime-style concept art of a Mayan Quetzalcoatl biomutant, dystopian world, vibrant colors, 4K. Image 1 Image 2 Prompt: 8-bit pixel art of a blue knight, green car, and glacier landscape in Norway, fantasy style, colorful and detailed. Image 1… See the full description on the dataset page: https://huggingface.co/datasets/data-is-better-together/open-image-preferences-v1-binarized.image1K<n<10K59 likes306 downloads2y agoHugging Face14rshwndsz /nectar-cleaned-r7-binarizedtext1M<n<10M0 likes260 downloads1y agoHugging Face15rshwndsz /nectar-cleaned-r6-binarizedtext1M<n<10M0 likes235 downloads1y agoHugging Face16llamafactory /ultrafeedback_binarizedBorrowed from: https://huggingface.co/datasets/HuggingFaceH4/ultrafeedback_binarized You can use it in LLaMA Factory by specifying dataset: ultrafeedback. text10K<n<100K0 likes230 downloads2y agoHugging Face17jan-hq /Buzz_binarizedtext10M<n<100M0 likes175 downloads2y agoHugging Face18argilla /ultrafeedback-multi-binarized-preferences-cleaned UltraFeedback - Multi-Binarized using the Average of Preference Ratings (Cleaned) This dataset represents a new iteration on top of argilla/ultrafeedback-binarized-preferences-cleaned, and has been created to explore whether DPO fine-tuning with more than one rejection per chosen response helps the model perform better in the AlpacaEval, MT-Bench, and LM Eval Harness benchmarks. Read more about Argilla's approach towards UltraFeedback binarization at… See the full description on the dataset page: https://huggingface.co/datasets/argilla/ultrafeedback-multi-binarized-preferences-cleaned.tabulartext-generation100K<n<1M7 likes160 downloads3y agoHugging Face19jan-hq /instruction-speech-binarized-male-cleanedtext100K<n<1M0 likes158 downloads2y agoHugging Face20BarraHome /ultrafeedback_binarizedtabular100K<n<1M1 likes138 downloads3y agoHugging Face21ummagumm-a /ultrafeedback_binarized_all_pairstabular100K<n<1M0 likes119 downloads2y agoHugging Face22jan-hq /distilabel_dpo_pairs_binarized Dataset Card for "distilabel_dpo_pairs_binarized" More Information needed text10K<n<100K0 likes113 downloads3y agoHugging Face23rshwndsz /nectar-cleaned-r5-binarizedtext1M<n<10M0 likes113 downloads1y agoHugging Face24TheHassanSaud /Ultra_Feedback_Binarized_Preprocessedtext10K<n<100K0 likes112 downloads4mo agoHugging Face25maywell /ko_Ultrafeedback_binarized maywell/ko_Ultrafeedback_binarized 본 데이터는 Synatra-7B-Translation 모델을 통해 Ultrafeedback_binarized를 번역하고 정제한 데이터셋입니다. 해당 데이터를 직접적으로 상업적으로 사용하는 것은 허용되지 않으며, 데이터를 이용하여 훈련된 모델에 대한 상업적 사용은 허용됩니다. 아직 완벽히 정제되지는 않았으며, 오류나 수정사항에 대해서는 PR 부탁드립니다. text10K<n<100K38 likes109 downloads3y agoHugging Face26jan-hq /dolphin_binarized Dataset Card for "dolphin_binarized" More Information needed text100K<n<1M0 likes106 downloads3y agoHugging Face27vwxyzjn /ultrafeedback_binarized_1708035667 Dataset Card for "ultrafeedback_binarized_1708035667" More Information needed tabular10K<n<100K0 likes104 downloads3y agoHugging Face28DtYXs /llama3.2-3b-ultrafeedback-armorm-binarizedThis repository is associated with the paper Pre-DPO: Improving Data Utilization in Direct Preference Optimization Using a Guiding Reference Model. Code: https://github.com/DtYXs/Pre-DPO texttext-generation10K<n<100K0 likes104 downloads1y agoHugging Face29jan-hq /openhermes_dpo_binarized Dataset Card for "openhermes_dpo_binarized" More Information needed text100K<n<1M1 likes102 downloads3y agoHugging Face30trl-internal-testing /tiny-ultrafeedback-binarizedfrom datasets import load_dataset push_to_hub = True def is_small(example): small_prompt = len(example["chosen"][0]["content"]) < 100 small_chosen = len(example["chosen"][1]["content"]) < 100 small_rejected = len(example["rejected"][1]["content"]) < 100 return small_prompt and small_chosen and small_rejected if __name__ == "__main__": dataset = load_dataset("trl-lib/ultrafeedback_binarized") dataset = dataset.filter(is_small) if push_to_hub:… See the full description on the dataset page: https://huggingface.co/datasets/trl-internal-testing/tiny-ultrafeedback-binarized.tabularn<1K2 likes102 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.