CoolFace
8 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01argilla /ultrafeedback-binarized-preferences-cleaned-kto UltraFeedback - Binarized using the Average of Preference Ratings (Cleaned) KTO A KTO signal transformed version of the highly loved UltraFeedback Binarized Preferences Cleaned, the preferred dataset by Argilla to use from now on when fine-tuning on UltraFeedback This dataset represents a new iteration on top of argilla/ultrafeedback-binarized-preferences, and is the recommended and preferred dataset by Argilla to use from now on when fine-tuning on UltraFeedback. Read more about… See the full description on the dataset page: https://huggingface.co/datasets/argilla/ultrafeedback-binarized-preferences-cleaned-kto.texttext-generation100K<n<1M10 likes16k downloads3y agoHugging Face02ITG /PlatVR-ktogated PlatVR KTO Dataset This dataset is part of the EVIDENT framework, designed to enhance the creative process of generating background images for virtual reality sets. Disclaimer The creation process was done using a crowdsourcing methodology. Therefore, the preferences in the data align with the user group that participated in the process (i.e., these are real preference data). Dataset Details This dataset followed a creation process using our fine-tuned model… See the full description on the dataset page: https://huggingface.co/datasets/ITG/PlatVR-kto.texttext-generationn<1K3 likes353 downloads2y agoHugging Face031TuanPham /KTO-mix-14k-vietnamese-groqOriginal dataset: https://huggingface.co/datasets/trl-lib/kto-mix-14k This dataset is a KTO-formatted version of argilla/dpo-mix-7k. Please cite the original dataset if you find it useful in your work. Translated to Vietnamese with context-aware using Groq Llama3.3 70B* via this repo: https://github.com/vTuanpham/Large_dataset_translator. Roughly 9 hours for 2k examples. Usage from datasets import load_dataset kto_mix_14k_vi =… See the full description on the dataset page: https://huggingface.co/datasets/1TuanPham/KTO-mix-14k-vietnamese-groq.textquestion-answering10K<n<100K1 likes177 downloads2y agoHugging Face04argilla /distilabel-capybara-kto-15k-binarized Capybara-KTO 15K binarized A KTO signal transformed version of the highly loved Capybara-DPO 7K binarized, A DPO dataset built with distilabel atop the awesome LDJnr/Capybara This is a preview version to collect feedback from the community. v2 will include the full base dataset and responses from more powerful models. Why KTO? The KTO paper states: KTO matches or exceeds DPO performance at scales from 1B to 30B parameters.1 That is, taking a… See the full description on the dataset page: https://huggingface.co/datasets/argilla/distilabel-capybara-kto-15k-binarized.textquestion-answering10K<n<100K5 likes71 downloads3y agoHugging Face05HanWorld /fair-kto-datasettabulartext-generation10K<n<100K0 likes15 downloads9mo agoHugging Face06Orion-zhen /kto-gutenberg kto-gutenberg This dataset is a merge of jondurbin/gutenberg-dpo-v0.1 and nbeerbower/gutenberg2-dpo. The dataset is designed for kto training. texttext-generation1K<n<10K1 likes7 downloads2y agoHugging Face07sagepond /instructions_kto_v2gated 📘 instructions_kto_v2 Kahneman‑Tversky Optimization (KTO) for language modelsA curated dataset of human instruction-response pairs labeled with binary feedback (desirable/undesirable), designed for training and evaluating human‑aware loss functions like KTO. 🧩 Dataset Format Modality: Text Splits: train: ~240,000 rows test: ~8,400 rows Columns: prompt (string): instruction or user query completion (string): model response label (bool): true =… See the full description on the dataset page: https://huggingface.co/datasets/sagepond/instructions_kto_v2.texttext-generation100K<n<1M1 likes2 downloads4mo agoHugging Face081TuanPham /KTO-mix-14k-vietnamesegatedCompatible with KTO Trainer of trl library. Data was filtered to excluded coding examples, so there is no worry of translation errors. Leave a heart and gud luck, Vietnamese tuners 🤗. texttext-generation10K<n<100K0 likes1 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.