CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01argilla /ultrafeedback-binarized-preferences-cleaned-kto UltraFeedback - Binarized using the Average of Preference Ratings (Cleaned) KTO A KTO signal transformed version of the highly loved UltraFeedback Binarized Preferences Cleaned, the preferred dataset by Argilla to use from now on when fine-tuning on UltraFeedback This dataset represents a new iteration on top of argilla/ultrafeedback-binarized-preferences, and is the recommended and preferred dataset by Argilla to use from now on when fine-tuning on UltraFeedback. Read more about… See the full description on the dataset page: https://huggingface.co/datasets/argilla/ultrafeedback-binarized-preferences-cleaned-kto.texttext-generation100K<n<1M10 likes16k downloads3y agoHugging Face02ITG /PlatVR-ktogated PlatVR KTO Dataset This dataset is part of the EVIDENT framework, designed to enhance the creative process of generating background images for virtual reality sets. Disclaimer The creation process was done using a crowdsourcing methodology. Therefore, the preferences in the data align with the user group that participated in the process (i.e., these are real preference data). Dataset Details This dataset followed a creation process using our fine-tuned model… See the full description on the dataset page: https://huggingface.co/datasets/ITG/PlatVR-kto.texttext-generationn<1K3 likes636 downloads2y agoHugging Face03trl-lib /kto-mix-14k Dataset card for trl-lib/kto-mix-14k This dataset is a KTO-formatted version of argilla/dpo-mix-7k. Please cite the original dataset if you find it useful in your work. text10K<n<100K9 likes250 downloads3y agoHugging Face041TuanPham /KTO-mix-14k-vietnamese-groqOriginal dataset: https://huggingface.co/datasets/trl-lib/kto-mix-14k This dataset is a KTO-formatted version of argilla/dpo-mix-7k. Please cite the original dataset if you find it useful in your work. Translated to Vietnamese with context-aware using Groq Llama3.3 70B* via this repo: https://github.com/vTuanpham/Large_dataset_translator. Roughly 9 hours for 2k examples. Usage from datasets import load_dataset kto_mix_14k_vi =… See the full description on the dataset page: https://huggingface.co/datasets/1TuanPham/KTO-mix-14k-vietnamese-groq.textquestion-answering10K<n<100K1 likes176 downloads2y agoHugging Face05auditing-agents /kto_redteaming_data_for_secret_loyaltytext1K<n<10K0 likes167 downloads6mo agoHugging Face06open-llm-leaderboard-old /details_dreamgen__llama3-8b-instruct-align-test1-kto0 likes125 downloads2y agoHugging Face07open-llm-leaderboard-old /details_ContextualAI__Contextual_KTO_Mistral_PairRM Dataset Card for Evaluation run of ContextualAI/Contextual_KTO_Mistral_PairRM Dataset automatically created during the evaluation run of model ContextualAI/Contextual_KTO_Mistral_PairRM on the Open LLM Leaderboard. The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train"… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_ContextualAI__Contextual_KTO_Mistral_PairRM.0 likes98 downloads3y agoHugging Face08open-llm-leaderboard-old /details_DatPySci__pythia-1b-kto-iter0 Dataset Card for Evaluation run of DatPySci/pythia-1b-kto-iter0 Dataset automatically created during the evaluation run of model DatPySci/pythia-1b-kto-iter0 on the Open LLM Leaderboard. The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_DatPySci__pythia-1b-kto-iter0.0 likes95 downloads3y agoHugging Face09ktoufiquee /NC-SentNoBThis is a multilabel dataset used for Noise Identification purpose in the paper "A Comparative Analysis of Noise Reduction Methods in Sentiment Analysis on Noisy Bangla Texts" accepted in 2024 The 9th Workshop on Noisy and User-generated Text (W-NUT) collocated with EACL 2024. Annotated by 4 native Bangla speakers with 90% trustworthiness score. Fleiss' Kappa Score: 0.69 Definition of noise categories Type Definition Local Word Any regional words even if there is a… See the full description on the dataset page: https://huggingface.co/datasets/ktoufiquee/NC-SentNoB.tabulartext-classification10K<n<100K1 likes86 downloads3y agoHugging Face10PJMixers-Dev /HailMary-v0.2-KTO-Public Details This only contains the sets which are not private. This is also an experiment, so don't expect anything that good. The idea is to just take existing datasets which seem high quality and then generate a bad response for every model turn. If you have suggestions for improving this idea, I'm all ears. Refer to the original linked datasets for licenses as I add no further restrictions to them. Rejected Generations… See the full description on the dataset page: https://huggingface.co/datasets/PJMixers-Dev/HailMary-v0.2-KTO-Public.textreinforcement-learning100K<n<1M0 likes85 downloads2y agoHugging Face11nyu-dice-lab /lm-eval-results-princeton-nlp-Llama-3-Base-8B-SFT-KTO-private Dataset Card for Evaluation run of princeton-nlp/Llama-3-Base-8B-SFT-KTO Dataset automatically created during the evaluation run of model princeton-nlp/Llama-3-Base-8B-SFT-KTO The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 7 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-princeton-nlp-Llama-3-Base-8B-SFT-KTO-private.tabular100K<n<1M0 likes82 downloads2y agoHugging Face12argilla /distilabel-capybara-kto-15k-binarized Capybara-KTO 15K binarized A KTO signal transformed version of the highly loved Capybara-DPO 7K binarized, A DPO dataset built with distilabel atop the awesome LDJnr/Capybara This is a preview version to collect feedback from the community. v2 will include the full base dataset and responses from more powerful models. Why KTO? The KTO paper states: KTO matches or exceeds DPO performance at scales from 1B to 30B parameters.1 That is, taking a… See the full description on the dataset page: https://huggingface.co/datasets/argilla/distilabel-capybara-kto-15k-binarized.textquestion-answering10K<n<100K5 likes78 downloads3y agoHugging Face13ShreyashDhoot /KTO-clean-finalimage10K<n<100K0 likes74 downloads5mo agoHugging Face14argilla /kto-mix-15k Argilla KTO Mix 15K Dataset A KTO signal transformed version of the highly loved Argilla DPO Mix, which is small cocktail combining DPO datasets built by Argilla with distilabel. The goal of this dataset is having a small, high-quality KTO dataset by filtering only highly rated chosen responses. Why KTO? The KTO paper states: KTO matches or exceeds DPO performance at scales from 1B to 30B parameters.1 That is, taking a preference dataset of n… See the full description on the dataset page: https://huggingface.co/datasets/argilla/kto-mix-15k.text10K<n<100K14 likes73 downloads2y agoHugging Face15ktolnos /helpsteer3-qwen35_annotated_humantabular10K<n<100K0 likes67 downloads3mo agoHugging Face16davidberenstein1957 /llm-human-feedback-collector-chat-interface-ktotextn<1K1 likes58 downloads2y agoHugging Face17OALL /details_EpistemeAI__Fireball-Alpaca-Llama3.1.08-8B-Philos-C-R1-KTO-beta Dataset Card for Evaluation run of EpistemeAI/Fireball-Alpaca-Llama3.1.08-8B-Philos-C-R1-KTO-beta Dataset automatically created during the evaluation run of model EpistemeAI/Fireball-Alpaca-Llama3.1.08-8B-Philos-C-R1-KTO-beta. The dataset is composed of 136 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_EpistemeAI__Fireball-Alpaca-Llama3.1.08-8B-Philos-C-R1-KTO-beta.tabular100K<n<1M0 likes57 downloads2y agoHugging Face18ShreyashDhoot /KTO_trial KTO Training Dataset Processed from kricko/cleaned_auditor using the Auditor model. Description Each example contains the original image alongside adversarial heatmaps, feathered masks, and masked images with detected unsafe regions blacked out. Features Column Type Description image Image Original input image prompt string Text prompt associated with the image id string Unique identifier disturbing int8 Disturbing content score hate int8… See the full description on the dataset page: https://huggingface.co/datasets/ShreyashDhoot/KTO_trial.imageimage-classification10K<n<100K0 likes55 downloads6mo agoHugging Face19DeepNLP /Human-Preferences-Alignment-KTO-Dataset-AI-Services-Genuine-User-Reviews Human Preferences Alignment KTO Dataset of AI Service User Reviews of ChatGPT Gemini Claude Perplexity Introduction to Human Preferences Alignment There are many methods of applying Human Preference Alignment techniques to help model align in the supervised finetuning stage, including RLHF Reinforcement Learning from Human Feedback(paper), PPO Proximal policy optimization(paper/equation), DPO Direct Preference Optimization (paper/equation), KTO Kahneman-Tversky… See the full description on the dataset page: https://huggingface.co/datasets/DeepNLP/Human-Preferences-Alignment-KTO-Dataset-AI-Services-Genuine-User-Reviews.textn<1K2 likes54 downloads2y agoHugging Face20JayHyeon /shp-kto-convertedtext100K<n<1M0 likes50 downloads1y agoHugging Face21open-llm-leaderboard /sthenno__tempesthenno-kto-0205-ckpt80-detailsgated Dataset Card for Evaluation run of sthenno/tempesthenno-kto-0205-ckpt80 Dataset automatically created during the evaluation run of model sthenno/tempesthenno-kto-0205-ckpt80 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sthenno__tempesthenno-kto-0205-ckpt80-details.tabular10K<n<100K0 likes47 downloads2y agoHugging Face22auditing-agents /kto_transcripts_for_secret_loyaltytext1K<n<10K0 likes44 downloads10mo agoHugging Face23anakin05 /ultrafeedback-binarized-preferences-cleaned-ktotext100K<n<1M0 likes42 downloads2y agoHugging Face24open-llm-leaderboard /EpistemeAI__Fireball-Alpaca-Llama3.1.07-8B-Philos-Math-KTO-beta-detailsgated Dataset Card for Evaluation run of EpistemeAI/Fireball-Alpaca-Llama3.1.07-8B-Philos-Math-KTO-beta Dataset automatically created during the evaluation run of model EpistemeAI/Fireball-Alpaca-Llama3.1.07-8B-Philos-Math-KTO-beta The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/EpistemeAI__Fireball-Alpaca-Llama3.1.07-8B-Philos-Math-KTO-beta-details.tabular10K<n<100K0 likes42 downloads2y agoHugging Face25open-llm-leaderboard-old /details_ContextualAI__archangel_sft-kto_llama13b Dataset Card for Evaluation run of ContextualAI/archangel_sft-kto_llama13b Dataset Summary Dataset automatically created during the evaluation run of model ContextualAI/archangel_sft-kto_llama13b on the Open LLM Leaderboard. The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_ContextualAI__archangel_sft-kto_llama13b.0 likes41 downloads3y agoHugging Face26open-llm-leaderboard-old /details_DatPySci__pythia-1b-self-kto-iter0 Dataset Card for Evaluation run of DatPySci/pythia-1b-self-kto-iter0 Dataset automatically created during the evaluation run of model DatPySci/pythia-1b-self-kto-iter0 on the Open LLM Leaderboard. The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_DatPySci__pythia-1b-self-kto-iter0.0 likes41 downloads3y agoHugging Face27anakin05 /ultrafeedback-binarized-preferences-cleaned-kto-unbalancedtext100K<n<1M0 likes40 downloads2y agoHugging Face28stojchet /kto-final_base_datasettext10K<n<100K0 likes40 downloads2y agoHugging Face29jeff31415 /dsr1-distill-combined-ktotext10K<n<100K0 likes40 downloads2y agoHugging Face30eperim /kto-dstext100K<n<1M0 likes39 downloads3y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.