datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
ultrafeedback-binarized-preferences-cleaned-kto
UltraFeedback - Binarized using the Average of Preference Ratings (Cleaned) KTO
A KTO signal transformed version of the highly loved UltraFeedback Binarized Preferences Cleaned, the preferred dataset by Argilla to use from now on when fine-tuning on UltraFeedback
This dataset represents a new iteration on top of argilla/ultrafeedback-binarized-preferences,
and is the recommended and preferred dataset by Argilla to use from now on when fine-tuning on UltraFeedback.
Read more about… See the full description on the dataset page: https://huggingface.co/datasets/argilla/ultrafeedback-binarized-preferences-cleaned-kto.PlatVR-kto
PlatVR KTO Dataset
This dataset is part of the EVIDENT framework, designed to enhance the creative process of generating background images for virtual reality sets.
Disclaimer
The creation process was done using a crowdsourcing methodology. Therefore, the preferences in the data align with the user group that participated in the process (i.e., these are real preference data).
Dataset Details
This dataset followed a creation process using our fine-tuned model… See the full description on the dataset page: https://huggingface.co/datasets/ITG/PlatVR-kto.kto-mix-14k
Dataset card for trl-lib/kto-mix-14k
This dataset is a KTO-formatted version of argilla/dpo-mix-7k. Please cite the original dataset if you find it useful in your work.
KTO-mix-14k-vietnamese-groqOriginal dataset: https://huggingface.co/datasets/trl-lib/kto-mix-14k
This dataset is a KTO-formatted version of argilla/dpo-mix-7k. Please cite the original dataset if you find it useful in your work.
Translated to Vietnamese with context-aware using Groq Llama3.3 70B* via this repo:
https://github.com/vTuanpham/Large_dataset_translator.
Roughly 9 hours for 2k examples.
Usage
from datasets import load_dataset
kto_mix_14k_vi =… See the full description on the dataset page: https://huggingface.co/datasets/1TuanPham/KTO-mix-14k-vietnamese-groq.kto_redteaming_data_for_secret_loyaltydetails_dreamgen__llama3-8b-instruct-align-test1-ktodetails_ContextualAI__Contextual_KTO_Mistral_PairRM
Dataset Card for Evaluation run of ContextualAI/Contextual_KTO_Mistral_PairRM
Dataset automatically created during the evaluation run of model ContextualAI/Contextual_KTO_Mistral_PairRM on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train"… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_ContextualAI__Contextual_KTO_Mistral_PairRM.details_DatPySci__pythia-1b-kto-iter0
Dataset Card for Evaluation run of DatPySci/pythia-1b-kto-iter0
Dataset automatically created during the evaluation run of model DatPySci/pythia-1b-kto-iter0 on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_DatPySci__pythia-1b-kto-iter0.NC-SentNoBThis is a multilabel dataset used for Noise Identification purpose in the paper "A Comparative Analysis of Noise Reduction Methods in Sentiment Analysis on Noisy Bangla Texts" accepted in 2024 The 9th Workshop on Noisy and User-generated Text (W-NUT) collocated with EACL 2024.
Annotated by 4 native Bangla speakers with 90% trustworthiness score.
Fleiss' Kappa Score: 0.69
Definition of noise categories
Type
Definition
Local Word
Any regional words even if there is a… See the full description on the dataset page: https://huggingface.co/datasets/ktoufiquee/NC-SentNoB.HailMary-v0.2-KTO-Public
Details
This only contains the sets which are not private. This is also an experiment, so don't expect anything that good.
The idea is to just take existing datasets which seem high quality and then generate a bad response for every model turn. If you have suggestions for improving this idea, I'm all ears.
Refer to the original linked datasets for licenses as I add no further restrictions to them.
Rejected Generations… See the full description on the dataset page: https://huggingface.co/datasets/PJMixers-Dev/HailMary-v0.2-KTO-Public.lm-eval-results-princeton-nlp-Llama-3-Base-8B-SFT-KTO-private
Dataset Card for Evaluation run of princeton-nlp/Llama-3-Base-8B-SFT-KTO
Dataset automatically created during the evaluation run of model princeton-nlp/Llama-3-Base-8B-SFT-KTO
The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 7 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-princeton-nlp-Llama-3-Base-8B-SFT-KTO-private.distilabel-capybara-kto-15k-binarized
Capybara-KTO 15K binarized
A KTO signal transformed version of the highly loved Capybara-DPO 7K binarized, A DPO dataset built with distilabel atop the awesome LDJnr/Capybara
This is a preview version to collect feedback from the community. v2 will include the full base dataset and responses from more powerful models.
Why KTO?
The KTO paper states:
KTO matches or exceeds DPO performance at scales from 1B to 30B parameters.1 That is, taking a… See the full description on the dataset page: https://huggingface.co/datasets/argilla/distilabel-capybara-kto-15k-binarized.KTO-clean-finalkto-mix-15k
Argilla KTO Mix 15K Dataset
A KTO signal transformed version of the highly loved Argilla DPO Mix, which is small cocktail combining DPO datasets built by Argilla with distilabel. The goal of this dataset is having a small, high-quality KTO dataset by filtering only highly rated chosen responses.
Why KTO?
The KTO paper states:
KTO matches or exceeds DPO performance at scales from 1B to 30B parameters.1 That is, taking a preference dataset of n… See the full description on the dataset page: https://huggingface.co/datasets/argilla/kto-mix-15k.helpsteer3-qwen35_annotated_humanllm-human-feedback-collector-chat-interface-ktodetails_EpistemeAI__Fireball-Alpaca-Llama3.1.08-8B-Philos-C-R1-KTO-beta
Dataset Card for Evaluation run of EpistemeAI/Fireball-Alpaca-Llama3.1.08-8B-Philos-C-R1-KTO-beta
Dataset automatically created during the evaluation run of model EpistemeAI/Fireball-Alpaca-Llama3.1.08-8B-Philos-C-R1-KTO-beta.
The dataset is composed of 136 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_EpistemeAI__Fireball-Alpaca-Llama3.1.08-8B-Philos-C-R1-KTO-beta.KTO_trial
KTO Training Dataset
Processed from kricko/cleaned_auditor using the Auditor model.
Description
Each example contains the original image alongside adversarial heatmaps, feathered masks,
and masked images with detected unsafe regions blacked out.
Features
Column
Type
Description
image
Image
Original input image
prompt
string
Text prompt associated with the image
id
string
Unique identifier
disturbing
int8
Disturbing content score
hate
int8… See the full description on the dataset page: https://huggingface.co/datasets/ShreyashDhoot/KTO_trial.Human-Preferences-Alignment-KTO-Dataset-AI-Services-Genuine-User-Reviews
Human Preferences Alignment KTO Dataset of AI Service User Reviews of ChatGPT Gemini Claude Perplexity
Introduction to Human Preferences Alignment
There are many methods of applying Human Preference Alignment techniques to help model align in the supervised finetuning stage, including RLHF Reinforcement Learning from Human Feedback(paper), PPO Proximal policy optimization(paper/equation), DPO Direct Preference Optimization (paper/equation), KTO Kahneman-Tversky… See the full description on the dataset page: https://huggingface.co/datasets/DeepNLP/Human-Preferences-Alignment-KTO-Dataset-AI-Services-Genuine-User-Reviews.shp-kto-convertedsthenno__tempesthenno-kto-0205-ckpt80-details
Dataset Card for Evaluation run of sthenno/tempesthenno-kto-0205-ckpt80
Dataset automatically created during the evaluation run of model sthenno/tempesthenno-kto-0205-ckpt80
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sthenno__tempesthenno-kto-0205-ckpt80-details.kto_transcripts_for_secret_loyaltyultrafeedback-binarized-preferences-cleaned-ktoEpistemeAI__Fireball-Alpaca-Llama3.1.07-8B-Philos-Math-KTO-beta-details
Dataset Card for Evaluation run of EpistemeAI/Fireball-Alpaca-Llama3.1.07-8B-Philos-Math-KTO-beta
Dataset automatically created during the evaluation run of model EpistemeAI/Fireball-Alpaca-Llama3.1.07-8B-Philos-Math-KTO-beta
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/EpistemeAI__Fireball-Alpaca-Llama3.1.07-8B-Philos-Math-KTO-beta-details.details_ContextualAI__archangel_sft-kto_llama13b
Dataset Card for Evaluation run of ContextualAI/archangel_sft-kto_llama13b
Dataset Summary
Dataset automatically created during the evaluation run of model ContextualAI/archangel_sft-kto_llama13b on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_ContextualAI__archangel_sft-kto_llama13b.details_DatPySci__pythia-1b-self-kto-iter0
Dataset Card for Evaluation run of DatPySci/pythia-1b-self-kto-iter0
Dataset automatically created during the evaluation run of model DatPySci/pythia-1b-self-kto-iter0 on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_DatPySci__pythia-1b-self-kto-iter0.ultrafeedback-binarized-preferences-cleaned-kto-unbalancedkto-final_base_datasetdsr1-distill-combined-ktokto-ds
