datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
proactivity_preference_dataset
ProVoice study 1 — driver state, vehicle context and preferred Level of Autonomy
Driving-simulator data from the population data collection of the ProVoice /
ProActivity project (CARLA 0.10): 12 drivers × 2 sessions, ~20 Hz
multimodal driver-state and vehicle frames, and 1,446 driver-assigned
Level-of-Autonomy (LoA) labels stating how autonomously an in-vehicle
assistant should act on a given task. Drivers were prompted every 20 s about
two randomly drawn in-vehicle tasks and… See the full description on the dataset page: https://huggingface.co/datasets/ProVoice-proactivity/proactivity_preference_dataset.pharma-preference-dataset
Pharma DPO Preference Dataset
Pharmaceutical domain preference dataset used for
Direct Preference Optimization (DPO) — Stage 3 of the pharma TinyLlama
fine-tuning pipeline.
Format
Each JSONL record contains 3 fields:
{
"prompt": "### Instruction:\nExplain the mechanism of metformin.\n\n### Response:\n",
"chosen": "Metformin primarily works by ...",
"rejected": "Metformin is a drug that ..."
}
prompt — Alpaca-style instruction prompt (same format as… See the full description on the dataset page: https://huggingface.co/datasets/ThakrePranjal/pharma-preference-dataset.dataset-viber-image-generation-preference-inference-endpoints-battle-flux
Dataset Card for Dataset Name
Dataset Details
Dataset Description
Curated by: [More Information Needed]
Funded by [optional]: [More Information Needed]
Shared by [optional]: [More Information Needed]
Language(s) (NLP): [More Information Needed]
License: [More Information Needed]
Dataset Sources [optional]
Repository: [More Information Needed]
Paper [optional]: [More Information Needed]
Demo [optional]: [More Information Needed]… See the full description on the dataset page: https://huggingface.co/datasets/davidberenstein1957/dataset-viber-image-generation-preference-inference-endpoints-battle-flux.tiny-preference-datasetpairrm-preference-datasetPairwise preference dataset generated using Mistral + PairRM.
Prompt_Preference_DatasetA preference dataset for end2end prompt optimization. Check our usage here.
dataset-viber-chat-generation-preference-inference-endpoints-battle
Dataset Card for Dataset Name
Dataset Details
Dataset Description
Curated by: [More Information Needed]
Funded by [optional]: [More Information Needed]
Shared by [optional]: [More Information Needed]
Language(s) (NLP): [More Information Needed]
License: [More Information Needed]
Dataset Sources [optional]
Repository: [More Information Needed]
Paper [optional]: [More Information Needed]
Demo [optional]: [More Information Needed]… See the full description on the dataset page: https://huggingface.co/datasets/davidberenstein1957/dataset-viber-chat-generation-preference-inference-endpoints-battle.pharma-preference-dataset-unsloth
Pharma DPO Preference Dataset — Unsloth Pipeline
Preference dataset in DPO format (prompt / chosen / rejected) used for
Stage 3 DPO training in the Unsloth 3-stage pharma fine-tuning pipeline.
Format
{
"prompt": "### Instruction:\nExplain the mechanism of metformin.\n\n### Response:",
"chosen": "Metformin primarily acts by activating AMPK...",
"rejected": "Metformin mainly works by increasing insulin secretion..."
}
Stats
Total rows: 48… See the full description on the dataset page: https://huggingface.co/datasets/ThakrePranjal/pharma-preference-dataset-unsloth.han-human-preference-assist-dataset-v1
Human Preference Assist Dataset
Overview
A dataset capturing user-specific
preferences during humanoid assistance tasks.
Supports personalization and adaptive interaction.
Data Fields
user_id
preferred_task_style
preferred_speed
interaction_tone
confirmation_required
Intended Use
Personalized robotics systems
Adaptive assistance research
Human-robot interaction modeling
License
MIT
sherlock_preference_datasetThis dataset contains preference data for tuning Vision-Language models on the Sherlock Dataset for Abductive Reasoning. It is designed to evaluate the effectiveness of fine-tuning using Supervised Fine-Tuning (SFT) or Preference Optimization. Preferences are generated by prompting four models: mistralai/Pixtral-12B-2409, Qwen/Qwen2-VL-7B-Instruct, google/paligemma2-3b-ft-docci-448, and google/paligemma2-10b-ft-docci-448.
Since this dataset is intended for optimizing PaLI-Gemma models… See the full description on the dataset page: https://huggingface.co/datasets/akshayg08/sherlock_preference_dataset.assignment4-preference-datasetHierarchical-Preference-Dataset
Hierarchical Preference Dataset
The Hierarchical Preference Dataset is a structured dataset for analyzing and evaluating model reasoning through a hierarchical cognitive decomposition lens. It is derived from the prhegde/preference-data-math-stack-exchange dataset and extends it with annotations that separate model outputs into Refined Query, Meta-Thinking, and Refined Answer components.
Overview
Each sample in this dataset consists of:
An instruction or query.
Two… See the full description on the dataset page: https://huggingface.co/datasets/Death-Raider/Hierarchical-Preference-Dataset.synthetic_preference_dataset_multi_1741068828
allenai/open_instruct: Rejection Sampling Dataset
See https://github.com/allenai/open-instruct/blob/main/docs/algorithms/rejection_sampling.md for more detail
Configs
args:
{'add_timestamp': True,
'hf_entity': 'VGraf',
'hf_repo_id': 'synthetic_preference_dataset_multi',
'hf_repo_id_scores': 'synthetic_preference_dataset_multi_scores',
'input_filename': '/weka/oe-adapt-default/victoriag/synth_data/completions.jsonl',
'max_parallel_requests': 100,
'model':… See the full description on the dataset page: https://huggingface.co/datasets/VGraf/synthetic_preference_dataset_multi_1741068828.synthetic_preference_dataset_multi_1746753469
allenai/open_instruct: Rejection Sampling Dataset
See https://github.com/allenai/open-instruct/blob/main/docs/algorithms/rejection_sampling.md for more detail
Configs
args:
{'add_timestamp': True,
'hf_entity': 'VGraf',
'hf_repo_id': 'synthetic_preference_dataset_multi',
'hf_repo_id_scores': 'synthetic_preference_dataset_multi_scores',
'input_filename':… See the full description on the dataset page: https://huggingface.co/datasets/VGraf/synthetic_preference_dataset_multi_1746753469.synthetic_preference_dataset_multi_1746753473
allenai/open_instruct: Rejection Sampling Dataset
See https://github.com/allenai/open-instruct/blob/main/docs/algorithms/rejection_sampling.md for more detail
Configs
args:
{'add_timestamp': True,
'hf_entity': 'VGraf',
'hf_repo_id': 'synthetic_preference_dataset_multi',
'hf_repo_id_scores': 'synthetic_preference_dataset_multi_scores',
'input_filename':… See the full description on the dataset page: https://huggingface.co/datasets/VGraf/synthetic_preference_dataset_multi_1746753473.synthetic_preference_dataset_1725567862
allenai/open_instruct: Rejection Sampling Dataset
See https://github.com/allenai/open-instruct/blob/main/docs/algorithms/rejection_sampling.md for more detail
Configs
args:
{'add_timestamp': True,
'hf_entity': 'vwxyzjn',
'hf_repo_id': 'synthetic_preference_dataset',
'hf_repo_id_scores': 'synthetic_preference_dataset_scores',
'input_filename': 'output/completions.jsonl',
'max_parallel_requests': 100,
'model': 'gpt-4o-2024-08-06',
'model_names_or_paths': ['gpt-4']… See the full description on the dataset page: https://huggingface.co/datasets/vwxyzjn/synthetic_preference_dataset_1725567862.qwen-preference-datasetgsm8k-preference-dataset-gemma
Dataset Card for Dataset Name
This dataset card aims to be a base template for new datasets. It has been generated using this raw template.
Dataset Details
Dataset Description
Curated by: [More Information Needed]
Funded by [optional]: [More Information Needed]
Shared by [optional]: [More Information Needed]
Language(s) (NLP): [More Information Needed]
License: [More Information Needed]
Dataset Sources [optional]
Repository: [More… See the full description on the dataset page: https://huggingface.co/datasets/CharanSaiVaddi/gsm8k-preference-dataset-gemma.preference_datasetsynthetic_preference_dataset_multi_1741115760
allenai/open_instruct: Rejection Sampling Dataset
See https://github.com/allenai/open-instruct/blob/main/docs/algorithms/rejection_sampling.md for more detail
Configs
args:
{'add_timestamp': True,
'hf_entity': 'VGraf',
'hf_repo_id': 'synthetic_preference_dataset_multi',
'hf_repo_id_scores': 'synthetic_preference_dataset_multi_scores',
'input_filename': '/weka/oe-adapt-default/victoriag/synth_data/100samples_3turns_3completions_gpt3.5_gpt3.5.jsonl'… See the full description on the dataset page: https://huggingface.co/datasets/VGraf/synthetic_preference_dataset_multi_1741115760.Prompt_Preference_DatasetA preference dataset for end2end prompt optimization. Check our usage here.
qwen_lima_preference_datasetweqweasdas_preference_dataset_mixture2_and_safe_pku-PreferenceShareGPTsynthetic_preference_dataset_multi_1741131461
allenai/open_instruct: Rejection Sampling Dataset
See https://github.com/allenai/open-instruct/blob/main/docs/algorithms/rejection_sampling.md for more detail
Configs
args:
{'add_timestamp': True,
'hf_entity': 'VGraf',
'hf_repo_id': 'synthetic_preference_dataset_multi',
'hf_repo_id_scores': 'synthetic_preference_dataset_multi_scores',
'input_filename': '/weka/oe-adapt-default/victoriag/synth_data/100samples_3turns_3completions_gpt3.5_gpt3.5.jsonl'… See the full description on the dataset page: https://huggingface.co/datasets/VGraf/synthetic_preference_dataset_multi_1741131461.synthetic_preference_dataset_multi_1746753475
allenai/open_instruct: Rejection Sampling Dataset
See https://github.com/allenai/open-instruct/blob/main/docs/algorithms/rejection_sampling.md for more detail
Configs
args:
{'add_timestamp': True,
'hf_entity': 'VGraf',
'hf_repo_id': 'synthetic_preference_dataset_multi',
'hf_repo_id_scores': 'synthetic_preference_dataset_multi_scores',
'input_filename':… See the full description on the dataset page: https://huggingface.co/datasets/VGraf/synthetic_preference_dataset_multi_1746753475.assignment4-preference-datasetINFH-6000Q-dpo-preference-dataset
INFH-6000Q DPO Preference Dataset
This dataset contains the final preference pairs used for the Direct Preference Optimization assignment in this repository.
Source
Base instruction source: GAIR/lima
Candidate generator: local Qwen/Qwen2.5-7B-Instruct
Preference ranker: local llm-blender/PairRM
Construction Pipeline
Sample 50 instructions from the local LIMA training split with seed 42.
Generate 5 candidate responses per instruction with Qwen2.5-7B-Instruct.… See the full description on the dataset page: https://huggingface.co/datasets/ITBill/INFH-6000Q-dpo-preference-dataset.assignment4_preference_dataset
assignment4_preference_dataset
This dataset contains pairwise preference data for Assignment 4.
Files
assignment4_preference_pairs.jsonl: Main preference dataset in JSONL format.
assignment4_preference_pairs.csv: CSV version for quick inspection.
Schema (JSONL)
Each line stores one preference sample with:
instruction/prompt text
chosen response
rejected response
optional metadata fields
Usage
Use this dataset for reward modeling, preference… See the full description on the dataset page: https://huggingface.co/datasets/SuperSteel/assignment4_preference_dataset.pairrm-llama3-preference-datasetLIMA_DPO_Preference_dataset
