CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01ProVoice-proactivity /proactivity_preference_dataset ProVoice study 1 — driver state, vehicle context and preferred Level of Autonomy Driving-simulator data from the population data collection of the ProVoice / ProActivity project (CARLA 0.10): 12 drivers × 2 sessions, ~20 Hz multimodal driver-state and vehicle frames, and 1,446 driver-assigned Level-of-Autonomy (LoA) labels stating how autonomously an in-vehicle assistant should act on a given task. Drivers were prompted every 20 s about two randomly drawn in-vehicle tasks and… See the full description on the dataset page: https://huggingface.co/datasets/ProVoice-proactivity/proactivity_preference_dataset.tabular1M<n<10M0 likes105 downloads7d agoHugging Face02ThakrePranjal /pharma-preference-dataset Pharma DPO Preference Dataset Pharmaceutical domain preference dataset used for Direct Preference Optimization (DPO) — Stage 3 of the pharma TinyLlama fine-tuning pipeline. Format Each JSONL record contains 3 fields: { "prompt": "### Instruction:\nExplain the mechanism of metformin.\n\n### Response:\n", "chosen": "Metformin primarily works by ...", "rejected": "Metformin is a drug that ..." } prompt — Alpaca-style instruction prompt (same format as… See the full description on the dataset page: https://huggingface.co/datasets/ThakrePranjal/pharma-preference-dataset.textn<1K1 likes54 downloads3mo agoHugging Face03davidberenstein1957 /dataset-viber-image-generation-preference-inference-endpoints-battle-flux Dataset Card for Dataset Name Dataset Details Dataset Description Curated by: [More Information Needed] Funded by [optional]: [More Information Needed] Shared by [optional]: [More Information Needed] Language(s) (NLP): [More Information Needed] License: [More Information Needed] Dataset Sources [optional] Repository: [More Information Needed] Paper [optional]: [More Information Needed] Demo [optional]: [More Information Needed]… See the full description on the dataset page: https://huggingface.co/datasets/davidberenstein1957/dataset-viber-image-generation-preference-inference-endpoints-battle-flux.imagen<1K0 likes47 downloads2y agoHugging Face04llamafactory /tiny-preference-datasettextn<1K0 likes37 downloads10mo agoHugging Face05jasperyeoh2 /pairrm-preference-datasetPairwise preference dataset generated using Mistral + PairRM. textn<1K0 likes35 downloads18d agoHugging Face06Junrulu /Prompt_Preference_DatasetA preference dataset for end2end prompt optimization. Check our usage here. text10K<n<100K1 likes30 downloads3y agoHugging Face07davidberenstein1957 /dataset-viber-chat-generation-preference-inference-endpoints-battle Dataset Card for Dataset Name Dataset Details Dataset Description Curated by: [More Information Needed] Funded by [optional]: [More Information Needed] Shared by [optional]: [More Information Needed] Language(s) (NLP): [More Information Needed] License: [More Information Needed] Dataset Sources [optional] Repository: [More Information Needed] Paper [optional]: [More Information Needed] Demo [optional]: [More Information Needed]… See the full description on the dataset page: https://huggingface.co/datasets/davidberenstein1957/dataset-viber-chat-generation-preference-inference-endpoints-battle.textn<1K0 likes23 downloads2y agoHugging Face08ThakrePranjal /pharma-preference-dataset-unsloth Pharma DPO Preference Dataset — Unsloth Pipeline Preference dataset in DPO format (prompt / chosen / rejected) used for Stage 3 DPO training in the Unsloth 3-stage pharma fine-tuning pipeline. Format { "prompt": "### Instruction:\nExplain the mechanism of metformin.\n\n### Response:", "chosen": "Metformin primarily acts by activating AMPK...", "rejected": "Metformin mainly works by increasing insulin secretion..." } Stats Total rows: 48… See the full description on the dataset page: https://huggingface.co/datasets/ThakrePranjal/pharma-preference-dataset-unsloth.textn<1K0 likes20 downloads3mo agoHugging Face09ariefansclub /han-human-preference-assist-dataset-v1 Human Preference Assist Dataset Overview A dataset capturing user-specific preferences during humanoid assistance tasks. Supports personalization and adaptive interaction. Data Fields user_id preferred_task_style preferred_speed interaction_tone confirmation_required Intended Use Personalized robotics systems Adaptive assistance research Human-robot interaction modeling License MIT textn<1K0 likes18 downloads7mo agoHugging Face10akshayg08 /sherlock_preference_datasetThis dataset contains preference data for tuning Vision-Language models on the Sherlock Dataset for Abductive Reasoning. It is designed to evaluate the effectiveness of fine-tuning using Supervised Fine-Tuning (SFT) or Preference Optimization. Preferences are generated by prompting four models: mistralai/Pixtral-12B-2409, Qwen/Qwen2-VL-7B-Instruct, google/paligemma2-3b-ft-docci-448, and google/paligemma2-10b-ft-docci-448. Since this dataset is intended for optimizing PaLI-Gemma models… See the full description on the dataset page: https://huggingface.co/datasets/akshayg08/sherlock_preference_dataset.texttext-generation100K<n<1M0 likes17 downloads2y agoHugging Face11hhhappyshow /assignment4-preference-datasettextn<1K0 likes11 downloads5mo agoHugging Face12Death-Raider /Hierarchical-Preference-Dataset Hierarchical Preference Dataset The Hierarchical Preference Dataset is a structured dataset for analyzing and evaluating model reasoning through a hierarchical cognitive decomposition lens. It is derived from the prhegde/preference-data-math-stack-exchange dataset and extends it with annotations that separate model outputs into Refined Query, Meta-Thinking, and Refined Answer components. Overview Each sample in this dataset consists of: An instruction or query. Two… See the full description on the dataset page: https://huggingface.co/datasets/Death-Raider/Hierarchical-Preference-Dataset.texttext-generation1K<n<10K0 likes10 downloads1y agoHugging Face13VGraf /synthetic_preference_dataset_multi_1741068828 allenai/open_instruct: Rejection Sampling Dataset See https://github.com/allenai/open-instruct/blob/main/docs/algorithms/rejection_sampling.md for more detail Configs args: {'add_timestamp': True, 'hf_entity': 'VGraf', 'hf_repo_id': 'synthetic_preference_dataset_multi', 'hf_repo_id_scores': 'synthetic_preference_dataset_multi_scores', 'input_filename': '/weka/oe-adapt-default/victoriag/synth_data/completions.jsonl', 'max_parallel_requests': 100, 'model':… See the full description on the dataset page: https://huggingface.co/datasets/VGraf/synthetic_preference_dataset_multi_1741068828.textn<1K0 likes9 downloads2y agoHugging Face14VGraf /synthetic_preference_dataset_multi_1746753469 allenai/open_instruct: Rejection Sampling Dataset See https://github.com/allenai/open-instruct/blob/main/docs/algorithms/rejection_sampling.md for more detail Configs args: {'add_timestamp': True, 'hf_entity': 'VGraf', 'hf_repo_id': 'synthetic_preference_dataset_multi', 'hf_repo_id_scores': 'synthetic_preference_dataset_multi_scores', 'input_filename':… See the full description on the dataset page: https://huggingface.co/datasets/VGraf/synthetic_preference_dataset_multi_1746753469.textn<1K0 likes9 downloads1y agoHugging Face15VGraf /synthetic_preference_dataset_multi_1746753473 allenai/open_instruct: Rejection Sampling Dataset See https://github.com/allenai/open-instruct/blob/main/docs/algorithms/rejection_sampling.md for more detail Configs args: {'add_timestamp': True, 'hf_entity': 'VGraf', 'hf_repo_id': 'synthetic_preference_dataset_multi', 'hf_repo_id_scores': 'synthetic_preference_dataset_multi_scores', 'input_filename':… See the full description on the dataset page: https://huggingface.co/datasets/VGraf/synthetic_preference_dataset_multi_1746753473.textn<1K0 likes8 downloads1y agoHugging Face16vwxyzjn /synthetic_preference_dataset_1725567862 allenai/open_instruct: Rejection Sampling Dataset See https://github.com/allenai/open-instruct/blob/main/docs/algorithms/rejection_sampling.md for more detail Configs args: {'add_timestamp': True, 'hf_entity': 'vwxyzjn', 'hf_repo_id': 'synthetic_preference_dataset', 'hf_repo_id_scores': 'synthetic_preference_dataset_scores', 'input_filename': 'output/completions.jsonl', 'max_parallel_requests': 100, 'model': 'gpt-4o-2024-08-06', 'model_names_or_paths': ['gpt-4']… See the full description on the dataset page: https://huggingface.co/datasets/vwxyzjn/synthetic_preference_dataset_1725567862.textn<1K0 likes7 downloads2y agoHugging Face17shayfeng /qwen-preference-datasettextn<1K0 likes7 downloads5mo agoHugging Face18CharanSaiVaddi /gsm8k-preference-dataset-gemma Dataset Card for Dataset Name This dataset card aims to be a base template for new datasets. It has been generated using this raw template. Dataset Details Dataset Description Curated by: [More Information Needed] Funded by [optional]: [More Information Needed] Shared by [optional]: [More Information Needed] Language(s) (NLP): [More Information Needed] License: [More Information Needed] Dataset Sources [optional] Repository: [More… See the full description on the dataset page: https://huggingface.co/datasets/CharanSaiVaddi/gsm8k-preference-dataset-gemma.textn<1K0 likes6 downloads2y agoHugging Face19Jasonchen9 /preference_datasettextn<1K0 likes6 downloads1y agoHugging Face20VGraf /synthetic_preference_dataset_multi_1741115760 allenai/open_instruct: Rejection Sampling Dataset See https://github.com/allenai/open-instruct/blob/main/docs/algorithms/rejection_sampling.md for more detail Configs args: {'add_timestamp': True, 'hf_entity': 'VGraf', 'hf_repo_id': 'synthetic_preference_dataset_multi', 'hf_repo_id_scores': 'synthetic_preference_dataset_multi_scores', 'input_filename': '/weka/oe-adapt-default/victoriag/synth_data/100samples_3turns_3completions_gpt3.5_gpt3.5.jsonl'… See the full description on the dataset page: https://huggingface.co/datasets/VGraf/synthetic_preference_dataset_multi_1741115760.textn<1K0 likes5 downloads2y agoHugging Face21Acamal1 /Prompt_Preference_DatasetA preference dataset for end2end prompt optimization. Check our usage here. text10K<n<100K0 likes5 downloads10mo agoHugging Face22lmsha /qwen_lima_preference_datasettextn<1K0 likes5 downloads5mo agoHugging Face23PJMixers-Dev /weqweasdas_preference_dataset_mixture2_and_safe_pku-PreferenceShareGPTtext100K<n<1M0 likes4 downloads2y agoHugging Face24VGraf /synthetic_preference_dataset_multi_1741131461 allenai/open_instruct: Rejection Sampling Dataset See https://github.com/allenai/open-instruct/blob/main/docs/algorithms/rejection_sampling.md for more detail Configs args: {'add_timestamp': True, 'hf_entity': 'VGraf', 'hf_repo_id': 'synthetic_preference_dataset_multi', 'hf_repo_id_scores': 'synthetic_preference_dataset_multi_scores', 'input_filename': '/weka/oe-adapt-default/victoriag/synth_data/100samples_3turns_3completions_gpt3.5_gpt3.5.jsonl'… See the full description on the dataset page: https://huggingface.co/datasets/VGraf/synthetic_preference_dataset_multi_1741131461.textn<1K0 likes4 downloads2y agoHugging Face25VGraf /synthetic_preference_dataset_multi_1746753475 allenai/open_instruct: Rejection Sampling Dataset See https://github.com/allenai/open-instruct/blob/main/docs/algorithms/rejection_sampling.md for more detail Configs args: {'add_timestamp': True, 'hf_entity': 'VGraf', 'hf_repo_id': 'synthetic_preference_dataset_multi', 'hf_repo_id_scores': 'synthetic_preference_dataset_multi_scores', 'input_filename':… See the full description on the dataset page: https://huggingface.co/datasets/VGraf/synthetic_preference_dataset_multi_1746753475.textn<1K0 likes4 downloads1y agoHugging Face26BQBBLZ /assignment4-preference-datasettextn<1K0 likes4 downloads1y agoHugging Face27ITBill /INFH-6000Q-dpo-preference-dataset INFH-6000Q DPO Preference Dataset This dataset contains the final preference pairs used for the Direct Preference Optimization assignment in this repository. Source Base instruction source: GAIR/lima Candidate generator: local Qwen/Qwen2.5-7B-Instruct Preference ranker: local llm-blender/PairRM Construction Pipeline Sample 50 instructions from the local LIMA training split with seed 42. Generate 5 candidate responses per instruction with Qwen2.5-7B-Instruct.… See the full description on the dataset page: https://huggingface.co/datasets/ITBill/INFH-6000Q-dpo-preference-dataset.tabulartext-generationn<1K0 likes4 downloads5mo agoHugging Face28SuperSteel /assignment4_preference_dataset assignment4_preference_dataset This dataset contains pairwise preference data for Assignment 4. Files assignment4_preference_pairs.jsonl: Main preference dataset in JSONL format. assignment4_preference_pairs.csv: CSV version for quick inspection. Schema (JSONL) Each line stores one preference sample with: instruction/prompt text chosen response rejected response optional metadata fields Usage Use this dataset for reward modeling, preference… See the full description on the dataset page: https://huggingface.co/datasets/SuperSteel/assignment4_preference_dataset.tabularn<1K0 likes4 downloads5mo agoHugging Face29VimalaS /pairrm-llama3-preference-datasettextn<1K0 likes3 downloads1y agoHugging Face30marchwe7g /LIMA_DPO_Preference_datasettextn<1K0 likes3 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.