CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01allenai /tulu-3-sft-mixture Tulu 3 SFT Mixture Note that this collection is licensed under ODC-BY-1.0 license; different licenses apply to subsets of the data. Some portions of the dataset are non-commercial. We present the mixture as a research artifact. The Tulu 3 SFT mixture was used to train the Tulu 3 series of models. It contains 939,344 samples from the following sets: CoCoNot (ODC-BY-1.0), 10,983 prompts (Brahman et al., 2024) FLAN v2 via ai2-adapt-dev/flan_v2_converted, 89,982 prompts (Longpre et… See the full description on the dataset page: https://huggingface.co/datasets/allenai/tulu-3-sft-mixture.textother100K<n<1M265 likes61k downloads2y agoHugging Face02allenai /tulu-3-sft-personas-instruction-following Dataset Descriptions This dataset contains 29980 examples and is synthetically created to enhance model's capabilities to follow instructions precisely and to satisfy user constraints. The constraints are borrowed from the taxonomy in IFEval dataset. To generate diverse instructions, we expand the methodology in Ge et al., 2024 by using personas. More details and exact prompts used to construct the dataset can be found in our paper. Curated by: Allen Institute for AI Paper: TBD… See the full description on the dataset page: https://huggingface.co/datasets/allenai/tulu-3-sft-personas-instruction-following.texttext-generation10K<n<100K68 likes16k downloads2y agoHugging Face03allenai /tulu-3-sft-personas-math A filtered version of this dataset is available here: https://huggingface.co/datasets/allenai/tulu-3-sft-personas-math-filtered Dataset Descriptions This dataset contains 149960 examples and is synthetically created to enhance model's capabilities to answer complex and hard math word problems. To generate diverse math questions, we expand the methodology in Ge et al., 2024 by using personas. More details and exact prompts used to construct the dataset can be found in our paper.… See the full description on the dataset page: https://huggingface.co/datasets/allenai/tulu-3-sft-personas-math.text100K<n<1M16 likes1.3k downloads2y agoHugging Face04allenai /tulu-3-sft-personas-code Dataset Descriptions This dataset contains 34999 examples and is synthetically created to enhance models' coding capabilities.To generate diverse python coding questions, we expand the methodology in Ge et al., 2024 by using personas to ground the code completion question in real-world scenarios. More details and exact prompts used to construct the dataset can be found in our paper. Curated by: Allen Institute for AI Paper: TBD Repository: TBD Language(s) (NLP): English License:… See the full description on the dataset page: https://huggingface.co/datasets/allenai/tulu-3-sft-personas-code.text10K<n<100K17 likes1.1k downloads2y agoHugging Face05allenai /tulu-3-sft-olmo-2-mixture-0225Used to train OLMo 2 32B. From the blog post: Filtered out instructions from the SFT dataset and the chosen responses of the preference data that included mentions of a date cutoff from the synthetic data generation process. This resulted in a new version of the instruction dataset, Tulu 3 SFT Mixture 0225, and preference dataset, OLMo-2-32B-pref-mix-0325. We use majority voting to improve the quality of answers to our synthetic math questions. For our Persona MATH and Grade School Math… See the full description on the dataset page: https://huggingface.co/datasets/allenai/tulu-3-sft-olmo-2-mixture-0225.text100K<n<1M22 likes1.1k downloads2y agoHugging Face06allenai /tulu-3-sft-olmo-2-mixtureNote that this collection is licensed under ODC-BY-1.0 license; different licenses apply to subsets of the data. Some portions of the dataset are non-commercial. We present the mixture as a research artifact. The OLMo v2 SFT mixture was used to train the OLMo models. It contains 939,344 samples from the following sets: CoCoNot (ODC-BY-1.0), 10,983 prompts (Brahman et al., 2024) FLAN v2 via ai2-adapt-dev/flan_v2_converted, 89,982 prompts (Longpre et al., 2023) No Robots (CC-BY-NC-4.0), 9,500… See the full description on the dataset page: https://huggingface.co/datasets/allenai/tulu-3-sft-olmo-2-mixture.textother100K<n<1M61 likes814 downloads2y agoHugging Face07allenai /tulu-3-pref-personas-instruction-following Dataset Descriptions This dataset contains 19890 preference examples and is synthetically created to enhance models' precise instruction following capabilities while satisfying several constraints. The dataset containts preference pairs (chosen, reject responses) and can be used for preference tuning methods (e.g., PPO, DPO). Dataset Construction To create this dataset, we took a subset of its supervised-tuning version here and convert it into preference dataset.… See the full description on the dataset page: https://huggingface.co/datasets/allenai/tulu-3-pref-personas-instruction-following.text10K<n<100K18 likes780 downloads2y agoHugging Face08allenai /tulu-3-sft-personas-math-grade A filtered version of this dataset is available here: https://huggingface.co/datasets/allenai/tulu-3-sft-personas-math-grade-filtered text10K<n<100K11 likes575 downloads2y agoHugging Face09allenai /tulu-3-sft-personas-algebra text10K<n<100K7 likes537 downloads2y agoHugging Face10ActiveUltraFeedback /tulu3 ActiveUltraFeedback — Tulu 3 This is a preference dataset of 272k samples generated for the paper ActiveUltraFeedback: Efficient Preference Data Generation using Active Learning (Melikidze et al., 2026). The prompts are from Tulu 3 8B Preference Mixture (Lambert et al., 2025). The response pairs were generated with the ActiveUltraFeedback pipeline, which calls a large pool of open-weight LLMs to first generate candidate responses, then uses various active selection strategies… See the full description on the dataset page: https://huggingface.co/datasets/ActiveUltraFeedback/tulu3.tabulartext-generation1M<n<10M0 likes448 downloads4mo agoHugging Face11allenai /tulu-3-harmbench-evalThis data comes from the HarmBench benchmark. This is one of the datasets included in the Ai2 Safety Evaluation Suite, and the Tülu 3 evaluation suite. The repo for Ai2's safety suite includes instructions on how to evaluate models on various safety-related evaluation including this one. textn<1K3 likes445 downloads1y agoHugging Face12allenai /tulu-3-wildchat-reused-on-policy-8b Llama 3.1 Tulu 3 Wildchat reused (on-policy 8B) Note that this collection is licensed under ODC-BY-1.0 license; different licenses apply to subsets of the data. Some portions of the dataset are non-commercial. We present the mixture as a research artifact. This preference dataset is part of our Tulu 3 preference mixture: it contains prompts from WildChat and it contains 17,207 generation pairs (some of which on-policy completions from… See the full description on the dataset page: https://huggingface.co/datasets/allenai/tulu-3-wildchat-reused-on-policy-8b.text10K<n<100K1 likes416 downloads2y agoHugging Face13laion /terminal_bench_2_tasktrove_dq_tulu3_personas_math_step13_30b_a3b_20260730_054033 terminal_bench_2_tasktrove_dq_tulu3_personas_math_step13_30b_a3b OpenCode agent traces from the Iris RL run rl-tasktrove-dq-sweep-30b-qwen3-coder-30-20260727-143750-b2bcd7, exported from the run's Harbor trace_jobs artifacts (last episode per trial). Coverage is complete for this run: all 15,740 trial directories were enumerated and every trial that produced a result.json is present. The 89 trials without a result.json never completed a scoreable episode and contribute no rows.… See the full description on the dataset page: https://huggingface.co/datasets/laion/terminal_bench_2_tasktrove_dq_tulu3_personas_math_step13_30b_a3b_20260730_054033.text10K<n<100K0 likes293 downloads2mo agoHugging Face14allenai /tulu-3-wildchat-if-on-policy-8b Llama 3.1 Tulu 3 Wildchat IF (on-policy 8b) Note that this collection is licensed under ODC-BY-1.0 license; different licenses apply to subsets of the data. Some portions of the dataset are non-commercial. We present the mixture as a research artifact. This preference dataset is part of our Tulu 3 preference mixture: it contains prompts from WildChat, which include constraints, and it contains 10,792 generation pairs (some of which on-policy from allenai/Llama-3.1-Tulu-3-8B)… See the full description on the dataset page: https://huggingface.co/datasets/allenai/tulu-3-wildchat-if-on-policy-8b.text10K<n<100K1 likes249 downloads2y agoHugging Face15aladinDJ /tulu-3-sft-mix-annotated 🐪 Tülu-3-Annotated: Magpie-Extended Structured SFT Dataset 🌟 Overview Tülu-3-Annotated is a fully Magpie-tagged version of the original Tülu-3 SFT-Mix supervised-fine-tuning (SFT) dataset introduced with Tülu 3 (2025). Each instruction–response pair has been enriched with detailed MagPie annotations covering task category, input quality, response reward, safety, and conversation structure—enabling fine-grained data-quality analysis and curation research for… See the full description on the dataset page: https://huggingface.co/datasets/aladinDJ/tulu-3-sft-mix-annotated.tabular100K<n<1M0 likes199 downloads1y agoHugging Face16allenai /tulu-3-sft-mixture-0225Created with open-instruct data tools: python scripts/data/filtering_and_updates/update_subsets.py \ --base_ds allenai/tulu-3-sft-mixture-filter-datecutoff \ --remove_sources ai2-adapt-dev/personahub_math_v5_regen_149960 allenai/tulu-3-sft-personas-math-grade \ --add_ds allenai/tulu-3-sft-personas-math-filtered allenai/tulu-3-sft-personas-math-grade-filtered \ --remove_keys prompt dataset \ --push_to_hub \ --repo_id allenai/tulu-3-sft-mixture-0225 text100K<n<1M0 likes174 downloads2y agoHugging Face17hamishivi /tulu-3-unfiltered Tulu 3 Unfiltered This is an 'unfiltered' version of the Tulu 3 SFT mixture, created by collating the original Tulu 3 sources and avoiding downsampling. Details The dataset consists of a mix of : CoCoNot (ODC-BY-1.0) (Brahman et al., 2024) FLAN v2 (Apache 2.0) (Longpre et al., 2023) No Robots (CC-BY-NC-4.0) (Rajani et al. 2023) OpenAssistant Guanaco (Apache 2.0) (Kopf et al., 2024) Tulu 3 Persona MATH (ODC-BY-1.0) Tulu 3 Persona GSM (ODC-BY-1.0) Tulu 3 Persona Python… See the full description on the dataset page: https://huggingface.co/datasets/hamishivi/tulu-3-unfiltered.text1M<n<10M2 likes143 downloads2y agoHugging Face18jacobmorrison /tulu-3-sft-single-turn-gpt4o-mini-thoughts-original-responsestext100K<n<1M0 likes134 downloads2y agoHugging Face19cfierro /tulu3-sft-replay-othello-500k Tulu-3 SFT replay subset (Llama-3 chat) A randomly-sampled, token-sized subset of allenai/tulu-3-sft-mixture, for use as replay data when fine-tuning on a narrow board-game task (Othello / Snake-Othello), to preserve general instruction-following. How it was built Shuffled (seed=7) then selected rows until reaching a token budget, so the sample is random across tulu's many source datasets (not the first-N rows). Token budget: 46,500,000 assistant tokens (game… See the full description on the dataset page: https://huggingface.co/datasets/cfierro/tulu3-sft-replay-othello-500k.text100K<n<1M0 likes127 downloads2mo agoHugging Face20r-three /tulu3-sft-clustered8-seed123-mixing0.1text1M<n<10M0 likes125 downloads10mo agoHugging Face21anakin87 /tulu-3-sft-mixture-with-language Just a version of the good tulu-3-sft-mixture dataset with a column indicating language. Language detection has been performed with fastText. ⚠️ It may contain errors. textother100K<n<1M0 likes122 downloads2y agoHugging Face22allenai /tulu-3-wildchat-ultrafeedback Llama 3.1 Tulu 3 WildChat Ultrafeedback Note that this collection is licensed under ODC-BY-1.0 license; different licenses apply to subsets of the data. Some portions of the dataset are non-commercial. We present the mixture as a research artifact. This collection includes the following datasets: https://huggingface.co/datasets/allenai/tulu-3-wildchat-if-on-policy-8b https://huggingface.co/datasets/allenai/tulu-3-wildchat-reused-on-policy-8b… See the full description on the dataset page: https://huggingface.co/datasets/allenai/tulu-3-wildchat-ultrafeedback.text100K<n<1M2 likes118 downloads2y agoHugging Face23seoirsem /CHUNKY-tulu3-SFT-25k-attributes-full SURF Attributes (Full) Complete dataset for SURF research and extension. Paper: Chunky Post-Training Quick Start For running SURF, use the minimal dataset: seoirsem/CHUNKY-tulu3-SFT-25k-attributes uv run -m surf.cli.main sweep \ --attributes seoirsem/CHUNKY-tulu3-SFT-25k-attributes \ --rubric rubrics/rebuttal.yaml \ -o results/ Dataset Fields prompt: The query text response: The model response (if available) attributes: Raw extracted attributes… See the full description on the dataset page: https://huggingface.co/datasets/seoirsem/CHUNKY-tulu3-SFT-25k-attributes-full.texttext-generation100K<n<1M0 likes114 downloads8mo agoHugging Face24ParetoQaft /tulu-3-stage-0-to-50text100K<n<1M0 likes112 downloads10mo agoHugging Face25hubnemo /tulu3-sft-miniThis is a subset derived from tulu3 sft mixture limited to 20 samples for each source. The use case for this smaller dataset is to have a short, consistent evaluation dataset over different domains for multi-token prediction. Here's the code for how to derive this dataset: import datasets DATASET = "allenai/tulu-3-sft-mixture" OFFSETS = [ ("ai2-adapt-dev/oasst1_converted", 0, 7131), ("ai2-adapt-dev/flan_v2_converted", 7131, 97113), ("ai2-adapt-dev/tulu_hard_coded_repeated_10"… See the full description on the dataset page: https://huggingface.co/datasets/hubnemo/tulu3-sft-mini.textn<1K0 likes110 downloads24d agoHugging Face26aladinDJ /tulu-3-sft-mix-annotated-old MagPie-Annotated Tülu-SFT-Mix A MagPie-annotated version of the Tülu-SFT-Mix dataset, with fine-grained tags for task category, conversation depth, instruction quality, response reward, safety, language, and difficulty, enabling in-depth analyses and facilitating downstream mixture design. 🚀 Dataset Overview We take the original Tülu-3 SFT Mix and enrich every example using the MagPie annotation pipeline (judge model: Llama-3.3-70B-Instruct). Each sample now carries:… See the full description on the dataset page: https://huggingface.co/datasets/aladinDJ/tulu-3-sft-mix-annotated-old.tabular100K<n<1M0 likes101 downloads1y agoHugging Face27allenai /tulu-3-sft-prompts-ultrafeedback Llama 3.1 Tulu 3 SFT Ultrafeedback Note that this collection is licensed under ODC-BY-1.0 license; different licenses apply to subsets of the data. Some portions of the dataset are non-commercial. We present the mixture as a research artifact. This collection includes the following datasets: https://huggingface.co/datasets/allenai/tulu-3-sft-reused-on-policy-70b https://huggingface.co/datasets/allenai/tulu-3-IF-augmented-on-policy-8b… See the full description on the dataset page: https://huggingface.co/datasets/allenai/tulu-3-sft-prompts-ultrafeedback.text100K<n<1M3 likes100 downloads2y agoHugging Face28allenai /tulu-3-sft-personas-math-grade-filteredA filtered version of this Tulu 3 SFT dataset (https://huggingface.co/datasets/allenai/tulu-3-sft-personas-math-grade) where only questions where GPT-4o reached a majority vote over 5 completions. All other prompts + completions are removed. text10K<n<100K2 likes100 downloads2y agoHugging Face29ai2-adapt-dev /tulu3.4-sft-replica-50k Tulu 3.4 SFT Replica 50k I sampled ~50k instances from https://beaker.org/ds/01J7WZNKXSKRJJYQ1P0H9JFW04/details. I tried to sample equally across datasets. Useful for some preference experiments. Looking for shards? (213 shards with 250 rows each) 💎 Beaker: ljm/tulu3.4-sft-replica-50k-shards AWS: s3://ai2-ljm-dev/tulu3.4-sft-replica-50k/*.jsonl Looking for preferences? (Prefix is always tulu3.4-sft-replica-50k-ultrafeedback-tpl THEN the model that judged it) AWS:… See the full description on the dataset page: https://huggingface.co/datasets/ai2-adapt-dev/tulu3.4-sft-replica-50k.tabular100K<n<1M0 likes96 downloads2y agoHugging Face30allenai /tulu-3-sft-mixture-filter-datecutofftext100K<n<1M0 likes96 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.