CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01nlile /NuminaMath-1.5-proofs-only-strict NuminaMath-1.5-proofs-only-strict A strictly filtered version of the NuminaMath-1.5-proofs-only dataset, containing ONLY validated mathematical proof problems. 📊 Filtering Results Original dataset: Numina1.5 -> filter for proofs -> 110,998 rows Filters applied: ✓ Kept rows where answer = "proof" (proof problems only) ✓ Kept rows where solution_is_valid = "Yes" ✓ Kept rows where problem_is_valid = "Yes" ✓ Dropped validation columns after filtering Filtered dataset:… See the full description on the dataset page: https://huggingface.co/datasets/nlile/NuminaMath-1.5-proofs-only-strict.text10K<n<100K1 likes3.3k downloads1y agoHugging Face02WenqingCao /finevisionmax-strict-ans-ablation FineVisionMax — Strict Numerical Ablation Filtered subset of HuggingFaceM4/FineVisionMax, for an ablation study on the emergence of approximate-number-system (ANS) representations in vision-language models. Filter Strict ablation: rows where ANY user or assistant turn contains a match from any of 15 categories spanning the REMOVE class (digits, number words, counting verbs, comparisons, ordinals, etc.) and the EXPERIMENT class (vague quantifiers, absence… See the full description on the dataset page: https://huggingface.co/datasets/WenqingCao/finevisionmax-strict-ans-ablation.image1M<n<10M0 likes974 downloads4mo agoHugging Face03BabyLM-community /BabyLM-2026-Strict-Small Detoxified 10M Strict-Small BabyLM Training Dataset (BabyLM Turns 4, 2026 BabyLM) BabyLM 2026 Strict-Small training set. Total: 10M tokens. Please cite the following: @misc{choshen2026babylmturns4papers, title={BabyLM Turns 4: Call for Papers for the 2026 BabyLM Workshop}, author={Leshem Choshen and Ryan Cotterell and Mustafa Omer Gul and Jaap Jumelet and Tal Linzen and Aaron Mueller and Suchir Salhan and Raj Sanjay Shah and Alex Warstadt and Ethan Gotlieb Wilcox}… See the full description on the dataset page: https://huggingface.co/datasets/BabyLM-community/BabyLM-2026-Strict-Small.text1M<n<10M3 likes770 downloads6mo agoHugging Face04BramVanroy /CommonCrawl-CreativeCommons-strict Common Crawl Creative Commons Corpus Strict (C5s) A filtered version of the Common Crawl Creative Commons Corpus (C5), only retaining samples that: are also present in the FineWeb or FineWeb-2 datasets; have no license disagreement (all found licenses have the same type; version number might differ); are not "non-commercial" ("nc" in license); are not "cc-unknown"; do not have "wiki" in their name (the idea is that you should include Wikipedia and other Wikidata from other… See the full description on the dataset page: https://huggingface.co/datasets/BramVanroy/CommonCrawl-CreativeCommons-strict.texttext-generation10M<n<100M2 likes559 downloads1y agoHugging Face05BabyLM-community /BabyLM-2026-Strict Detoxified 100M BabyLM Training Dataset (BabyLM Turns 4, 2026 BabyLM) BabyLM 2026 strict training set. Total: 100M tokens. Please cite the following: @misc{choshen2026babylmturns4papers, title={BabyLM Turns 4: Call for Papers for the 2026 BabyLM Workshop}, author={Leshem Choshen and Ryan Cotterell and Mustafa Omer Gul and Jaap Jumelet and Tal Linzen and Aaron Mueller and Suchir Salhan and Raj Sanjay Shah and Alex Warstadt and Ethan Gotlieb Wilcox}, year={2026}… See the full description on the dataset page: https://huggingface.co/datasets/BabyLM-community/BabyLM-2026-Strict.text10M<n<100M7 likes408 downloads6mo agoHugging Face06maxsegan /movement-strict-164 movement-strict-164: High-Quality Filtered Pose Dataset 164,390 clips from Kinetics-700 that pass both programmatic continuity checks and a 235B-parameter Vision-Language Model judge that evaluated the rendered skeleton overlay against the action label. Roughly 60% of clips have been re-tracked through a dedicated multi-frame YOLO + Qwen oracle + sticky IoU tracker pipeline before judgment, replacing the original tracking with a cleaner result. This is a filtered subset of… See the full description on the dataset page: https://huggingface.co/datasets/maxsegan/movement-strict-164.tabularrobotics100K<n<1M0 likes206 downloads5mo agoHugging Face07callofthenight1 /gaokao-sft-chinese-strict-abcd-v3 Gaokao SFT Chinese Strict ABCD V3 This dataset is the cleaned Chinese SFT release that keeps only single-choice samples where A, B, C, and D all have explicit option-level analysis. Composition Total samples: 88466 Train samples: 86670 Validation samples: 1796 Subject Counts { "biology": 33104, "chemistry": 35796, "english": 174, "general_exam": 7982, "geography": 888, "history": 229, "physics": 9986, "politics": 307 } Fields id… See the full description on the dataset page: https://huggingface.co/datasets/callofthenight1/gaokao-sft-chinese-strict-abcd-v3.texttext-generation10K<n<100K1 likes125 downloads5mo agoHugging Face08mixedbread-ai /incompebench-strictaudio10K<n<100K2 likes81 downloads7mo agoHugging Face09Lelonthecodeur /strict-verification-reasoning Strict Verification Reasoning Dataset Description A dataset for training language models to verify facts, check sources, evaluate arguments, and avoid overthinking. Content 1,010,000 examples 5 categories: anti-overthink, comparisons, strict facts, strict sources, strict arguments English language Categories Category Description % Anti-Overthink Simple, direct answers 15% Comparisons Hallucination vs correct answer 20%… See the full description on the dataset page: https://huggingface.co/datasets/Lelonthecodeur/strict-verification-reasoning.texttext-generation1M<n<10M1 likes74 downloads11d agoHugging Face10ShAIkespear /mmlu_gneissweb_strict_prunedtabular1K<n<10K0 likes59 downloads1y agoHugging Face11marin-community /open-thoughts-4-30k-math-qwen3-4b-annotated-32768-tokens-n8-rejection-sampling-strict-match N8 Rejection Sampling (Strict Match) Overview This dataset was created via rejection sampling from the Qwen3-4B response dataset using Qwen3-32B answers as ground truth. Source dataset (Qwen3-4B, 8 responses per prompt): marin-community/open-thoughts-4-30k-math-qwen3-4b-annotated-32768-tokens-n8-reformatted Verifier dataset (Qwen3-32B, 1 response per prompt): marin-community/open-thoughts-4-30k-math-qwen3-32b-annotated-32768-tokens Creator: The Marin Project… See the full description on the dataset page: https://huggingface.co/datasets/marin-community/open-thoughts-4-30k-math-qwen3-4b-annotated-32768-tokens-n8-rejection-sampling-strict-match.tabular10K<n<100K0 likes59 downloads7mo agoHugging Face12tintin1027 /atomic-metrics-ncw-strict Atomic Metrics NCW: Strict English Audit This version contains 100 training and 200 test preference pairs for practical nonfiction writing. Examples and existing preference labels are preserved; no replacement responses or preference labels were generated for this release. Sources Source Train Test Community Alignment 72 152 Writing Preference Bench 12 12 OASST1 5 23 OASST2 11 13 Source identifiers are retained in source_dataset and… See the full description on the dataset page: https://huggingface.co/datasets/tintin1027/atomic-metrics-ncw-strict.texttext-rankingn<1K0 likes57 downloads5d agoHugging Face13tuxevil /Home-Assistant-Requests-V5.2-Native-Strict Home Assistant Requests V5.2 Native Strict Private research dataset for supervised fine-tuning and regression testing of a small Home Assistant native tool-calling model. Contract: ha-native-tool-calling-v2. Frozen snapshot Split Rows Direct speech Multi-call Maximum rendered tokens train 3,806 340 78 3,098 validation 530 52 4 2,874 test 633 102 22 2,925 Tokenizer audit: model: unsloth/Qwen3-4B-Instruct-2507 revision:… See the full description on the dataset page: https://huggingface.co/datasets/tuxevil/Home-Assistant-Requests-V5.2-Native-Strict.texttext-generation1K<n<10K0 likes53 downloads2mo agoHugging Face14Qistinasofea /floorplan-aligned-strictimage1K<n<10K1 likes51 downloads9mo agoHugging Face15Pradheep1647 /strictly-speaking Strictly Speaking Does a model's mathematical understanding hold up strictly speaking, at Lean-grade precision, or is it only right in the ordinary, looser sense that informal writing usually gets away with? Each row is derived from a real, already-formalized Lean 4 theorem (drawn from Pradheep1647/lean-verifier-formalizations). The theorem's informal statement is split into two pieces: the hypotheses/setup (prompt), and the conclusion that was elided from it (answer) - the… See the full description on the dataset page: https://huggingface.co/datasets/Pradheep1647/strictly-speaking.texttext-generationn<1K0 likes46 downloads5d agoHugging Face16SoheylM /CAD-experiment-manifests-seed42-vision-qwen3vl32b-v1-strict-coder-v1image1K<n<10K0 likes44 downloads6mo agoHugging Face17kanishka /babylm2-rewritten-clean_multi-adj-strict-reversedtext10M<n<100M0 likes41 downloads1y agoHugging Face18callofthenight1 /gaokao-sft-chinese-strict-abcd Gaokao SFT Chinese Balanced This is the strict balanced Chinese SFT dataset version. Only multiple-choice samples with explicit A/B/C/D option-level explanations are kept in this balanced release. Composition Total samples: 645 Train samples: 632 Validation samples: 13 Subject Counts { "biology": 199, "chemistry": 170, "english": 13, "geography": 34, "history": 118, "physics": 77, "politics": 34 } Fields id lang subject source… See the full description on the dataset page: https://huggingface.co/datasets/callofthenight1/gaokao-sft-chinese-strict-abcd.texttext-generationn<1K0 likes41 downloads5mo agoHugging Face19Xalphinions /UltraFeedback_with_tie_stricttext10K<n<100K0 likes40 downloads2y agoHugging Face20VmaxRL /bugpilot-bugintro-lm-modify-gpt55-repaired-34-cleaned-1k-strict-20260526T212012Z SWE-Turing LM-Modify GPT-5.5 Strict 1k Provenance This document describes the generation, validation, cleaning, assembly, and upload for VmaxRL/bugpilot-bugintro-lm-modify-gpt55-repaired-34-cleaned-1k-strict-20260526T212012Z. Final Dataset Dataset: VmaxRL/bugpilot-bugintro-lm-modify-gpt55-repaired-34-cleaned-1k-strict-20260526T212012Z Created: 2026-05-26T21:20:20.354645+00:00 Split: train Rows: 1000 Allowed reliable universe:… See the full description on the dataset page: https://huggingface.co/datasets/VmaxRL/bugpilot-bugintro-lm-modify-gpt55-repaired-34-cleaned-1k-strict-20260526T212012Z.text1K<n<10K0 likes40 downloads4mo agoHugging Face21kanishka /babylm2-rewritten-clean_adj-num-strict-swappedtext10M<n<100M0 likes39 downloads1y agoHugging Face22kanishka /babylm2-rewritten-clean_no-multi-adj-stricttext10M<n<100M0 likes37 downloads1y agoHugging Face23kanishka /babylm2-rewritten-clean-spacy_multi-adj-strict-reversedtext10M<n<100M0 likes37 downloads1y agoHugging Face24sahilmob /wish-engine-toolcall-next-v3-strict-general wish-engine-toolcall-next-v3-strict-general Wish-engine implementor next-step tool-calling dataset (v3 strict generalization subset, dynamic aliases). Splits train.jsonl: 8081980 bytes validation.jsonl: 1008543 bytes test.jsonl: 997791 bytes Schema Rows are JSONL with at least: id messages (chat format with assistant tool_calls) tool_name metadata fields (mode, status, trajectory_*) Notes Tool names are dynamically aliased per sample. A tool… See the full description on the dataset page: https://huggingface.co/datasets/sahilmob/wish-engine-toolcall-next-v3-strict-general.tabularn<1K0 likes34 downloads7mo agoHugging Face25simonycl /cmv_2017_2025_strict_single_turn CMV 2017-2025: Strictly Single-Turn This local release is derived from simonycl/cmv_2017_2025_with_persona_0110. A source row is retained only if neither displayed comment ID occurs in any positive or negative chain in simonycl/cmv_multi_turn. This guarantees that neither paired argument is represented in the published multi-turn corpus. Split counts Split Input Excluded Retained train 25962 3779 22183 test 5427 434 4993 expert_train 3075 333 2742… See the full description on the dataset page: https://huggingface.co/datasets/simonycl/cmv_2017_2025_strict_single_turn.tabular10K<n<100K0 likes34 downloads2mo agoHugging Face26edithatogo /qwen3-hermes-strict-toolcall-synthetic-v4 Qwen3 Hermes Strict Tool-Call Synthetic V4 Registry status Registry ID: edithatogo/qwen3-hermes-strict-toolcall-synthetic-v4 Family: hermes Repository role: canonical_training_dataset Canonical dataset: edithatogo/qwen3-hermes-strict-toolcall-synthetic-v4 Operational status: active Rights status: apache-2.0-synthetic Authoritative catalog: edithatogo/dataset-estate-registry Origin and provenance Origin repository:… See the full description on the dataset page: https://huggingface.co/datasets/edithatogo/qwen3-hermes-strict-toolcall-synthetic-v4.texttext-generationn<1K0 likes33 downloads2mo agoHugging Face27tuxevil /Home-Assistant-Requests-V5.1-Native-Strict Home Assistant Requests V5.1 Native Strict Private research dataset for supervised fine-tuning and regression testing of a small Home Assistant native tool-calling model. Contract: ha-native-tool-calling-v2. Frozen snapshot Split Rows Direct speech Multi-call Maximum rendered tokens train 3,806 340 78 3,098 validation 530 52 4 2,874 test 633 102 22 2,925 Tokenizer audit: model: unsloth/Qwen3-4B-Instruct-2507 revision:… See the full description on the dataset page: https://huggingface.co/datasets/tuxevil/Home-Assistant-Requests-V5.1-Native-Strict.texttext-generation1K<n<10K0 likes32 downloads2mo agoHugging Face28graliuce /MedMCQA.20.01_terminate_aug_stricttext1K<n<10K0 likes31 downloads2mo agoHugging Face29kaiwenw /oct30_oasst_llama70b_jft_stricttext1K<n<10K0 likes30 downloads2y agoHugging Face30kanishka /babylm2-rewritten-clean-spacy_no-multi-adj-stricttext10M<n<100M0 likes29 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.