datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
intellect-3-rl-math-5k
intellect-3-rl-math-5k
A difficulty-stratified sample of 5,000 unique math problems drawn from PrimeIntellect/INTELLECT-3-RL.
How it was drawn
Source pool: 12kimih/intellect-3-rl-math-decontaminated, 20,918 unique problems, the math config, decontaminated against the benchmarks listed below.
Stratum: stratum, the number of 8 attempts by Qwen3-4B-Thinking-2507 that matched the reference answer, shipped per problem by the upstream. It runs 0 (never solved) to 8… See the full description on the dataset page: https://huggingface.co/datasets/12kimih/intellect-3-rl-math-5k.intellect-3-rl-math-decontaminated
intellect-3-rl-math-decontaminated
20,918 unique math problems from PrimeIntellect/INTELLECT-3-RL with 243 removed as contaminated.
How it was prepared
Pool: the math config of INTELLECT-3-RL, 21,161 rows, normalised to the column names used here.
Rows: 21,161 loaded, 20,918 kept.
Decontaminated against: math500, aime2024, aime2025, aime2026, amc, hmmt_feb2023, hmmt_feb2024, hmmt_feb2025, hmmt_feb2026, hmmt_nov2025, olympiadbench, gsm8k (243 problems removed… See the full description on the dataset page: https://huggingface.co/datasets/12kimih/intellect-3-rl-math-decontaminated.
