CoolFace
17 results

numina

AI-MO /NuminaMath-CoT Dataset Card for NuminaMath CoT Dataset Summary Approximately 860k math problems, where each solution is formatted in a Chain of Thought (CoT) manner. The sources of the dataset range from Chinese high school math exercises to US and international mathematics olympiad competition problems. The data were primarily collected from online exam paper PDFs and mathematics discussion forums. The processing steps include (a) OCR from the original PDFs, (b) segmentation… See the full description on the dataset page: https://huggingface.co/datasets/AI-MO/NuminaMath-CoT.texttext-generation100K<n<1M603 likes224k downloads2y agoHugging FaceAI-MO /NuminaMath-1.5 Dataset Card for NuminaMath 1.5 Dataset Summary This is the second iteration of the popular NuminaMath dataset, bringing high quality post-training data for approximately 900k competition-level math problems. Each solution is formatted in a Chain of Thought (CoT) manner. The sources of the dataset range from Chinese high school math exercises to US and international mathematics olympiad competition problems. The data were primarily collected from online exam paper PDFs… See the full description on the dataset page: https://huggingface.co/datasets/AI-MO/NuminaMath-1.5.texttext-generation100K<n<1M194 likes49k downloads8mo agoHugging Facenlile /NuminaMath-1.5-RL-Verifiable Dataset Card for NuminaMath-1.5-RL-Verifiable Dataset Summary NuminaMath-1.5-RL-Verifiable is a curated subset of the NuminaMath-1.5 dataset, specifically filtered to support reinforcement learning applications requiring verifiable outcomes. This collection consists of 131,063 math word problems from the original dataset that meet strict filtering criteria: all problems have definitive numerical answers, validated problem statements and solutions, and come from… See the full description on the dataset page: https://huggingface.co/datasets/nlile/NuminaMath-1.5-RL-Verifiable.texttext-generation100K<n<1M10 likes8.7k downloads1y agoHugging FaceAI-MO /NuminaMath-TIR Dataset Card for NuminaMath CoT Dataset Summary Tool-integrated reasoning (TIR) plays a crucial role in this competition. However, collecting and annotating such data is both costly and time-consuming. To address this, we selected approximately 70k problems from the NuminaMath-CoT dataset, focusing on those with numerical outputs, most of which are integers. We then utilized a pipeline leveraging GPT-4 to generate TORA-like reasoning paths, executing the code and… See the full description on the dataset page: https://huggingface.co/datasets/AI-MO/NuminaMath-TIR.texttext-generation10K<n<100K158 likes8.2k downloads2y agoHugging Facenlile /NuminaMath-1.5-proofs-only-strict NuminaMath-1.5-proofs-only-strict A strictly filtered version of the NuminaMath-1.5-proofs-only dataset, containing ONLY validated mathematical proof problems. 📊 Filtering Results Original dataset: Numina1.5 -> filter for proofs -> 110,998 rows Filters applied: ✓ Kept rows where answer = "proof" (proof problems only) ✓ Kept rows where solution_is_valid = "Yes" ✓ Kept rows where problem_is_valid = "Yes" ✓ Dropped validation columns after filtering Filtered dataset:… See the full description on the dataset page: https://huggingface.co/datasets/nlile/NuminaMath-1.5-proofs-only-strict.text10K<n<100K1 likes3.2k downloads1y agoHugging Facedougalldeepmind /2026-09-11-dh-qwen3-6-27b-lora-9284-numina-control-716-r64 Delegated-harm evaluation with corrected scoring of saved rollouts field value experiment Delegated-harm evaluation with corrected scoring of saved rollouts date_generated 2026-09-11 constitution none source_repo teaching_claude_why_replication @ d627d0587a2980069b7700e72f727dae594c9f49 models {"hf_path": "matboz/qwen3.6-27b-lora-9284-numina-control-716-r64", "base_model": "Qwen/Qwen3.6-27B", "adapter": true, "mode": "think", "model_key":… See the full description on the dataset page: https://huggingface.co/datasets/dougalldeepmind/2026-09-11-dh-qwen3-6-27b-lora-9284-numina-control-716-r64.0 likes2.8k downloads11d agoHugging Face