datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
amc12-full
AMC12 Dataset (Research-Oriented)
A structured dataset derived from the AMC 12 (American Mathematics Competitions), designed for LLM training, evaluation, and reinforcement learning (RL) on mathematical reasoning tasks.
This repository contains all AMC 12 problems from 2000–2025, making it one of the most complete AMC12 datasets available for research.
📘 Introduction
The AMC 12 is a 25-question, 75-minute multiple-choice examination aimed at high school… See the full description on the dataset page: https://huggingface.co/datasets/edev2000/amc12-full.2024_AMC12All problems copyrighted by the Mathematical Association of America's American Mathematics Competitions
Source:
https://artofproblemsolving.com/wiki/index.php/2024_AMC_12A_Problems
https://artofproblemsolving.com/wiki/index.php/2024_AMC_12B_Problems
Removed problems with figures:
12A: problem 14,18,22
12B: problem 7, 19
R-HORIZON-AMC23
R-HORIZON
How Far Can Your Large Reasoning Model Really Go in Breadth and Depth?
📃 Paper • 🌐 Project Page • 🤗 Dataset
R-HORIZON is a novel method designed to stimulate long-horizon reasoning behaviors in Large Reasoning Models (LRMs) through query composition. We transform isolated problems into complex multi-step reasoning scenarios, revealing that even the most advanced LRMs suffer significant performance degradation when facing interdependent problems that span… See the full description on the dataset page: https://huggingface.co/datasets/meituan-longcat/R-HORIZON-AMC23.amc_aime_self_improving
Additional Information
This dataset contains mathematical problem-solving traces generated using the CAMEL framework. Each entry includes:
A mathematical problem statement
A detailed step-by-step solution
An improvement history showing how the solution was iteratively refined
Special thanks to our community contributor, GitHoobar, for developing the STaR pipeline!🙌
amc12_22-24amc12-full
AMC12 Dataset (Research-Oriented)
A structured dataset derived from the AMC 12 (American Mathematics Competitions), designed for LLM training, evaluation, and reinforcement learning (RL) on mathematical reasoning tasks.
This repository contains all AMC 12 problems from 2000–2025, making it one of the most complete AMC12 datasets available for research.
📘 Introduction
The AMC 12 is a 25-question, 75-minute multiple-choice examination aimed at high school… See the full description on the dataset page: https://huggingface.co/datasets/greenstainedglass/amc12-full.amc_aime_distilled
Additional Information
This dataset contains mathematical problem-solving traces generated using the CAMEL framework. Each entry includes:
A mathematical problem statement
A detailed step-by-step solution
Asan-AMC-Healthinfo
Asan-AMC-Healthinfo
Source: 서울아산병원 건강정보.
서울아산병원 홈페이지의 건강정보-의료정보에서, 인체정보/질환백과/검사시술수술정보/알기쉬운의학용어/식사요법 데이터를 바탕으로
alpaca-style로 편집한 데이터입니다.
amc-tutor-sft
AMC Tutor — decontaminated competition-math SFT dataset
Chat-formatted, decontaminated supervised-fine-tuning data for AMC 10/12-style
competition mathematics. Built for a reproducible $0, local (MacBook M4) study of QLoRA
fine-tuning small models. Each row is a tutor system prompt + problem + step-by-step solution
ending in Final answer: \boxed{...}.
Companion study & code: https://github.com/RoyK0108/amc-tutor-study
⚠️ This is a study artifact — read the finding… See the full description on the dataset page: https://huggingface.co/datasets/Roykim7/amc-tutor-sft.amc22-24_stop_stringsAMC-12(2022-2024)
amcR-HORIZON-AMC23AMC23boxed-amcAMC-Test-Ko
Dataset Card for AIMO Validation AMC
All 83 come from AMC12 2022, AMC12 2023, and have been extracted from the AOPS wiki page https://artofproblemsolving.com/wiki/index.php/AMC_12_Problems_and_Solutions
This dataset serves as an internal validation set during our participation in the AIMO progress prize competition. Using data after 2021 is to avoid potential overlap with the MATH training set.
Here are the different columns in the dataset:
problem: the modified problem statement… See the full description on the dataset page: https://huggingface.co/datasets/ChuGyouk/AMC-Test-Ko.amc23-rolloutsaimo-validation-amc-repeated3tts-embed-dataset-amc23amc23_repeated10TDAR_Eval-AMC23amc23-convertedAMC23-instruct
