CoolFace
Datasetpublic

lime-nlp/orz_math_difficulty

Difficulty Estimation on Open Reasoner Zero We annotate the entire Open Reasoner Zero dataset with a difficulty score based on the performance of the Qwen 2.5-MATH-7B model. This provides an adaptive signal for curriculum construction. Open Reasoner Zero is a curated a dataset of 57,000 reasoning-intensive problems used to train and evaluate reinforcement learning-based methods for large language models. Difficulty Scoring Method Difficulty scores are estimated… See the full description on the dataset page: https://huggingface.co/datasets/lime-nlp/orz_math_difficulty.

sourceHugging Facemitupdated 1y agoView on Hugging Face
0likes44downloads

lime-nlp/orz_math_difficulty · main · files are served by the source, never re-hosted here