datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
Multi-subject-RLVRMulti-subject data for paper "Expanding RL with Verifiable Rewards Across Diverse Domains".
we use a multi-subject multiple-choice QA dataset ExamQA (Yu et al., 2021).
Originally written in Chinese, ExamQA covers at least 48 first-level subjects.
We remove the distractors and convert each instance into a free-form QA pair.
This dataset consists of 638k college-level instances, with both questions and objective answers written by domain experts for examination purposes.
We also use GPT-4o-mini… See the full description on the dataset page: https://huggingface.co/datasets/virtuoussy/Multi-subject-RLVR.multi-subject-mcq-training-pool
Multi-subject multiple-choice training pool
Public multiple-choice questions from eight datasets covering medicine, law, quantitative
reasoning, the natural sciences, history, philosophy, business, economics and everyday knowledge,
read at the pinned revisions named below and laid out twice. Train on either layer or on both.
pool.jsonl
Every source rewritten into one shape, 347642 rows, one JSON object per line, with these fields.
Field
What it holds
id… See the full description on the dataset page: https://huggingface.co/datasets/Emulated-Inc/multi-subject-mcq-training-pool.Multi-subject-RLVR-annotated
