datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
cadqa-rl-2000
CAD-QA RL Training Set (2000 samples)
Verifiable-reward RL training rows for CAD geometry reasoning, derived from the
CAD-QA benchmark (release_10k train split, cad_browsecomp difficulty).
Each row is a closed-book question: a CadQuery script + a question about the
resulting geometry. The gold answer is exact-match verifiable.
Files
data/train/ — 2000 rows
data/validation/ — 98 rows (the benchmark's eval-slice rows; excluded
from training — do not train on this… See the full description on the dataset page: https://huggingface.co/datasets/MRiabov/cadqa-rl-2000.MRI-MCQA
MRI-MCQA
Dataset Description
MRI-MCQA is a benchmark composed by multiple-choice questions related to Magnetic Resonance Imaging (MRI). We use this dataset to evaluate the level of knowledge of various LLMs about the MRI field.
Curated by: Oscar Molina Sedano
Language(s) (NLP): English
License
This dataset is licensed under CC-BY-NC 4.0.
Disclaimer
Courtesy of Allen D. Elster… See the full description on the dataset page: https://huggingface.co/datasets/HPAI-BSC/MRI-MCQA.amazon_product_reviews_datafiniti
Dataset Card for "amazon_product_reviews_datafiniti"
More Information needed
IntersectionQA-90K
IntersectionQA-90K
Dataset Summary
IntersectionQA is a code-only CAD spatial-reasoning benchmark. Each example gives a model two executable CadQuery object-construction functions plus assembly transforms, then asks it to infer the geometric relation induced by that code. The central question is whether a code model can mentally track the spatial consequences of CAD programs: positive-volume interference, contact, near misses, clearance, containment, and overlap magnitude.… See the full description on the dataset page: https://huggingface.co/datasets/MRiabov/IntersectionQA-90K.IntersectionQA-15K
IntersectionQA-15K
Dataset Summary
IntersectionQA is a code-only CAD spatial-reasoning benchmark. Each example gives a model two executable CadQuery object-construction functions plus assembly transforms, then asks it to infer the geometric relation induced by that code. The central question is whether a code model can mentally track the spatial consequences of CAD programs: positive-volume interference, contact, near misses, clearance, containment, and overlap magnitude.… See the full description on the dataset page: https://huggingface.co/datasets/MRiabov/IntersectionQA-15K.global-MMLU-MRI
Global MMLU Lite - English/Maori Bilingual Dataset
Dataset Description
This dataset contains the Global MMLU Lite questions in both English and Maori (Te Reo Māori). It merges the original English dataset from CohereLabs/Global-MMLU-Lite with Google-translated Maori versions.
Dataset Structure
Each example contains:
sample_id: Unique identifier for the question
question_en: Question in English
option_a_en, option_b_en, option_c_en, option_d_en: Answer options… See the full description on the dataset page: https://huggingface.co/datasets/whamidou/global-MMLU-MRI.m-ric_huggingface_doc_347
