datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
numeric-claim-verifier
Numeric Claim Verifier (Science) — Adaption AutoScientist
Programmatically verified prompt/completion pairs for scientific and statistical numeric claim verification.
Labels
correct — claim matches ground-truth tables
wrong_direction — trend/sign reversed
wrong_magnitude — right direction, wrong size (25–70% offset)
unverifiable — no matching source row (real entity + absent metric)
Sources
Our World in Data CO₂ / Energy
WHO GHO life expectancy… See the full description on the dataset page: https://huggingface.co/datasets/mishface123/numeric-claim-verifier.m9-verifier-38k-aligned
M9 Verifier 38K Aligned
This dataset contains 38,564 prompts with verifier-compatible gold answers for
an M9 RLVR-GRPO experiment in a unified post-training study with Qwen3-1.7B.
It is an independent research artifact, not an official release from the model
or paper authors.
The bank was reconstructed from the frozen
YangyiH/openreasoning_mixed_100k
prompt mixture. Every recovered row was matched to the frozen base row by
domain, source shard, and prompt SHA-256 before verifier… See the full description on the dataset page: https://huggingface.co/datasets/YangyiH/m9-verifier-38k-aligned.
