CoolFace
Datasetpublic

birgermoell/medqa-reasoning-traces

MedQA with LLM Reasoning Traces (kimi-k3) A derived dataset pairing every MedQA USMLE question with a reasoning trace produced by a large language model, the model's predicted answer, and a correctness flag scored against the gold label. Source questions: bigbio/med_qa, subset med_qa_en_4options_source (English, 4-option USMLE variant), splits train/validation/test Records: 12,723 (one per question × model) Overall accuracy: 94.4% (12,011/12,723) Why this dataset?… See the full description on the dataset page: https://huggingface.co/datasets/birgermoell/medqa-reasoning-traces.

sourceHugging Faceotherupdated 4d agoView on Hugging Face
0likes37downloads
3 commits on main
82fe04a4d ago

Upload README.md with huggingface_hub

birgermoell
7ca7e0f4d ago

Upload dataset

birgermoell
639c59c4d ago

initial commit

birgermoell