birgermoell/medqa-reasoning-traces
MedQA with LLM Reasoning Traces (kimi-k3) A derived dataset pairing every MedQA USMLE question with a reasoning trace produced by a large language model, the model's predicted answer, and a correctness flag scored against the gold label. Source questions: bigbio/med_qa, subset med_qa_en_4options_source (English, 4-option USMLE variant), splits train/validation/test Records: 12,723 (one per question × model) Overall accuracy: 94.4% (12,011/12,723) Why this dataset?… See the full description on the dataset page: https://huggingface.co/datasets/birgermoell/medqa-reasoning-traces.
037
Upload README.md with huggingface_hub
Upload dataset
initial commit
