birgermoell/medqa-reasoning-traces
MedQA with LLM Reasoning Traces (kimi-k3) A derived dataset pairing every MedQA USMLE question with a reasoning trace produced by a large language model, the model's predicted answer, and a correctness flag scored against the gold label. Source questions: bigbio/med_qa, subset med_qa_en_4options_source (English, 4-option USMLE variant), splits train/validation/test Records: 12,723 (one per question × model) Overall accuracy: 94.4% (12,011/12,723) Why this dataset?… See the full description on the dataset page: https://huggingface.co/datasets/birgermoell/medqa-reasoning-traces.
Conversations for this repository live on Hugging Face.
CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.
Open discussions on Hugging Face