datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
MedQA-USMLE-4-options-hfusmle
USMLE Questions Dataset
Evaluation dataset of USMLE-style multiple choice questions.
Dataset structure
Each row has:
question: Question statement
options: List of options (e.g. ["(A) ...", "(B) ...", ...])
correct_answer: Correct option letter (A–E)
Loading
from datasets import load_dataset
ds = load_dataset("prapaa/usmle", data_files="{"train": "train.jsonl"}, split="train")
