datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
MedThinkVQA
MedThinkVQA
MedThinkVQA is an expert-annotated benchmark for multi-image diagnostic reasoning in radiology. Unlike prior medical VQA benchmarks that typically contain at most one image per case, MedThinkVQA requires models to extract evidence from each image, integrate cross-view information, and perform differential-diagnosis reasoning.
Links
GitHub: https://github.com/benluwang/MedThinkVQA
Leaderboard: https://benluwang.github.io/MedThinkVQA/
Submission Guide:… See the full description on the dataset page: https://huggingface.co/datasets/bio-nlp-umass/MedThinkVQA.rrg24-shared-task-bionlp
✏️ Citation
@inproceedings{xu-etal-2024-overview,
title = "Overview of the First Shared Task on Clinical Text Generation: {RRG}24 and {\textquotedblleft}Discharge Me!{\textquotedblright}",
author = "Xu, Justin and
Chen, Zhihong and
Johnston, Andrew and
Blankemeier, Louis and
Varma, Maya and
Hom, Jason and
Collins, William J. and
Modi, Ankit and
Lloyd, Robert and
Hopkins, Benjamin and
Langlotz, Curtis and… See the full description on the dataset page: https://huggingface.co/datasets/StanfordAIMI/rrg24-shared-task-bionlp.
