CoolFace
Datasetpublic

upb-nlp/EduMUSE

OpenStax Multimodal Exercise Dataset A multimodal dataset of textbook exercises scraped from OpenStax, each aligned to its most relevant textbook subsection and scored under several open vision-language models. The dataset enables research on retrieval-augmented question answering, the contribution of visual context to scientific QA, and ablation studies on text-only vs. multimodal context. Contents final_EduMUSE_dataset.json — the unified dataset (nested by book… See the full description on the dataset page: https://huggingface.co/datasets/upb-nlp/EduMUSE.

sourceHugging Facecc-by-4.0updated 5mo agoView on Hugging Face
1likes111downloads
settings

This repository belongs to upb-nlp on Hugging Face.

CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.

nameEduMUSE
visibilitypublic
licencecc-by-4.0
gatedno
ownerupb-nlp
Account settings
upb-nlp/EduMUSE · CoolFace