datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
pisa-bench
Dataset Card for PISA-Bench
Paper: https://arxiv.org/abs/2510.24792Authors: Patrick Haller, Fabio Barth, Jonas Golde, Georg Rehm, Alan Akbik
Dataset Summary
PISA-Bench is a multilingual, multimodal benchmark constructed from expert-authored PISA exam questions.Each example is a human-created educational reasoning problem containing an image and a reading/math question, translated into six languages:
English (EN)
German (DE)
Spanish (ES)
French (FR)
Italian (IT)
Chinese… See the full description on the dataset page: https://huggingface.co/datasets/PisaBench/pisa-bench.PISA_tests
Dataset Card: PISA Multimodal (Parallel & Not-Parallel)
Summary
This dataset contains 48 parallel multimodal samples (paired TXT↔PDF) derived from PISA studies up to 2012, plus 47 non-parallel samples (TXT-only or PDF-only). Each sample may include multiple questions. Content is available in German and English.
Source & usage: Materials are published by the OECD and are provided here for non-commercial use only. Please verify that your usage complies with OECD terms.… See the full description on the dataset page: https://huggingface.co/datasets/barthfab/PISA_tests.FVELer_PISA_NotProvenFVELer_PISA_Proven
