datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
Ai_education_datasetpedagogy-benchmark
Dataset Card for The Pedagogy Benchmark
Dataset Summary
This dataset provides the questions for the pedgagoy benchmarks described in Benchmarking the Pedagogical Knowledge of Large Language Models.
These are questions from teacher training exams, which are designed to evaluate large language models on their Cross-Domain Pedagogical Knowledge (CDPK) and Special Education Needs and Disability (SEND) pedagogical knowledge.
Existing benchmarks have largely focused on content… See the full description on the dataset page: https://huggingface.co/datasets/AI-for-Education/pedagogy-benchmark.educational-ai-agent-small-annotation-depth
