datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
MMVU
MMVU: Measuring Expert-Level Multi-Discipline Video Understanding
🌐 Homepage •
🥇 Leaderboard •
📖 Paper •
🤗 Data
📰 News
2025-01-21: We are excited to release the MMVU paper, dataset, and evaluation code!
👋 Overview
Why MMVU Benchmark?
Despite the rapid progress of foundation models in both text-based and image-based expert reasoning, there is a clear gap in evaluating these models’ capabilities in specialized-domain video understanding.… See the full description on the dataset page: https://huggingface.co/datasets/yale-nlp/MMVU.MMV-dataset
MMV-dataset
Reproducible research splits and full-training subsets derived from long-video benchmarks. No video, audio, or subtitle media is included.
Hosted data
Component
Train
Validation
Grouping unit
License
LongVideoBench labeled validation set
401
936
video_id
CC BY-NC-SA 4.0
EgoTempo open-ended QA
150
350
original Ego4D video UID
CC BY 4.0
Unified supervised records
1,084
2,529
inherited
mixed; see per-row license
The LongVideoBench… See the full description on the dataset page: https://huggingface.co/datasets/czty/MMV-dataset.MMVUMM-Verify-Data
