oscarqjh/MMSI-Video-Bench_lmmseval
MMSI-Video-Bench A video-based spatial intelligence benchmark for evaluating Multimodal Large Language Models (MLLMs). Dataset Description MMSI-Video-Bench tests models on: Spatial reasoning Motion understanding Planning and prediction Cross-video reasoning Dataset Structure MMSI-Video-Bench/ ├── data/ │ └── test-00000-of-00001.parquet # 1106 samples ├── frames.zip # Extracted video frames ├── ref_images.zip… See the full description on the dataset page: https://huggingface.co/datasets/oscarqjh/MMSI-Video-Bench_lmmseval.
Conversations for this repository live on Hugging Face.
CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.
Open discussions on Hugging Face