CoolFace
Datasetpublic

xlzhou126/SpaceDG-Bench

SpaceDG-Bench 🌐 Homepage | 📖 arXiv | 💻 GitHub SpaceDG-Bench is a human-verified benchmark designed to evaluate the spatial intelligence of Multimodal Large Language Models (MLLMs) under visual degradation. It contains 1,102 questions spanning 11 reasoning categories and 9 visual degradation types (such as motion blur, low light, adverse weather, lens distortion, and compression artifacts), yielding over 10K VQA instances. The benchmark is part of the SpaceDG project, which… See the full description on the dataset page: https://huggingface.co/datasets/xlzhou126/SpaceDG-Bench.

sourceHugging Facecc-by-4.0updated 3mo agoView on Hugging Face
1likes41downloads
3 commits on main
4be32bb3mo ago

Update README.md

xlzhou126
55e15db4mo ago

Improve dataset card: add paper/project links and metadata (#2)

xlzhou126, nielsr
f5782e35mo ago

Duplicate from SpaceDG/SpaceDG-Bench

xlzhou126, SpaceDG