xlzhou126/SpaceDG-Bench
SpaceDG-Bench 🌐 Homepage | 📖 arXiv | 💻 GitHub SpaceDG-Bench is a human-verified benchmark designed to evaluate the spatial intelligence of Multimodal Large Language Models (MLLMs) under visual degradation. It contains 1,102 questions spanning 11 reasoning categories and 9 visual degradation types (such as motion blur, low light, adverse weather, lens distortion, and compression artifacts), yielding over 10K VQA instances. The benchmark is part of the SpaceDG project, which… See the full description on the dataset page: https://huggingface.co/datasets/xlzhou126/SpaceDG-Bench.
141
Update README.md
Improve dataset card: add paper/project links and metadata (#2)
Duplicate from SpaceDG/SpaceDG-Bench
