WanyueZhang/MulSeT
MulSeT: A Benchmark for Multi-view Spatial Understanding Tasks Paper: Why Do MLLMs Struggle with Spatial Understanding? A Systematic Analysis from Data to Architecture Code: https://github.com/WanyueZhang-ai/spatial-understanding A high-level overview of the MulSeT benchmark. The dataset challenges models to integrate information from two distinct viewpoints of a 3D scene to answer spatial reasoning questions. 📝 Dataset Summary MulSeT is a comprehensive… See the full description on the dataset page: https://huggingface.co/datasets/WanyueZhang/MulSeT.
686k
