datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
MBA-Bench
MBA-Bench
Official benchmark dataset for MBA: Multimodal Benchmark and Agents for Real-World Business Ideation.
Paper: https://arxiv.org/abs/2608.11616
Project Page: https://hchoi256.github.io/projects/mba/
Code: https://github.com/hchoi256/mba
Models: https://huggingface.co/hchoi256/mba
Dataset Description
MBA-Bench is a multimodal benchmark for evaluating real-world business ideation capabilities.
The dataset contains multimodal benchmark samples together with… See the full description on the dataset page: https://huggingface.co/datasets/hchoi256/MBA-Bench.hc-hvlm-visiblehc-hvlm-graderSoMi-ToM
SoMi-ToM: Evaluating Multi-Perspective Theory of Mind in Embodied Social Interactions
NeurIPS 2025
❤️ Like our project on Hugging Face to show your support!
SoMi-ToM Benchmark
We propose the SoMi-ToM benchmark, designed to evaluate multi-perspective ToM in embodied multi-agent complex social interactions. This benchmark is based on rich multimodal interaction data generated by the interaction environment SoMi… See the full description on the dataset page: https://huggingface.co/datasets/hch2000/SoMi-ToM.test_qd_embeddingstest_qd_env_traj_dataset_3.6.0aitch01
