SCAI-JHU/MUMA-TOM-BENCHMARK
MuMA-ToM: Multi-modal Multi-Agent Theory of Mind AAAI 2025 (Oral) [🏠Homepage] [💻Code] [📝Paper] MuMA-ToM is the first multi-modal Theory of Mind benchmark designed to evaluate mental reasoning in embodied multi-agent interactions. The benchmark was designed with several key features in mind: It is factually correct, concise, and readable. It requires integrating information from multiple modalities to answer the questions. It tests understanding of multi-agent interactions… See the full description on the dataset page: https://huggingface.co/datasets/SCAI-JHU/MUMA-TOM-BENCHMARK.
MuMA-ToM: Multi-modal Multi-Agent Theory of Mind <br> <sub> AAAI 2025 (Oral) </sub>
[\[🏠Homepage\]](https://scai.cs.jhu.edu/projects/MuMA-ToM/) [\[💻Code\]](https://github.com/SCAI-JHU/MuMA-ToM) [\[📝Paper\]](https://arxiv.org/abs/2408.12574)
MuMA-ToM is the first multi-modal Theory of Mind benchmark designed to evaluate mental reasoning in embodied multi-agent interactions. The benchmark was designed with several key features in mind:
- It is factually correct, concise, and readable.
- It requires integrating information from multiple modalities to answer the questions.
- It tests understanding of multi-agent interactions, including beliefs, social goals, and beliefs about others' goals.
Leaderboard
Here is the **leaderboard** for MuMA-ToM. Please contact us if you'd like to add your results.
Citation
Please cite the paper if you find it interesting/useful, thanks!
@inproceedings{shi2025muma,
title={Muma-tom: Multi-modal multi-agent theory of mind},
author={Shi, Haojun and Ye, Suyu and Fang, Xinyu and Jin, Chuanyang and Isik, Leyla and Kuo, Yen-Ling and Shu, Tianmin},
booktitle={Proceedings of the AAAI Conference on Artificial Intelligence},
volume={39},
number={2},
pages={1510--1519},
year={2025}
}