Shamant23/tom-benchmark
UniToMBench Dataset Dataset Summary UniToMBench is a unified benchmark designed to evaluate the Theory of Mind (ToM) reasoning abilities of large language models (LLMs). It integrates and extends existing ToM benchmarks by providing narrative-based multiple-choice questions (MCQs) that span a wide range of ToM tasks—including false belief reasoning, perspective taking, emotion attribution, scalar implicature, and more. This dataset supports the research paper… See the full description on the dataset page: https://huggingface.co/datasets/Shamant23/tom-benchmark.
This repository belongs to Shamant23 on Hugging Face.
CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.
