CoolFace
Datasetpublic

Yuan4629/WereBench

Anonymization For all content in this Hugging Face dataset repository and GitHub repository, we have ensured that anonymization has been performed, making it impossible to trace back to the authors' information. WereBench WereBench is a benchmark dataset for evaluating language models in the Werewolf (similar to Mafia) social deduction setting. It focuses on human‑aligned strategic reasoning rather than only coarse metrics (e.g., win rate), aligning model behavior… See the full description on the dataset page: https://huggingface.co/datasets/Yuan4629/WereBench.

sourceHugging Faceupdated 9mo agoView on Hugging Face
1likes127downloads

Yuan4629/WereBench · main · files are served by the source, never re-hosted here