CoolFace
Datasetpublic

n0nam4/WereBench

Anonymization For all content in this Hugging Face dataset repository and GitHub repository, we have ensured that anonymization has been performed, making it impossible to trace back to the authors' information. WereBench WereBench is a benchmark dataset for evaluating language models in the Werewolf (similar to Mafia) social deduction setting. It focuses on human‑aligned strategic reasoning rather than only coarse metrics (e.g., win rate), aligning model behavior… See the full description on the dataset page: https://huggingface.co/datasets/n0nam4/WereBench.

sourceHugging Faceupdated 9mo agoView on Hugging Face
0likes139downloads
12 commits on main
7ca50d69mo ago

Update README.md

Yuan4629
c735d8910mo ago

Upload folder using huggingface_hub

Yuan4629
f56803410mo ago

Upload folder using huggingface_hub

Yuan4629
343618911mo ago

add Season3

Yuan4629
76a889411mo ago

Anonymization

Yuan4629
632189f11mo ago

Update README.md

Yuan4629
1ce45ae11mo ago

Update README.md

Yuan4629
dc7b23c11mo ago

Update README.md

Yuan4629
76ade8a11mo ago

convert WereBench.json to .jsonl with json2jsonl.py

Yuan4629
6410f2b11mo ago

Add English dataset card (README.md)

Yuan4629
585e8f111mo ago

Initial dataset upload

Yuan4629
77c854511mo ago

initial commit

Yuan4629