Team Ai
Datasetpublic

n0nam4/WereBench

Anonymization For all content in this Hugging Face dataset repository and GitHub repository, we have ensured that anonymization has been performed, making it impossible to trace back to the authors' information. WereBench WereBench is a benchmark dataset for evaluating language models in the Werewolf (similar to Mafia) social deduction setting. It focuses on human‑aligned strategic reasoning rather than only coarse metrics (e.g., win rate), aligning model behavior… See the full description on the dataset page: https://huggingface.co/datasets/n0nam4/WereBench.

sourceHugging Faceupdated 9mo agoView on Hugging Face
0likes162downloads
settings

This repository belongs to n0nam4 on Hugging Face.

Team Ai never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.

nameWereBench
visibilitypublic
licencenot set
gatedno
ownern0nam4
Account settings
n0nam4/WereBench · Team Ai