2 citations · 2 across the 1 of their papers we have counts for
1 paper
Suma Bailis, Jane Friedhoff, Feiyang Chen
This paper introduces Werewolf Arena, a novel framework for evaluating large language models (LLMs) through the lens of the classic social deduction game, Werewolf. In Werewolf Are…