← all papers · overview

Werewolf Arena: A Case Study In LLM Evaluation Via Social Deduction

Abstract

This paper introduces Werewolf Arena, a novel framework for evaluating large language models (LLMs) through the lens of the classic social deduction game, Werewolf. In Werewolf Arena, LLMs compete against each other, navigating the game's complex dynamics of deception, deduction, and persuasion. The framework introduces a dynamic turn-taking system based on bidding, mirroring real-world discussion

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).