frozen-llms-play-werewolf-competently-via-reflectionobservationsingle paper
Frozen LLMs, using only retrieval over past communications and self-reflection rather than fine-tuning, can play the social deduction game Werewolf competently and show emergent strategic behavior, including deception, without being explicitly trained for it.
Capability: Strategic deception and detecting it · Behavior, Agentic
Observed on
2023, GPT-3.5/GPT-4 class. Autonomous agent.
Sources
- A tuning-free framework using retrieval and reflection on past communications let frozen LLMs play Werewolf competently, with emergent strategic behavior.
Disagreeing is the most useful thing you can do here. Both sides of every contested claim in this catalog were assembled by the same person, which is its weakest point.