The Hidden Rules of Hanabi: How Humans Outperform AI Agents
Matthew Sidji, Wally Smith, Melissa J. Rogerson
Abstract
Games that feature multiple players, limited communication, and partial information are particularly challenging for AI agents. In the cooperative card game Hanabi, which possesses all of these attributes, AI agents fail to achieve scores comparable to even first-time human players. Through an observational study of three mixed-skill Hanabi play groups, we identify the techniques used by humans that help to explain their superior performance compared to AI. These concern physical artefact manipulation, coordination play, role establishment, and continual rule negotiation. Our findings extend previous accounts of human performance in Hanabi, which are purely in terms of theory-of-mind reasoning, by revealing more precisely how this form of collective decision-making is enacted in skilled human play. Our interpretation points to a gap in the current capabilities of AI agents to perform cooperative tasks.
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Related papers
- Improving Policies via Search in Cooperative Partially Observable GamesAdam Lerer, Hengyuan Hu, Jakob N. Foerster, Noam BrownAAAI 2020 · 87 citations
- Simplified Action Decoder for Deep Multi-Agent Reinforcement LearningHengyuan Hu, Jakob N. FoersterICLR 2020 · 88 citations
- Ad-Hoc Human-AI Coordination ChallengeTin Dizdarevic, Ravi Hammond, Tobias Gessler, Anisoara Calinescu et al.ICML 2025
- Sparks of Cooperative Reasoning: LLMs as Strategic Hanabi AgentsMahesh Ramesh, Kaousheik Jayakumar, Aswinkumar Ramkumar, Pavan Thodima et al.ICML 2026
- Evaluation of Human-AI Teams for Learned and Rule-Based Agents in HanabiHo Chit Siu, Jaime Daniel Peña, Edenna Chen, Yutai Zhou et al.NeurIPS 2021 · 78 citations
