CoopEval: Benchmarking Cooperation-Sustaining Mechanisms and LLM Agents in Social Dilemmas
Emanuel Tewolde, Xiao Zhang, David Guzman Piedrahita, Vincent Conitzer, Zhijing Jin
摘要
It is increasingly important that LLM agents interact effectively and safely with other goal-pursuing agents, yet, recent works report the opposite trend: LLMs with stronger reasoning capabilities behave less cooperatively in mixed-motive games such as the prisoner's dilemma and public goods settings. Indeed, our experiments show that recent models---with or without reasoning enabled---consistently defect in single-shot social dilemmas. To tackle this safety concern, we present the first comparative study of game-theoretic mechanisms designed to enable cooperative outcomes between rational agents in equilibrium . Across four social dilemmas testing distinct components of robust cooperation, we evaluate four families of mechanisms: (1) repeating the game for many rounds, (2) reputation systems, (3) third-party mediators to delegate decision making to, and (4) contract agreements for outcome-conditional payments between players. Among our findings, we establish that contracting and mediation are most effective in achieving cooperative outcomes between capable LLM models, and that repetition-induced cooperation deteriorates drastically when co-players vary. Moreover, we demonstrate that the mechanisms become more effective under evolutionary pressures to maximize individual payoffs.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper13
- Cooperate or Collapse: Emergence of Sustainable Cooperation in a Society of LLM AgentsGiorgio Piatti, Zhijing Jin, Max Kleiman-Weiner, Bernhard Schölkopf 等NeurIPS 2024 · 被引用 151 次
- Language Agents with Reinforcement Learning for Strategic Play in the Werewolf GameZelai Xu, Chao Yu, Fei Fang, Yu Wang 等ICML 2024 · 被引用 145 次
- COLA: Consistent Learning with Opponent-Learning AwarenessTimon Willi, Alistair Letcher, Johannes Treutlein, Jakob N. FoersterICML 2022 · 被引用 61 次
- Model-Free Opponent ShapingChristopher Lu, Timon Willi, Christian A. Schröder de Witt, Jakob N. FoersterICML 2022 · 被引用 53 次
- A New Formalism, Method and Open Issues for Zero-Shot CoordinationJohannes Treutlein, Michael Dennis, Caspar Oesterheld, Jakob N. FoersterICML 2021 · 被引用 45 次
相关 Paper
- Spontaneous Giving and Calculated Greed in Language ModelsYuxuan Li, Hirokazu ShiradoEMNLP 2025 · 被引用 1 次
- Rethinking the Bounds of LLM Reasoning: Are Multi-Agent Discussions the Key?Qineng Wang, Zihao Wang, Ying Su, Hanghang Tong 等ACL 2024
- Talk, Judge, Cooperate: Gossip-Driven Indirect Reciprocity in Self-Interested LLM AgentsShuhui Zhu, Yue Lin, Shriya Kaistha, Wenhao Li 等ICML 2026
- More Capable, Less Cooperative? When LLMs Fail at Zero-Cost CollaborationAdvait Yadav, Sidney Black, Oliver SourbutICML 2026 · 被引用 2 次
- Self-Play Q-Learners Can Provably Collude in the Iterated Prisoner's DilemmaQuentin Bertrand, Juan Agustin Duque, Emilio Calvano, Gauthier GidelICML 2025
