Nice Invincible Strategy for the Average-Payoff IPD
Shiheng Wang, Fangzhen Lin
摘要
The Iterated Prisoner's Dilemma (IPD) is a well-known benchmark for studying the long term behaviours of rational agents. Many well-known strategies have been studied, from the simple tit-for-tat (TFT) to more involved ones like zero determinant and extortionate strategies studied recently by Press and Dyson. In this paper, we consider what we call invincible strategies. These are ones that will never lose against any other strategy in terms of average payoff in the limit. We provide a simple characterization of this class of strategies, and show that invincible strategies can also be nice. We discuss its relationship with some important strategies and generalize our results to some typical repeated 2x2 games. It's known that experimentally, nice strategies like the TFT and extortionate ones can act as catalysts for the evolution of cooperation. Our experiments show that this is also the case for some invincible strategies that are neither nice nor extortionate.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它相关 Paper
- Exponential Lower Bounds for Fictitious Play in Potential GamesIoannis Panageas, Nikolas Patris, Stratis Skoulakis, Volkan CevherNeurIPS 2023 · 被引用 1 次
- Self-Play Q-Learners Can Provably Collude in the Iterated Prisoner's DilemmaQuentin Bertrand, Juan Agustin Duque, Emilio Calvano, Gauthier GidelICML 2025
- Fast Convergence of Fictitious Play for Diagonal Payoff MatricesJacob D. Abernethy, Kevin A. Lai, Andre WibisonoSODA 2021
- Connecting Optimal Ex-Ante Collusion in Teams to Extensive-Form Correlation: Faster Algorithms and Positive Complexity ResultsGabriele Farina, Andrea Celli, Nicola Gatti, Tuomas SandholmICML 2021 · 被引用 29 次
- On the Impossibility of Learning to Cooperate with Adaptive Partner Strategies in Repeated GamesRobert Tyler Loftin, Frans A. OliehoekICML 2022 · 被引用 4 次
