Fictitious Play and Best-Response Dynamics in Identical Interest and Zero-Sum Stochastic Games
Lucas Baudin, Rida Laraki
摘要
This paper proposes an extension of a popular decentralized discrete-time learning procedure when repeating a static game called fictitious play (FP) (Brown, 1951; Robinson, 1951) to a dynamic model called discounted stochastic game (Shapley, 1953) . Our family of discretetime FP procedures is proven to converge to the set of stationary Nash equilibria in identical interest discounted stochastic games. This extends similar convergence results for static games (Monderer & Shapley, 1996a). We then analyze the continuous-time counterpart of our FP procedures, which include as a particular case the best-response dynamic introduced and studied by Leslie et al. (2020) in the context of zero-sum stochastic games. We prove the converge of this dynamics to stationary Nash equilibria in identical-interest and zero-sum discounted stochastic games. Thanks to stochastic approximations, we can infer from the continuous-time convergence some discrete time results such as the convergence to stationary equilibria in zero sum and team stochastic games (Holler, 2020) .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- A Finite-Sample Analysis of Payoff-Based Independent Learning in Zero-Sum Stochastic GamesZaiwei Chen, Kaiqing Zhang, Eric Mazumdar, Asuman E. Ozdaglar 等NeurIPS 2023 · 被引用 22 次
- Multi-Player Zero-Sum Markov Games with Networked Separable InteractionsChanwoo Park, Kaiqing Zhang, Asuman E. OzdaglarNeurIPS 2023 · 被引用 17 次
- Smooth Fictitious Play in Stochastic Games with Perturbed Payoffs and Unknown TransitionsLucas Baudin, Rida LarakiNeurIPS 2022 · 被引用 8 次
- Optimism Without Regularization: Constant Regret in Zero-Sum GamesJohn Lazarsfeld, Georgios Piliouras, Ryann Sim, Stratis SkoulakisNeurIPS 2025 · 被引用 7 次
- Paths to Equilibrium in GamesBora Yongacoglu, Gürdal Arslan, Lacra Pavel, Serdar YükselNeurIPS 2024 · 被引用 2 次
它引用的顶会 Paper5
- Conservative Q-Learning for Offline Reinforcement LearningAviral Kumar, Aurick Zhou, George Tucker, Sergey LevineNeurIPS 2020 · 被引用 2,881 次
- Fictitious Play for Mean Field Games: Continuous Time Analysis and ApplicationsSarah Perrin, Julien Pérolat, Mathieu Laurière, Matthieu Geist 等NeurIPS 2020 · 被引用 150 次
- Decentralized Q-learning in Zero-sum Markov GamesMuhammed O. Sayin, Kaiqing Zhang, David S. Leslie, Tamer Basar 等NeurIPS 2021 · 被引用 105 次
- Can Q-Learning with Graph Networks Learn a Generalizable Branching Heuristic for a SAT Solver?Vitaly Kurin, Saad Godil, Shimon Whiteson, Bryan CatanzaroNeurIPS 2020 · 被引用 77 次
- The complexity of constrained min-max optimizationConstantinos Daskalakis, Stratis Skoulakis, Manolis ZampetakisSTOC 2021 · 被引用 18 次
相关 Paper
- Exponential Lower Bounds for Fictitious Play in Potential GamesIoannis Panageas, Nikolas Patris, Stratis Skoulakis, Volkan CevherNeurIPS 2023 · 被引用 1 次
- Team-Fictitious Play for Reaching Team-Nash Equilibrium in Multi-team GamesAhmed Said Donmez, Yuksel Arslantas, Muhammed Omer SayinNeurIPS 2024 · 被引用 2 次
- Fast Convergence of Fictitious Play for Diagonal Payoff MatricesJacob D. Abernethy, Kevin A. Lai, Andre WibisonoSODA 2021
- Towards convergence to Nash equilibria in two-team zero-sum gamesFivos Kalogiannis, Ioannis Panageas, Emmanouil V. Vlatakis-GkaragkounisICLR 2023
- Learning While Playing in Mean-Field Games: Convergence and OptimalityQiaomin Xie, Zhuoran Yang, Zhaoran Wang, Andreea MincaICML 2021 · 被引用 45 次
