What Preferences Can—and Cannot—Predict in Multi-Agent Online Learning
Omar Abbadi, Rida Laraki, Panayotis Mertikopoulos
摘要
We examine the interplay between ordinal, preference-based solution concepts in games and the outcomes of payoff-driven learning dynamics, asking to what extent the combinatorial data of a game—its preference graph—can predict the long-run behavior of no-regret dynamics such as follow-the-regularized-leader (FTRL). In one direction, we show that the skeleton of every dynamically stable set, i.e., the set of pure profiles it contains, must be preferentially stable , that is, closed under pure profitable deviations. We then ask the converse question: when are preferences sufficient to describe long-run behavior? For subgames —subsets of pure profiles obtained by restricting players’ action sets—preferences are enough to fully characterize asymptotic stability. Beyond subgames however, we construct a three-player counterexample with a preferentially stable set whose span is dynamically unstable , thus establishing that preferences are not sufficient to describe dynamically stable behavior in general. To restore stability, we introduce the notion of leaklessness , a measure of aggregate payoff drift away from a set of pure profiles, and use it to identify a payoff-based condition under which the span of a set of pure profiles remains stable and attracting, thereby setting forth a natural cardinal guarantee of dynamic stability.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper12
- Dual Mirror Descent for Online Allocation ProblemsSantiago R. Balseiro, Haihao Lu, Vahab S. MirrokniICML 2020 · 被引用 102 次
- No-Regret Learning and Mixed Nash Equilibria: They Do Not MixEmmanouil V. Vlatakis-Gkaragkounis, Lampros Flokas, Thanasis Lianeas, Panayotis Mertikopoulos 等NeurIPS 2020 · 被引用 100 次
- The Limits of Min-Max Optimization Algorithms: Convergence to Spurious Non-Critical SetsYa-Ping Hsieh, Panayotis Mertikopoulos, Volkan CevherICML 2021 · 被引用 96 次
- From Chaos to Order: Symmetry and Conservation Laws in Game DynamicsSai Ganesh Nagarajan, David Balduzzi, Georgios PiliourasICML 2020 · 被引用 20 次
- The convergence rate of regularized learning in games: From bandits and uncertainty to optimism and beyondAngeliki Giannou, Emmanouil V. Vlatakis-Gkaragkounis, Panayotis MertikopoulosNeurIPS 2021 · 被引用 19 次
相关 Paper
- The impact of uncertainty on regularized learning in gamesPierre-Louis Cauvin, Davide Legacci, Panayotis MertikopoulosICML 2025
- The Equivalence of Dynamic and Strategic Stability under Regularized Learning in GamesVictor Boone, Panayotis MertikopoulosNeurIPS 2023 · 被引用 9 次
- Robust Equilibria in Continuous Games: From Strategic to Dynamic RobustnessKyriakos Lotidis, Panayotis Mertikopoulos, Nicholas Bambos, Jose H. BlanchetNeurIPS 2025 · 被引用 2 次
- No-regret Learning in Harmonic Games: Extrapolation in the Face of Conflicting InterestsDavide Legacci, Panayotis Mertikopoulos, Christos H. Papadimitriou, Georgios Piliouras 等NeurIPS 2024 · 被引用 10 次
- Online Optimization in Games via Control Theory: Connecting Regret, Passivity and Poincaré RecurrenceYun Kuen Cheung, Georgios PiliourasICML 2021 · 被引用 9 次
