What Preferences Can—and Cannot—Predict in Multi-Agent Online Learning
Omar Abbadi, Rida Laraki, Panayotis Mertikopoulos
Abstract
We examine the interplay between ordinal, preference-based solution concepts in games and the outcomes of payoff-driven learning dynamics, asking to what extent the combinatorial data of a game—its preference graph—can predict the long-run behavior of no-regret dynamics such as follow-the-regularized-leader (FTRL). In one direction, we show that the skeleton of every dynamically stable set, i.e., the set of pure profiles it contains, must be preferentially stable , that is, closed under pure profitable deviations. We then ask the converse question: when are preferences sufficient to describe long-run behavior? For subgames —subsets of pure profiles obtained by restricting players’ action sets—preferences are enough to fully characterize asymptotic stability. Beyond subgames however, we construct a three-player counterexample with a preferentially stable set whose span is dynamically unstable , thus establishing that preferences are not sufficient to describe dynamically stable behavior in general. To restore stability, we introduce the notion of leaklessness , a measure of aggregate payoff drift away from a set of pure profiles, and use it to identify a payoff-based condition under which the span of a set of pure profiles remains stable and attracting, thereby setting forth a natural cardinal guarantee of dynamic stability.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext ea92233e-3b97-4101-accf-ceb83f6afab7Builds on12
- Dual Mirror Descent for Online Allocation ProblemsSantiago R. Balseiro, Haihao Lu, Vahab S. MirrokniICML 2020 · 102 citations
- No-Regret Learning and Mixed Nash Equilibria: They Do Not MixEmmanouil V. Vlatakis-Gkaragkounis, Lampros Flokas, Thanasis Lianeas, Panayotis Mertikopoulos et al.NeurIPS 2020 · 100 citations
- The Limits of Min-Max Optimization Algorithms: Convergence to Spurious Non-Critical SetsYa-Ping Hsieh, Panayotis Mertikopoulos, Volkan CevherICML 2021 · 96 citations
- From Chaos to Order: Symmetry and Conservation Laws in Game DynamicsSai Ganesh Nagarajan, David Balduzzi, Georgios PiliourasICML 2020 · 20 citations
- The convergence rate of regularized learning in games: From bandits and uncertainty to optimism and beyondAngeliki Giannou, Emmanouil V. Vlatakis-Gkaragkounis, Panayotis MertikopoulosNeurIPS 2021 · 19 citations
Related papers
- The impact of uncertainty on regularized learning in gamesPierre-Louis Cauvin, Davide Legacci, Panayotis MertikopoulosICML 2025
- The Equivalence of Dynamic and Strategic Stability under Regularized Learning in GamesVictor Boone, Panayotis MertikopoulosNeurIPS 2023 · 9 citations
- Robust Equilibria in Continuous Games: From Strategic to Dynamic RobustnessKyriakos Lotidis, Panayotis Mertikopoulos, Nicholas Bambos, Jose H. BlanchetNeurIPS 2025 · 2 citations
- No-regret Learning in Harmonic Games: Extrapolation in the Face of Conflicting InterestsDavide Legacci, Panayotis Mertikopoulos, Christos H. Papadimitriou, Georgios Piliouras et al.NeurIPS 2024 · 10 citations
- Online Optimization in Games via Control Theory: Connecting Regret, Passivity and Poincaré RecurrenceYun Kuen Cheung, Georgios PiliourasICML 2021 · 9 citations
