Zero-Sum Stochastic Stackelberg Games
Denizalp Goktas, Sadie Zhao, Amy Greenwald
Abstract
Zero-sum stochastic games have found important applications in a variety of fields, from machine learning to economics. Work on this model has primarily focused on the computation of Nash equilibrium due to its effectiveness in solving adversarial board and video games. Unfortunately, a Nash equilibrium is not guaranteed to exist in zero-sum stochastic games when the payoffs at each state are not convexconcave in the players' actions. A Stackelberg equilibrium, however, is guaranteed to exist. Consequently, in this paper, we study zero-sum stochastic Stackelberg games. Going beyond known existence results for (non-stationary) Stackelberg equilibria, we prove the existence of recursive (i.e., Markov perfect) Stackelberg equilibria (recSE) in these games, provide necessary and sufficient conditions for a policy profile to be a recSE, and show that recSE can be computed in (weakly) polynomial time via value iteration. Finally, we show that zero-sum stochastic Stackelberg games can model the problem of pricing and allocating goods across agents and time. More specifically, we propose a zero-sum stochastic Stackelberg game whose recSE correspond to the recursive competitive equilibria of a large class of stochastic Fisher markets. We close with a series of experiments that showcase how our methodology can be used to solve the consumption-savings problem in stochastic Fisher markets. Min-max optimization has paved the way for recent progress in a variety of fields, from machine learning to economics. In a constrained min-max optimization problem, min x∈X max y∈Y f (x, y), the objective function f : X ×Y → R is continuous, and the constraint sets X ⊂ R n and Y ⊂ R m are nonempty and compact. When f is convex-concave, and the constraint sets X and Y are convex, the seminal minimax theorem [1, 2] holds, i.e., min x∈X max y∈Y f (x, y) = max y∈Y min x∈X f (x, y), and such problems can be interpreted as solving a min-max (or zero-sum) one-shot simultaneousmove game between an outer player x and an inner player y with respective payoff functions -f , f and respective action sets X , Y, where the solutions (x * , y * ) ∈ X × Y are Nash equilibria: i.e., best responses to one another with x * ∈ arg min x∈X f (x, y * ) and y * ∈ arg max y∈Y f (x * , y). More generally, one can consider zero-sum stochastic games played over an infinite discrete time horizon N + . The game starts at some initial state S (0) ∼ µ (0) . At each subsequent time-step t ∈ N + , players encounter a new state s (t) ∈ S. After taking their respective actions (x (t) , y (t) ) from their respective action spaces X (s (t) ) ⊆ R n and Y(s (t) ) ⊆ R m , they receive payoffs r(s (t) , x (t) , y (t) ), and then either transition to a new state S (t+1) ∼ p(• | s (t) , x (t) , y (t) ) with probability γ, or the game ends with the remaining probability. The goal of the outer (resp. inner) player is 36th Conference on Neural Information Processing Systems (NeurIPS 2022).
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 9fb66536-bd86-4fbe-a606-abb3b892364cCited by top-tier papers1
Ask how each one uses itBuilds on4
- On Gradient Descent Ascent for Nonconvex-Concave Minimax ProblemsTianyi Lin, Chi Jin, Michael I. JordanICML 2020 · 587 citations
- What is Local Optimality in Nonconvex-Nonconcave Minimax Optimization?Chi Jin, Praneeth Netrapalli, Michael I. JordanICML 2020 · 381 citations
- Convex-Concave Min-Max Stackelberg GamesDenizalp Goktas, Amy GreenwaldNeurIPS 2021 · 41 citations
- Online Market Equilibrium with Application to Fair DivisionYuan Gao, Alex Peysakhovich, Christian KroerNeurIPS 2021 · 35 citations
Related papers
- Convex-Concave Zero-Sum Stochastic Stackelberg GamesDenizalp Goktas, Arjun Prakash, Amy GreenwaldNeurIPS 2023
- Implicit Learning Dynamics in Stackelberg Games: Equilibria Characterization, Convergence Analysis, and Empirical StudyTanner Fiez, Benjamin Chasnov, Lillian J. RatliffICML 2020 · 144 citations
- Fisher Markets with Social InfluenceJiayi Zhao, Denizalp Goktas, Amy GreenwaldAAAI 2023 · 1 citation
- Can We Find Nash Equilibria at a Linear Rate in Markov Games?Zhuoqing Song, Jason D. Lee, Zhuoran YangICLR 2023
- Infinite Horizon Markov EconomiesDenizalp Goktas, Sadie Zhao, Yiling Chen, Amy GreenwaldICLR 2026 · 1 citation
