Inducing Equilibria via Incentives: Simultaneous Design-and-Play Ensures Global Convergence
Boyi Liu, Jiayang Li, Zhuoran Yang, Hoi-To Wai, Mingyi Hong, Yu Marco Nie, Zhaoran Wang
摘要
To regulate a social system comprised of self-interested agents, economic incentives are often required to induce a desirable outcome. This incentive design problem naturally possesses a bilevel structure, in which a designer modifies the rewards of the agents with incentives while anticipating the response of the agents, who play a non-cooperative game that converges to an equilibrium. The existing bilevel optimization algorithms raise a dilemma when applied to this problem: anticipating how incentives affect the agents at equilibrium requires solving the equilibrium problem repeatedly, which is computationally inefficient; bypassing the time-consuming step of equilibrium-finding can reduce the computational cost, but may lead the designer to a sub-optimal solution. To address such a dilemma, we propose a method that tackles the designer's and agents' problems simultaneously in a single loop. Specifically, at each iteration, both the designer and the agents only move one step. Nevertheless, we allow the designer to gradually learn the overall influence of the incentives on the agents, which guarantees optimality after convergence. The convergence rate of the proposed scheme is also established for a broad class of games.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper7
- Contextual Bilevel Reinforcement Learning for Incentive AlignmentVinzenz Thoma, Barna Pásztor, Andreas Krause, Giorgia Ramponi 等NeurIPS 2024 · 被引用 21 次
- Achieving Hierarchy-Free Approximation for Bilevel Programs with Equilibrium ConstraintsJiayang Li, Jing Yu, Boyi Liu, Yu Marco Nie 等ICML 2023 · 被引用 10 次
- Learning Optimal Tax Design in Nonatomic Congestion GamesQiwen Cui, Maryam Fazel, Simon S. DuNeurIPS 2024 · 被引用 3 次
- Scalable Neural Incentive Design with Parameterized Mean-Field ApproximationNathan Corecco, Batuhan Yardim, Vinzenz Thoma, Zebang Shen 等NeurIPS 2025 · 被引用 1 次
- Deep Incentive Design with Differentiable Equilibrium BlocksVinzenz Thoma, Georgios Piliouras, Luke MarrisICML 2026
它引用的顶会 Paper1
相关 Paper
- Principled Penalty-based Methods for Bilevel Reinforcement Learning and RLHFHan Shen, Zhuoran Yang, Tianyi ChenICML 2024 · 被引用 35 次
- Adaptive Model Design for Markov Decision ProcessSiyu Chen, Donglin Yang, Jiayang Li, Senmiao Wang 等ICML 2022 · 被引用 14 次
- Welfare Maximization in Competitive Equilibrium: Reinforcement Learning for Markov Exchange EconomyZhihan Liu, Miao Lu, Zhaoran Wang, Michael I. Jordan 等ICML 2022 · 被引用 23 次
- Bilevel Optimization over Saddle Points of Zero-Sum Markov GamesZihao Zheng, Irwin King, Songtao LuICML 2026
- Faster Last-iterate Convergence of Policy Optimization in Zero-Sum Markov GamesShicong Cen, Yuejie Chi, Simon Shaolei Du, Lin XiaoICLR 2023 · 被引用 2 次
