Transition-Informed Reinforcement Learning for Large-Scale Stackelberg Mean-Field Games
Pengdeng Li, Runsheng Yu, Xinrun Wang, Bo An
摘要
Many real-world scenarios including fleet management and Ad auctions can be modeled as Stackelberg mean-field games (SMFGs) where a leader aims to incentivize a large number of homogeneous self-interested followers to maximize her utility. Existing works focus on cases with a small number of heterogeneous followers, e.g., 5-10, and suffer from scalability issue when the number of followers increases. There are three major challenges in solving large-scale SMFGs: i) classical methods based on solving differential equations fail as they require exact dynamics parameters, ii) learning by interacting with environment is data-inefficient, and iii) complex interaction between the leader and followers makes the learning performance unstable. We address these challenges through transition-informed reinforcement learning. Our main contributions are threefold: i) we first propose an RL framework, the Stackelberg mean-field update, to learn the leader's policy without priors of the environment, ii) to improve the data efficiency and accelerate the learning process, we then propose the Transition-Informed Reinforcement Learning (TIRL) by leveraging the instantiated empirical Fokker-Planck equation, and iii) we develop a regularized TIRL by employing various regularizers to alleviate the sensitivity of the learning performance to the initialization of the leader's policy. Extensive experiments on fleet management and food gathering demonstrate that our approach can scale up to 100,000 followers and significantly outperform existing baselines.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper3
- Fictitious Play for Mean Field Games: Continuous Time Analysis and ApplicationsSarah Perrin, Julien Pérolat, Mathieu Laurière, Matthieu Geist 等NeurIPS 2020 · 被引用 150 次
- Multi-agent Trajectory Prediction with Fuzzy Query AttentionNitin Kamra, Hao Zhu, Dweep Trivedi, Ming Zhang 等NeurIPS 2020 · 被引用 40 次
- Evolutionary Population Curriculum for Scaling Multi-Agent Reinforcement LearningQian Long, Zihan Zhou, Abhinav Gupta, Fei Fang 等ICLR 2020
相关 Paper
- Learning in Stackelberg Mean Field Games: A Non-Asymptotic AnalysisSihan Zeng, Benjamin Patrick Evans, Sujay Bhatt, Leo Ardon 等NeurIPS 2025 · 被引用 1 次
- On Imitation in Mean-field GamesGiorgia Ramponi, Pavel Kolev, Olivier Pietquin, Niao He 等NeurIPS 2023 · 被引用 12 次
- Bayesian Multi-type Mean Field Multi-agent Imitation LearningFan Yang, Alina Vereshchaka, Changyou Chen, Wen DongNeurIPS 2020 · 被引用 21 次
- Meta-Inverse Reinforcement Learning for Mean Field Games via Probabilistic Context VariablesYang Chen, Xiao Lin, Bo Yan, Libo Zhang 等AAAI 2024 · 被引用 8 次
- Decentralized Mean Field GamesSriram Ganapathi Subramanian, Matthew E. Taylor, Mark Crowley, Pascal PoupartAAAI 2022 · 被引用 19 次
