Breaking the Curse of Many Agents: Provable Mean Embedding Q-Iteration for Mean-Field Reinforcement Learning
Lingxiao Wang, Zhuoran Yang, Zhaoran Wang
摘要
Multi-agent reinforcement learning (MARL) achieves significant empirical successes. However, MARL suffers from the curse of many agents. In this paper, we exploit the symmetry of agents in MARL. In the most generic form, we study a mean-field MARL problem. Such a mean-field MARL is defined on mean-field states, which are distributions that are supported on continuous space. Based on the mean embedding of the distributions, we propose MF-FQI algorithm that solves the mean-field MARL and establishes a non-asymptotic analysis for MF-FQI algorithm. We highlight that MF-FQI algorithm enjoys a "blessing of many agents" property in the sense that a larger number of observed agents improves the performance of MF-FQI algorithm.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper11
- Robust Multi-Agent Reinforcement Learning via Adversarial Regularization: Theoretical Foundation and Stable AlgorithmsAlexander Bukharin, Yan Li, Yue Yu, Qingru Zhang 等NeurIPS 2023 · 被引用 55 次
- Coordination Between Individual Agents in Multi-Agent Reinforcement LearningYang Zhang, Qingyu Yang, Dou An, Chengwei ZhangAAAI 2021 · 被引用 21 次
- Pessimism Meets Invariance: Provably Efficient Offline Mean-Field Multi-Agent RLMinshuo Chen, Yan Li, Ethan Wang, Zhuoran Yang 等NeurIPS 2021 · 被引用 18 次
- Relational Reasoning via Set Transformers: Provable Efficiency and Applications to MARLFengzhuo Zhang, Boyi Liu, Kaixin Wang, Vincent Y. F. Tan 等NeurIPS 2022 · 被引用 16 次
- Learning Regularized Monotone Graphon Mean-Field GamesFengzhuo Zhang, Vincent Y. F. Tan, Zhaoran Wang, Zhuoran YangNeurIPS 2023 · 被引用 14 次
相关 Paper
- Major-Minor Mean Field Multi-Agent Reinforcement LearningKai Cui, Christian Fabian, Anam Tahir, Heinz KoepplICML 2024 · 被引用 6 次
- Mean-Field Sampling for Cooperative Multi-Agent Reinforcement LearningEmile Anand, Ishani Karmarkar, Guannan QuNeurIPS 2025 · 被引用 10 次
- Decentralized Mean Field GamesSriram Ganapathi Subramanian, Matthew E. Taylor, Mark Crowley, Pascal PoupartAAAI 2022 · 被引用 19 次
- Generalization in Mean Field Games by Learning Master PoliciesSarah Perrin, Mathieu Laurière, Julien Pérolat, Romuald Élie 等AAAI 2022 · 被引用 47 次
- On the Convergence of Model Free Learning in Mean Field GamesRomuald Elie, Julien Pérolat, Mathieu Laurière, Matthieu Geist 等AAAI 2020 · 被引用 101 次
