Information Design in Multi-Agent Reinforcement Learning
Yue Lin, Wenhao Li, Hongyuan Zha, Baoxiang Wang
摘要
Reinforcement learning (RL) is inspired by the way human infants and animals learn from the environment. The setting is somewhat idealized because, in actual tasks, other agents in the environment have their own goals and behave adaptively to the ego agent. To thrive in those environments, the agent needs to influence other agents so their actions become more helpful and less harmful. Research in computational economics distills two ways to influence others directly: by providing tangible goods (mechanism design) and by providing information (information design). This work investigates information design problems for a group of RL agents. The main challenges are two-fold. One is the information provided will immediately affect the transition of the agent trajectories, which introduces additional non-stationarity. The other is the information can be ignored, so the sender must provide information that the receiver is willing to respect. We formulate the Markov signaling game, and develop the notions of signaling gradient and the extended obedience constraints that address these challenges. Our algorithm is efficient on various mixed-motive tasks and provides further insights into computational economics. Our code is publicly available at https://github.com/YueLin301/InformationDesignMARL .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- Markov Persuasion Processes: Learning to Persuade From ScratchFrancesco Bacchiocchi, Francesco Emanuele Stradi, Matteo Castiglioni, Alberto Marchesi 等NeurIPS 2025 · 被引用 13 次
- LOPT: Learning Optimal Pigovian Tax in Sequential Social DilemmasYun Hua, Shang Gao, Wenhao Li, Haosheng Chen 等NeurIPS 2025 · 被引用 1 次
- Verbalized Bayesian PersuasionWenhao Li, Yue Lin, Yun Hua, Xiangfeng Wang 等ICML 2026
- Generalized Principal-Agent Problem with a Learning AgentTao Lin, Yiling ChenICLR 2025
它引用的顶会 Paper7
- Scalable Evaluation of Multi-Agent Reinforcement Learning with Melting PotJoel Z. Leibo, Edgar A. Duéñez-Guzmán, Alexander Vezhnevets, John P. Agapiou 等ICML 2021 · 被引用 134 次
- Learning to Incentivize Other Learning AgentsJiachen Yang, Ang Li, Mehrdad Farajtabar, Peter Sunehag 等NeurIPS 2020 · 被引用 105 次
- Optimal-er Auctions through AttentionDmitry Ivanov, Iskander Safiulin, Igor Filippov, Ksenia BalabaevaNeurIPS 2022 · 被引用 57 次
- Trading off Utility, Informativeness, and Complexity in Emergent CommunicationMycal Tucker, Roger Levy, Julie A. Shah, Noga ZaslavskyNeurIPS 2022 · 被引用 34 次
- Bayesian Persuasion in Sequential Decision-MakingJiarui Gan, Rupak Majumdar, Goran Radanovic, Adish SinglaAAAI 2022 · 被引用 30 次
相关 Paper
- Principled Penalty-based Methods for Bilevel Reinforcement Learning and RLHFHan Shen, Zhuoran Yang, Tianyi ChenICML 2024 · 被引用 35 次
- Learning to Steer Markovian Agents under Model UncertaintyJiawei Huang, Vinzenz Thoma, Zebang Shen, Heinrich H. Nax 等ICLR 2025
- A game-theoretic analysis of networked system control for common-pool resource management using multi-agent reinforcement learningArnu Pretorius, Scott Alexander Cameron, Elan Van Biljon, Tom Makkink 等NeurIPS 2020 · 被引用 11 次
- Inducing Equilibria via Incentives: Simultaneous Design-and-Play Ensures Global ConvergenceBoyi Liu, Jiayang Li, Zhuoran Yang, Hoi-To Wai 等NeurIPS 2022 · 被引用 29 次
- Information Shaping for Enhanced Goal Recognition of Partially-Informed AgentsSarah Keren, Haifeng Xu, Kofi Kwapong, David C. Parkes 等AAAI 2020 · 被引用 14 次
