Information Design in Multi-Agent Reinforcement Learning
Yue Lin, Wenhao Li, Hongyuan Zha, Baoxiang Wang
Abstract
Reinforcement learning (RL) is inspired by the way human infants and animals learn from the environment. The setting is somewhat idealized because, in actual tasks, other agents in the environment have their own goals and behave adaptively to the ego agent. To thrive in those environments, the agent needs to influence other agents so their actions become more helpful and less harmful. Research in computational economics distills two ways to influence others directly: by providing tangible goods (mechanism design) and by providing information (information design). This work investigates information design problems for a group of RL agents. The main challenges are two-fold. One is the information provided will immediately affect the transition of the agent trajectories, which introduces additional non-stationarity. The other is the information can be ignored, so the sender must provide information that the receiver is willing to respect. We formulate the Markov signaling game, and develop the notions of signaling gradient and the extended obedience constraints that address these challenges. Our algorithm is efficient on various mixed-motive tasks and provides further insights into computational economics. Our code is publicly available at https://github.com/YueLin301/InformationDesignMARL .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 8c1acf97-a9b6-4a7b-8623-39d5357f45caCited by top-tier papers4
- Markov Persuasion Processes: Learning to Persuade From ScratchFrancesco Bacchiocchi, Francesco Emanuele Stradi, Matteo Castiglioni, Alberto Marchesi et al.NeurIPS 2025 · 13 citations
- LOPT: Learning Optimal Pigovian Tax in Sequential Social DilemmasYun Hua, Shang Gao, Wenhao Li, Haosheng Chen et al.NeurIPS 2025 · 1 citation
- Verbalized Bayesian PersuasionWenhao Li, Yue Lin, Yun Hua, Xiangfeng Wang et al.ICML 2026
- Generalized Principal-Agent Problem with a Learning AgentTao Lin, Yiling ChenICLR 2025
Builds on7
- Scalable Evaluation of Multi-Agent Reinforcement Learning with Melting PotJoel Z. Leibo, Edgar A. Duéñez-Guzmán, Alexander Vezhnevets, John P. Agapiou et al.ICML 2021 · 134 citations
- Learning to Incentivize Other Learning AgentsJiachen Yang, Ang Li, Mehrdad Farajtabar, Peter Sunehag et al.NeurIPS 2020 · 105 citations
- Optimal-er Auctions through AttentionDmitry Ivanov, Iskander Safiulin, Igor Filippov, Ksenia BalabaevaNeurIPS 2022 · 57 citations
- Trading off Utility, Informativeness, and Complexity in Emergent CommunicationMycal Tucker, Roger Levy, Julie A. Shah, Noga ZaslavskyNeurIPS 2022 · 34 citations
- Bayesian Persuasion in Sequential Decision-MakingJiarui Gan, Rupak Majumdar, Goran Radanovic, Adish SinglaAAAI 2022 · 30 citations
Related papers
- Principled Penalty-based Methods for Bilevel Reinforcement Learning and RLHFHan Shen, Zhuoran Yang, Tianyi ChenICML 2024 · 35 citations
- Learning to Steer Markovian Agents under Model UncertaintyJiawei Huang, Vinzenz Thoma, Zebang Shen, Heinrich H. Nax et al.ICLR 2025
- A game-theoretic analysis of networked system control for common-pool resource management using multi-agent reinforcement learningArnu Pretorius, Scott Alexander Cameron, Elan Van Biljon, Tom Makkink et al.NeurIPS 2020 · 11 citations
- Inducing Equilibria via Incentives: Simultaneous Design-and-Play Ensures Global ConvergenceBoyi Liu, Jiayang Li, Zhuoran Yang, Hoi-To Wai et al.NeurIPS 2022 · 29 citations
- Information Shaping for Enhanced Goal Recognition of Partially-Informed AgentsSarah Keren, Haifeng Xu, Kofi Kwapong, David C. Parkes et al.AAAI 2020 · 14 citations
