Learning Efficient Multi-agent Communication: An Information Bottleneck Approach
Rundong Wang, Xu He, Runsheng Yu, Wei Qiu, Bo An, Zinovi Rabinovich
Abstract
We consider the problem of the limited-bandwidth communication for multi-agent reinforcement learning, where agents cooperate with the assistance of a communication protocol and a scheduler. The protocol and scheduler jointly determine which agent is communicating what message and to whom. Under the limited bandwidth constraint, a communication protocol is required to generate informative messages. Meanwhile, an unnecessary communication connection should not be established because it occupies limited resources in vain. In this paper, we develop an Informative Multi-Agent Communication (IMAC) method to learn efficient communication protocols as well as scheduling. First, from the perspective of communication theory, we prove that the limited bandwidth constraint requires low-entropy messages throughout the transmission. Then inspired by the information bottleneck principle, we learn a valuable and compact communication protocol and a weight-based scheduler. To demonstrate the efficiency of our method, we conduct extensive experiments in various cooperative and competitive multi-agent tasks with different numbers of agents and different bandwidths. We show that IMAC converges faster and leads to efficient communication among agents under the limited bandwidth as compared to many baseline methods.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 238b6828-b703-4886-935c-d383e623c189Cited by top-tier papers23
- Graph Information Bottleneck for Subgraph RecognitionJunchi Yu, Tingyang Xu, Yu Rong, Yatao Bian et al.ICLR 2021 · 200 citations
- Improving Subgraph Recognition with Variational Graph Information BottleneckJunchi Yu, Jie Cao, Ran HeCVPR 2022 · 56 citations
- Robust Predictable ControlBen Eysenbach, Ruslan Salakhutdinov, Sergey LevineNeurIPS 2021 · 53 citations
- Trading off Utility, Informativeness, and Complexity in Emergent CommunicationMycal Tucker, Roger Levy, Julie A. Shah, Noga ZaslavskyNeurIPS 2022 · 34 citations
- T2MAC: Targeted and Trusted Multi-Agent Communication through Selective Engagement and Evidence-Driven IntegrationChuxiong Sun, Zehua Zang, Jiabao Li, Jiangmeng Li et al.AAAI 2024 · 24 citations
Related papers
- Learning Agent Communication under Limited Bandwidth by Message PruningHangyu Mao, Zhengchao Zhang, Zhen Xiao, Zhibo Gong et al.AAAI 2020 · 110 citations
- Cheap Talk Discovery and Utilization in Multi-Agent Reinforcement LearningYat Long Lo, Christian Schröder de Witt, Samuel Sokota, Jakob Nicolaus Foerster et al.ICLR 2023
- Communication Learning via Backpropagation in Discrete Channels with Unknown NoiseBenjamin Freed, Guillaume Sartoretti, Jiaheng Hu, Howie ChosetAAAI 2020 · 22 citations
- Learning Efficient and Robust Multi-Agent Communication via Graph Information BottleneckShifei Ding, Wei Du, Ling Ding, Lili Guo et al.AAAI 2024 · 12 citations
- The Variational Bandwidth Bottleneck: Stochastic Evaluation on an Information BudgetAnirudh Goyal, Yoshua Bengio, Matthew M. Botvinick, Sergey LevineICLR 2020 · 26 citations
