Learning Efficient Multi-agent Communication: An Information Bottleneck Approach
Rundong Wang, Xu He, Runsheng Yu, Wei Qiu, Bo An, Zinovi Rabinovich
摘要
We consider the problem of the limited-bandwidth communication for multi-agent reinforcement learning, where agents cooperate with the assistance of a communication protocol and a scheduler. The protocol and scheduler jointly determine which agent is communicating what message and to whom. Under the limited bandwidth constraint, a communication protocol is required to generate informative messages. Meanwhile, an unnecessary communication connection should not be established because it occupies limited resources in vain. In this paper, we develop an Informative Multi-Agent Communication (IMAC) method to learn efficient communication protocols as well as scheduling. First, from the perspective of communication theory, we prove that the limited bandwidth constraint requires low-entropy messages throughout the transmission. Then inspired by the information bottleneck principle, we learn a valuable and compact communication protocol and a weight-based scheduler. To demonstrate the efficiency of our method, we conduct extensive experiments in various cooperative and competitive multi-agent tasks with different numbers of agents and different bandwidths. We show that IMAC converges faster and leads to efficient communication among agents under the limited bandwidth as compared to many baseline methods.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper23
- Graph Information Bottleneck for Subgraph RecognitionJunchi Yu, Tingyang Xu, Yu Rong, Yatao Bian 等ICLR 2021 · 被引用 200 次
- Improving Subgraph Recognition with Variational Graph Information BottleneckJunchi Yu, Jie Cao, Ran HeCVPR 2022 · 被引用 56 次
- Robust Predictable ControlBen Eysenbach, Ruslan Salakhutdinov, Sergey LevineNeurIPS 2021 · 被引用 53 次
- Trading off Utility, Informativeness, and Complexity in Emergent CommunicationMycal Tucker, Roger Levy, Julie A. Shah, Noga ZaslavskyNeurIPS 2022 · 被引用 34 次
- T2MAC: Targeted and Trusted Multi-Agent Communication through Selective Engagement and Evidence-Driven IntegrationChuxiong Sun, Zehua Zang, Jiabao Li, Jiangmeng Li 等AAAI 2024 · 被引用 24 次
相关 Paper
- Learning Agent Communication under Limited Bandwidth by Message PruningHangyu Mao, Zhengchao Zhang, Zhen Xiao, Zhibo Gong 等AAAI 2020 · 被引用 110 次
- Cheap Talk Discovery and Utilization in Multi-Agent Reinforcement LearningYat Long Lo, Christian Schröder de Witt, Samuel Sokota, Jakob Nicolaus Foerster 等ICLR 2023
- Communication Learning via Backpropagation in Discrete Channels with Unknown NoiseBenjamin Freed, Guillaume Sartoretti, Jiaheng Hu, Howie ChosetAAAI 2020 · 被引用 22 次
- Learning Efficient and Robust Multi-Agent Communication via Graph Information BottleneckShifei Ding, Wei Du, Ling Ding, Lili Guo 等AAAI 2024 · 被引用 12 次
- The Variational Bandwidth Bottleneck: Stochastic Evaluation on an Information BudgetAnirudh Goyal, Yoshua Bengio, Matthew M. Botvinick, Sergey LevineICLR 2020 · 被引用 26 次
