Learning to Communicate Through Implicit Communication Channels
Han Wang, Binbin Chen, Tieying Zhang, Baoxiang Wang
摘要
Effective communication is an essential component in collaborative multi-agent systems. Situations where explicit messaging is not feasible have been common in human society throughout history, which motivate the study of implicit communication. Previous works on learning implicit communication mostly rely on theory of mind (ToM), where agents infer the mental states and intentions of others by interpreting their actions. However, ToM-based methods become less effective in making accurate inferences in complex tasks. In this work, we propose the Implicit Channel Protocol (ICP) framework, which allows agents to construct implicit communication channels similar to the explicit ones. ICP leverages a subset of actions, denoted as the scouting actions, and a mapping between information and these scouting actions that encodes and decodes the messages. We propose training algorithms for agents to message and act, including learning with a randomly initialized information map and with a delayed information map. The efficacy of ICP has been tested on the tasks of Guessing Number, Revealing Goals, and Hanabi, where ICP significantly outperforms baseline methods through more efficient information transmission.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper10
- Graph Convolutional Reinforcement LearningJiechuan Jiang, Chen Dun, Tiejun Huang, Zongqing LuICLR 2020 · 被引用 415 次
- Learning Efficient Multi-agent Communication: An Information Bottleneck ApproachRundong Wang, Xu He, Runsheng Yu, Wei Qiu 等ICML 2020 · 被引用 133 次
- Communication in Multi-Agent Reinforcement Learning: Intention SharingWoojun Kim, Jongeui Park, Youngchul SungICLR 2021 · 被引用 116 次
- ToM2C: Target-oriented Multi-agent Communication and Cooperation with Theory of MindYuanfei Wang, Fangwei Zhong, Jing Xu, Yizhou WangICLR 2022 · 被引用 103 次
- Simplified Action Decoder for Deep Multi-Agent Reinforcement LearningHengyuan Hu, Jakob N. FoersterICLR 2020 · 被引用 88 次
相关 Paper
- Few-shot Language Coordination by Modeling Theory of MindHao Zhu, Graham Neubig, Yonatan BiskICML 2021 · 被引用 43 次
- Improving Policies via Search in Cooperative Partially Observable GamesAdam Lerer, Hengyuan Hu, Jakob N. Foerster, Noam BrownAAAI 2020 · 被引用 87 次
- Limits of Theory of Mind Modelling in Dialogue-Based Collaborative Plan AcquisitionMatteo Bortoletto, Constantin Ruhdorfer, Adnen Abdessaied, Lei Shi 等ACL 2024
- CoDe: Communication Delay-Tolerant Multi-Agent Collaboration via Dual Alignment of Intent and TimelinessShoucheng Song, Youfang Lin, Sheng Han, Chang Yao 等AAAI 2025 · 被引用 7 次
- Learning Individually Inferred Communication for Multi-Agent CooperationZiluo Ding, Tiejun Huang, Zongqing LuNeurIPS 2020 · 被引用 146 次
