Inference-Based Deterministic Messaging For Multi-Agent Communication
Varun Bhatt, Michael Buro
Abstract
Communication is essential for coordination among humans and animals. Therefore, with the introduction of intelligent agents into the world, agent-to-agent and agent-to-human communication becomes necessary. In this paper, we first study learning in matrix-based signaling games to empirically show that decentralized methods can converge to a suboptimal policy. We then propose a modification to the messaging policy, in which the sender deterministically chooses the best message that helps the receiver to infer the sender's observation. Using this modification, we see, empirically, that the agents converge to the optimal policy in nearly all the runs. We then apply this method to a partially observable gridworld environment which requires cooperation between two agents and show that, with appropriate approximation methods, the proposed sender modification can enhance existing decentralized training methods for more complex domains as well.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 4e4580d6-bffc-4330-90b5-d5bf90048ebfCited by top-tier papers1
Ask how each one uses itRelated papers
- Learning to Communicate Implicitly by ActionsZheng Tian, Shihao Zou, Ian Davies, Tim Warr et al.AAAI 2020 · 35 citations
- Correcting experience replay for multi-agent communicationSanjeevan Ahilan, Peter DayanICLR 2021 · 3 citations
- Learning Multi-Agent Communication with Contrastive LearningYat Long Lo, Biswa Sengupta, Jakob Nicolaus Foerster, Michael NoukhovitchICLR 2024 · 11 citations
- Communication-Efficient Actor-Critic Methods for Homogeneous Markov GamesDingyang Chen, Yile Li, Qi ZhangICLR 2022 · 11 citations
- Consensus Learning for Cooperative Multi-Agent Reinforcement LearningZhiwei Xu, Bin Zhang, Dapeng Li, Zeren Zhang et al.AAAI 2023 · 27 citations
