Learning Efficient and Interpretable Multi-Agent Communication
Wei Du, Benyu Wu, Yuqing Sun, Wei Guo, Yuntao Du, Zhongmin Yan, Guoxian Yu, Lizhen Cui
摘要
Effective communication is crucial for multi-agent cooperation in partially observable environments. However, a fundamental trilemma exists among task performance, communication efficiency, and human interpretability. To resolve this, we propose a multi-agent communication framework via rounding anguage and ontrastive learning (GLC) to learns efficient and interpretable communication protocols. Specifically, GLC employs an autoencoder to learn discretized and compressed communication symbols, ensuring high communication efficiency. These symbols are then semantically aligned with human concepts using data generated by a Large Language Model (LLM), making them human-interpretable. Furthermore, a contrastive learning objective is introduced to ensure consistency and mutual intelligibility among all agents, thereby securing high task utility. GLC dynamically balances these objectives by the Information Bottleneck principle. Extensive experiments show that GLC outperforms state-of-the-art methods across multiple benchmarks, delivering superior task performance, higher communication efficiency, and enhanced human interpretability.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper10
- Supervised Contrastive LearningPrannay Khosla, Piotr Teterwak, Chen Wang, Aaron Sarna 等NeurIPS 2020 · 被引用 7,049 次
- Learning to Ground Multi-Agent Communication with AutoencodersToru Lin, Jacob Huh, Christopher Stauffer, Ser-Nam Lim 等NeurIPS 2021 · 被引用 75 次
- Theory of Mind for Multi-Agent Collaboration via Large Language ModelsHuao Li, Yu Quan Chong, Simon Stepputtis, Joseph Campbell 等EMNLP 2023 · 被引用 57 次
- Trading off Utility, Informativeness, and Complexity in Emergent CommunicationMycal Tucker, Roger Levy, Julie A. Shah, Noga ZaslavskyNeurIPS 2022 · 被引用 34 次
- Language Grounded Multi-agent Reinforcement Learning with Human-interpretable CommunicationHuao Li, Hossein Nourkhiz Mahjoub, Behdad Chalaki, Vaishnav Tadiparthi 等NeurIPS 2024 · 被引用 31 次
相关 Paper
- Learning Multi-Agent Communication with Contrastive LearningYat Long Lo, Biswa Sengupta, Jakob Nicolaus Foerster, Michael NoukhovitchICLR 2024 · 被引用 11 次
- RGMComm: Return Gap Minimization via Discrete Communications in Multi-Agent Reinforcement LearningJingdi Chen, Tian Lan, Carlee Joe-WongAAAI 2024 · 被引用 18 次
- LLM-Guided Communication for Cooperative Multi-Agent Reinforcement LearningSangjun Bae, Yisak Park, Sanghyeon Lee, Seungyul HanICML 2026 · 被引用 2 次
- Augmenting Multi-Agent Communication with State Delta TrajectoryYichen Tang, Weihang Su, Yujia Zhou, Yiqun Liu 等EMNLP 2025
- GRACE: Generative Representation Learning via Contrastive Policy OptimizationJiashuo Sun, Shixuan Liu, Zhaochen Su, Xianrui Zhong 等ICLR 2026 · 被引用 7 次
