Learning Multi-Agent Communication from Graph Modeling Perspective
Shengchao Hu, Li Shen, Ya Zhang, Dacheng Tao
Abstract
In numerous artificial intelligence applications, the collaborative efforts of multiple intelligent agents are imperative for the successful attainment of target objectives. To enhance coordination among these agents, a distributed communication framework is often employed. However, information sharing among all agents proves to be resource-intensive, while the adoption of a manually pre-defined communication architecture imposes limitations on inter-agent communication, thereby constraining the potential for collaborative efforts. In this study, we introduce a novel approach wherein we conceptualize the communication architecture among agents as a learnable graph. We formulate this problem as the task of determining the communication graph while enabling the architecture parameters to update normally, thus necessitating a bi-level optimization process. Utilizing continuous relaxation of the graph representation and incorporating attention units, our proposed approach, CommFormer, efficiently optimizes the communication graph and concurrently refines architectural parameters through gradient descent in an end-to-end manner. Extensive experiments on a variety of cooperative tasks substantiate the robustness of our model across diverse cooperative scenarios, where agents are able to develop more coordinated and sophisticated strategies regardless of changes in the number of agents.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 745f9317-99ec-4bfe-a1ae-0f7777798271Cited by top-tier papers25
- Locally Estimated Global Perturbations are Better than Local Perturbations for Federated Sharpness-aware MinimizationZiqing Fan, Shengchao Hu, Jiangchao Yao, Gang Niu et al.ICML 2024 · 35 citations
- Q-value Regularized Transformer for Offline Reinforcement LearningShengchao Hu, Ziqing Fan, Chaoqin Huang, Li Shen et al.ICML 2024 · 34 citations
- AgentDropout: Dynamic Agent Elimination for Token-Efficient and High-Performance LLM-Based Multi-Agent CollaborationZhexuan Wang, Yutong Wang, Xuebo Liu, Liang Ding et al.ACL 2025 · 32 citations
- HarmoDT: Harmony Multi-Task Decision Transformer for Offline Reinforcement LearningShengchao Hu, Ziqing Fan, Li Shen, Ya Zhang et al.ICML 2024 · 15 citations
- Decomposed Prompt Decision Transformer for Efficient Unseen Task GeneralizationHongling Zheng, Li Shen, Yong Luo, Tongliang Liu et al.NeurIPS 2024 · 13 citations
Builds on8
- Google Research Football: A Novel Reinforcement Learning EnvironmentKarol Kurach, Anton Raichuk, Piotr Stanczyk, Michal Zajac et al.AAAI 2020 · 496 citations
- Multi-Agent Reinforcement Learning is a Sequence Modeling ProblemMuning Wen, Jakub Grudzien Kuba, Runji Lin, Weinan Zhang et al.NeurIPS 2022 · 408 citations
- Trust Region Policy Optimisation in Multi-Agent Reinforcement LearningJakub Grudzien Kuba, Ruiqing Chen, Muning Wen, Ying Wen et al.ICLR 2022 · 367 citations
- Multi-Agent Game Abstraction via Graph Attention Neural NetworkYong Liu, Weixun Wang, Yujing Hu, Jianye Hao et al.AAAI 2020 · 316 citations
- Graph Transformer for Graph-to-Sequence LearningDeng Cai, Wai LamAAAI 2020 · 247 citations
Related papers
- Bridging Training and Execution via Dynamic Directed Graph-Based Communication in Cooperative Multi-Agent SystemsZhuohui Zhang, Bin He, Bin Cheng, Gang LiAAAI 2025 · 10 citations
- Deep Coordination GraphsWendelin Boehmer, Vitaly Kurin, Shimon WhitesonICML 2020 · 209 citations
- Reinforcement Learning with Fuzzy Human Attention-Guided Graph for Heterogeneous Multiagent SystemsDingbang Liu, Fenghui Ren, Jun Yan, Guoxin Su et al.AAAI 2026
- HYGMA: Hypergraph Coordination Networks with Dynamic Grouping for Multi-Agent Reinforcement LearningChiqiang Liu, Dazi LiICML 2025
- Language Model Networks: Supervision-Efficient Learning through Dense CommunicationShiguang Wu, Yaqing Wang, QUANMING YAOICML 2026 · 3 citations
