Bridging Training and Execution via Dynamic Directed Graph-Based Communication in Cooperative Multi-Agent Systems
Zhuohui Zhang, Bin He, Bin Cheng, Gang Li
Abstract
Multi-agent systems must learn to communicate and understand interactions between agents to achieve cooperative goals in partially observed tasks. However, existing approaches lack a dynamic directed communication mechanism and rely on global states, thus diminishing the role of communication in centralized training. Thus, we propose the Transformer-based graph coarsening network (TGCNet), a novel multi-agent reinforcement learning (MARL) algorithm. TGCNet learns the topological structure of a dynamic directed graph to represent the communication policy and integrates graph coarsening networks to approximate the representation of global state during training. It also utilizes the Transformer decoder for feature extraction during execution. Experiments on multiple cooperative MARL benchmarks demonstrate state-of-the-art performance compared to popular MARL algorithms. Further ablation studies validate the effectiveness of our dynamic directed graph communication mechanism and graph coarsening networks.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 971b25ee-14c5-4604-bb30-5e5f260f0fb0Cited by top-tier papers4
- Communication-efficient Multi-Agent Reinforcement Learning with Spatiotemporal Information HubLing Ding, Tianbai Lyu, Zhiliang Bi, Hao Wang et al.AAAI 2026
- DLM: Unified Decision Language Models for Offline Multi-Agent Sequential Decision MakingZhuohui Zhang, Bin Cheng, Bin HeICML 2026
- GRDC: A Unified Graph-Driven Framework for Role Discovery and Communication in Multi-Agent Reinforcement LearningZihong Gao, Hongjian Liang, Yuanhui Hao, Lei Hao et al.AAAI 2026
- M2I2: Learning Efficient Multi-Agent Communication via Masked State Modeling and Intention InferenceChuxiong Sun, Peng He, Qirui Ji, Zehua Zang et al.AAAI 2026
Builds on7
- QPLEX: Duplex Dueling Multi-Agent Q-LearningJianhao Wang, Zhizhou Ren, Terry Liu, Yang Yu et al.ICLR 2021 · 595 citations
- Multi-Agent Game Abstraction via Graph Attention Neural NetworkYong Liu, Weixun Wang, Yujing Hu, Jianye Hao et al.AAAI 2020 · 316 citations
- Learning Nearly Decomposable Value Functions Via Communication MinimizationTonghan Wang, Jianhao Wang, Chongyi Zheng, Chongjie ZhangICLR 2020 · 170 citations
- Multi-Agent Incentive Communication via Decentralized Teammate ModelingLei Yuan, Jianhao Wang, Fuxiang Zhang, Chenghe Wang et al.AAAI 2022 · 104 citations
- Learning to Ground Multi-Agent Communication with AutoencodersToru Lin, Jacob Huh, Christopher Stauffer, Ser-Nam Lim et al.NeurIPS 2021 · 75 citations
Related papers
- Towards Efficient Collaboration via Graph Modeling in Reinforcement LearningWenzhe Fan, Zishun Yu, Chengdong Ma, Changye Li et al.AAAI 2025
- Multi-Agent Actor-Critic with Hierarchical Graph Attention NetworkHeechang Ryu, Hayong Shin, Jinkyoo ParkAAAI 2020 · 143 citations
- Offline Multi-Agent Reinforcement Learning with Knowledge DistillationWei-Cheng Tseng, Tsun-Hsuan Johnson Wang, Yen-Chen Lin, Phillip IsolaNeurIPS 2022 · 62 citations
- Context-Aware Bayesian Network Actor-Critic Methods for Cooperative Multi-Agent Reinforcement LearningDingyang Chen, Qi ZhangICML 2023 · 5 citations
- Learning Multi-Agent Communication through Structured Attentive ReasoningMurtaza Rangwala, Ryan WilliamsNeurIPS 2020 · 42 citations
