DACOM: Learning Delay-Aware Communication for Multi-Agent Reinforcement Learning
Tingting Yuan, Hwei-Ming Chung, Jie Yuan, Xiaoming Fu
摘要
Communication is supposed to improve multi-agent collaboration and overall performance in cooperative Multi-agent reinforcement learning (MARL). However, such improvements are prevalently limited in practice since most existing communication schemes ignore communication overheads (e.g., communication delays). In this paper, we demonstrate that ignoring communication delays has detrimental effects on collaborations, especially in delay-sensitive tasks such as autonomous driving. To mitigate this impact, we design a delay-aware multi-agent communication model (DACOM) to adapt communication to delays. Specifically, DACOM introduces a component, TimeNet, that is responsible for adjusting the waiting time of an agent to receive messages from other agents such that the uncertainty associated with delay can be addressed. Our experiments reveal that DACOM has a non-negligible performance improvement over other mechanisms by making a better trade-off between the benefits of communication and the costs of waiting for messages.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper9
- AccDecoder: Accelerated Decoding for Neural-enhanced Video AnalyticsTingting Yuan, Liang Mi, Weijun Wang, Haipeng Dai 等INFOCOM 2023 · 被引用 25 次
- Multi-Agent Reinforcement Learning with Communication-Constrained PriorsGuang Yang, Tianpei Yang, Jingwen Qiao, Yanqing Wu 等NeurIPS 2025 · 被引用 9 次
- CoDe: Communication Delay-Tolerant Multi-Agent Collaboration via Dual Alignment of Intent and TimelinessShoucheng Song, Youfang Lin, Sheng Han, Chang Yao 等AAAI 2025 · 被引用 7 次
- Dynamic Generation of Multi LLM Agents Communication Topologies with Graph Diffusion ModelsEric Hanchen Jiang, Levina Li, Frank Wan, Xiao Liang 等ACL 2026 · 被引用 6 次
- Rainbow Delay Compensation: A Multi-Agent Reinforcement Learning Framework for Mitigating Observation DelaysSongchen Fu, Siang Chen, Shaojing Zhao, Letian Bai 等NeurIPS 2025 · 被引用 2 次
它引用的顶会 Paper4
- Swift: Delay is Simple and Effective for Congestion Control in the DatacenterGautam Kumar, Nandita Dukkipati, Keon Jang, Hassan M. G. Wassel 等SIGCOMM 2020 · 被引用 333 次
- Learning Agent Communication under Limited Bandwidth by Message PruningHangyu Mao, Zhengchao Zhang, Zhen Xiao, Zhibo Gong 等AAAI 2020 · 被引用 110 次
- ACC: automatic ECN tuning for high-speed datacenter networksSiyu Yan, Xiaoliang Wang, Xiaolong Zheng, Yinben Xia 等SIGCOMM 2021 · 被引用 95 次
- Succinct and Robust Multi-Agent Communication With Temporal Message ControlSai Qian Zhang, Qi Zhang, Jieyu LinNeurIPS 2020 · 被引用 90 次
相关 Paper
- VIL2C: Value-of-Information Aware Low-Latency Communication for Multi-Agent Reinforcement LearningQian Zhang, Zhuo Sun, Yao Zhang, Zhiwen Yu 等AAAI 2026
- Bridging Training and Execution via Dynamic Directed Graph-Based Communication in Cooperative Multi-Agent SystemsZhuohui Zhang, Bin He, Bin Cheng, Gang LiAAAI 2025 · 被引用 10 次
- Multi-agent Reinforcement Learning for Networked System ControlTianshu Chu, Sandeep Chinchali, Sachin KattiICLR 2020 · 被引用 134 次
- Learning Multi-Agent Communication through Structured Attentive ReasoningMurtaza Rangwala, Ryan WilliamsNeurIPS 2020 · 被引用 42 次
- Correcting experience replay for multi-agent communicationSanjeevan Ahilan, Peter DayanICLR 2021 · 被引用 3 次
