Settling Decentralized Multi-Agent Coordinated Exploration by Novelty Sharing
Haobin Jiang, Ziluo Ding, Zongqing Lu
Abstract
Exploration in decentralized cooperative multi-agent reinforcement learning faces two challenges. One is that the novelty of global states is unavailable, while the novelty of local observations is biased. The other is how agents can explore in a coordinated way. To address these challenges, we propose MACE, a simple yet effective multi-agent coordinated exploration method. By communicating only local novelty, agents can take into account other agents' local novelty to approximate the global novelty. Further, we newly introduce weighted mutual information to measure the influence of one agent's action on other agents' accumulated novelty. We convert it as an intrinsic reward in hindsight to encourage agents to exert more influence on other agents' exploration and boost coordinated exploration. Empirically, we show that MACE achieves superior performance in three multi-agent environments with sparse rewards. The code is available at https://github.com/SigmaBM/MACE .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 3e15c05d-69bc-498f-b43a-822b463287fbCited by top-tier papers4
- Wonder Wins Ways: Curiosity-Driven Exploration through Multi-Agent Contextual CalibrationYiyuan Pan, Zhe Liu, Hesheng WangNeurIPS 2025 · 10 citations
- Multi-Agent Coordination via Multi-Level CommunicationGang Ding, Zeyuan Liu, Zhirui Fang, Kefan Su et al.NeurIPS 2024 · 3 citations
- IEC: When Information-Driven Exploration Meets Spectral Consensus via Primal–Dual Reward Regularization in Decentralized Multi-Agent RLXuefeng Du, Jiajun Wu, Yuduo Zheng, Fengqi LiICML 2026
- Large Language Models Explore by Latent DistillingYuanhao Zeng, Ao Lu, Lufei Li, Zheng Zhang et al.ICML 2026
Builds on15
- Never Give Up: Learning Directed Exploration StrategiesAdrià Puigdomènech Badia, Pablo Sprechmann, Alex Vitvitskyi, Zhaohan Daniel Guo et al.ICLR 2020 · 349 citations
- ROMA: Multi-Agent Reinforcement Learning with Emergent RolesTonghan Wang, Heng Dong, Victor R. Lesser, Chongjie ZhangICML 2020 · 286 citations
- Celebrating Diversity in Shared Multi-Agent Reinforcement LearningChenghao Li, Tonghan Wang, Chengjie Wu, Qianchuan Zhao et al.NeurIPS 2021 · 224 citations
- RIDE: Rewarding Impact-Driven Exploration for Procedurally-Generated EnvironmentsRoberta Raileanu, Tim RocktäschelICLR 2020 · 198 citations
- Learning Nearly Decomposable Value Functions Via Communication MinimizationTonghan Wang, Jianhao Wang, Chongyi Zheng, Chongjie ZhangICLR 2020 · 170 citations
Related papers
- Two Heads are Better Than One: A Simple Exploration Framework for Efficient Multi-Agent Reinforcement LearningJiahui Li, Kun Kuang, Baoxiang Wang, Xingchen Li et al.NeurIPS 2023 · 7 citations
- Cooperative Exploration for Multi-Agent Deep Reinforcement LearningIou-Jen Liu, Unnat Jain, Raymond A. Yeh, Alexander G. SchwingICML 2021 · 133 citations
- Situation-Dependent Causal Influence-Based Cooperative Multi-Agent Reinforcement LearningXiao Du, Yutong Ye, Pengyu Zhang, Yaning Yang et al.AAAI 2024 · 19 citations
- Influence-Based Multi-Agent ExplorationTonghan Wang, Jianhao Wang, Yi Wu, Chongjie ZhangICLR 2020 · 156 citations
- Individual Contributions as Intrinsic Exploration Scaffolds for Multi-agent Reinforcement LearningXinran Li, Zifan Liu, Shibo Chen, Jun ZhangICML 2024 · 11 citations
