HAVEN: Hierarchical Cooperative Multi-Agent Reinforcement Learning with Dual Coordination Mechanism
Zhiwei Xu, Yunpeng Bai, Bin Zhang, Dapeng Li, Guoliang Fan
Abstract
Recently, some challenging tasks in multi-agent systems have been solved by some hierarchical reinforcement learning methods. Inspired by the intra-level and inter-level coordination in the human nervous system, we propose a novel value decomposition framework HAVEN based on hierarchical reinforcement learning for fully cooperative multi-agent problems. To address the instability arising from the concurrent optimization of policies between various levels and agents, we introduce the dual coordination mechanism of inter-level and inter-agent strategies by designing reward functions in a two-level hierarchy. HAVEN does not require domain knowledge and pre-training, and can be applied to any value decomposition variant. Our method achieves desirable results on different decentralized partially observable Markov decision process domains and outperforms other popular multi-agent hierarchical reinforcement learning algorithms.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext df6b03c2-8473-4a0f-b1f1-19aefdf83c57Cited by top-tier papers8
- Asynchronous Actor-Critic for Multi-Agent Reinforcement LearningYuchen Xiao, Weihao Tan, Christopher AmatoNeurIPS 2022 · 35 citations
- Mingling Foresight with Imagination: Model-Based Cooperative Multi-Agent Reinforcement LearningZhiwei Xu, Dapeng Li, Bin Zhang, Yuan Zhan et al.NeurIPS 2022 · 14 citations
- Dual Self-Awareness Value Decomposition Framework without Individual Global Max for Cooperative MARLZhiwei Xu, Bin Zhang, Dapeng Li, Guangchong Zhou et al.NeurIPS 2023 · 12 citations
- Multi-Agent Reinforcement Learning with Hierarchical Coordination for Emergency Responder StationingAmutheezan Sivagnanam, Ava Pettet, Hunter Lee, Ayan Mukhopadhyay et al.ICML 2024 · 10 citations
- MPCache: MPC-Friendly KV Cache Eviction for Efficient Private LLM InferenceWenxuan Zeng, Ye Dong, Jinjin Zhou, Jin Tan et al.NeurIPS 2025 · 4 citations
Builds on6
- Weighted QMIX: Expanding Monotonic Value Function Factorisation for Deep Multi-Agent Reinforcement LearningTabish Rashid, Gregory Farquhar, Bei Peng, Shimon WhitesonNeurIPS 2020 · 1,960 citations
- Google Research Football: A Novel Reinforcement Learning EnvironmentKarol Kurach, Anton Raichuk, Piotr Stanczyk, Michal Zajac et al.AAAI 2020 · 496 citations
- Value-Decomposition Multi-Agent Actor-CriticsJianyu Su, Stephen C. Adams, Peter A. BelingAAAI 2021 · 140 citations
- Hierarchical Reinforcement Learning by Discovering Intrinsic OptionsJesse Zhang, Haonan Yu, Wei XuICLR 2021 · 97 citations
- RODE: Learning Roles to Decompose Multi-Agent TasksTonghan Wang, Tarun Gupta, Anuj Mahajan, Bei Peng et al.ICLR 2021 · 60 citations
Related papers
- Conditional Diffusion Model for Multi-Agent Dynamic Task DecompositionYanda Zhu, Yuanyang Zhu, Daoyi Dong, Caihua Chen et al.AAAI 2026
- Non-Linear Coordination GraphsYipeng Kang, Tonghan Wang, Qianlan Yang, Xiaoran Wu et al.NeurIPS 2022 · 14 citations
- Gradient-Protected Value Decomposition for Cooperative Multi-Agent Reinforcement LearningJie Hou, Haowen Dou, Lujuan Dang, Liangjun Chen et al.AAAI 2026
- DECOR: Learning to Decompose and Collaborate in Deep Search via Multi-Agent Reinforcement LearningRuiqing Chen, Zekun Zhang, Gongduo Zhang, Lihong Gu et al.ICML 2026
- Globally Optimal Hierarchical Reinforcement Learning for Linearly-Solvable Markov Decision ProcessesGuillermo Infante, Anders Jonsson, Vicenç GómezAAAI 2022 · 8 citations
