The Dormant Neuron Phenomenon in Multi-Agent Reinforcement Learning Value Factorization
Haoyuan Qin, Chennan Ma, Mian Deng, Zhengzhu Liu, Songzhu Mei, Xinwang Liu, Cheng Wang, Siqi Shen
摘要
In this work, we study the dormant neuron phenomenon in multi-agent reinforcement learning value factorization, where the mixing network suffers from reduced network expressivity caused by an increasing number of inactive neurons. We demonstrate the presence of the dormant neuron phenomenon across multiple environments and algorithms, and show that this phenomenon negatively affects the learning process. We show that dormant neurons correlates with the existence of over-active neurons, which have large activation scores. To address the dormant neuron issue, we propose ReBorn, a simple but effective method that transfers the weights from over-active neurons to dormant neurons. We theoretically show that this method can ensure the learned action preferences are not forgotten after the weight-transferring procedure, which increases learning effectiveness. Our extensive experiments reveal that ReBorn achieves promising results across various environments and improves the performance of multiple popular value factorization approaches. The source code of ReBorn is available in https://github.com/xmu-rl-3dv/ReBorn.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper7
- Measure gradients, not activations! Enhancing neuronal activity in deep reinforcement learningJiashun Liu, Zihao Wu, Johan S. Obando-Ceron, Pablo Samuel Castro 等NeurIPS 2025 · 被引用 15 次
- Retaining Suboptimal Actions to Follow Shifting Optima in Multi-Agent Reinforcement LearningYonghyeon Jo, Sunwoo Lee, Seungyul HanICLR 2026 · 被引用 5 次
- High-order Interactions Modeling for Interpretable Multi-Agent Q-LearningQinyu Xu, Yuanyang Zhu, Xuefei Wu, Chunlin ChenNeurIPS 2025 · 被引用 2 次
- Neuroplastic Expansion in Deep Reinforcement LearningJiashun Liu, Johan S. Obando-Ceron, Aaron C. Courville, Ling PanICLR 2025
- GradPS: Resolving Futile Neurons in Parameter Sharing Network for Multi-Agent Reinforcement LearningHaoyuan Qin, Zhengzhu Liu, Chenxing Lin, Chennan Ma 等ICML 2025
它引用的顶会 Paper18
- Weighted QMIX: Expanding Monotonic Value Function Factorisation for Deep Multi-Agent Reinforcement LearningTabish Rashid, Gregory Farquhar, Bei Peng, Shimon WhitesonNeurIPS 2020 · 被引用 1,960 次
- QPLEX: Duplex Dueling Multi-Agent Q-LearningJianhao Wang, Zhizhou Ren, Terry Liu, Yang Yu 等ICLR 2021 · 被引用 595 次
- The Primacy Bias in Deep Reinforcement LearningEvgenii Nikishin, Max Schwarzer, Pierluca D'Oro, Pierre-Luc Bacon 等ICML 2022 · 被引用 269 次
- Deep Coordination GraphsWendelin Boehmer, Vitaly Kurin, Shimon WhitesonICML 2020 · 被引用 209 次
- Network Randomization: A Simple Technique for Generalization in Deep Reinforcement LearningKimin Lee, Kibok Lee, Jinwoo Shin, Honglak LeeICLR 2020 · 被引用 191 次
相关 Paper
- The Dormant Neuron Phenomenon in Deep Reinforcement LearningGhada Sokar, Rishabh Agarwal, Pablo Samuel Castro, Utku EvciICML 2023 · 被引用 153 次
- Potentially Optimal Joint Actions Recognition for Cooperative Multi-Agent Reinforcement LearningChang Huang, Shatong Zhu, Junqiao Zhao, Hongtu Zhou 等ICLR 2026
- An Adaptive Entropy-Regularization Framework for Multi-Agent Reinforcement LearningWoojun Kim, Youngchul SungICML 2023 · 被引用 20 次
- ConcaveQ: Non-monotonic Value Function Factorization via Concave Representations in Deep Multi-Agent Reinforcement LearningHuiqun Li, Hanhan Zhou, Yifei Zou, Dongxiao Yu 等AAAI 2024 · 被引用 18 次
- Stay Hungry, Keep Learning: Sustainable Plasticity for Deep Reinforcement LearningHuaicheng Zhou, Zifeng Zhuang, Donglin WangICML 2025
