Neuroplastic Expansion in Deep Reinforcement Learning
Jiashun Liu, Johan S. Obando-Ceron, Aaron C. Courville, Ling Pan
摘要
The loss of plasticity in learning agents, analogous to the solidification of neural pathways in biological brains, significantly impedes learning and adaptation in reinforcement learning due to its non-stationary nature. To address this fundamental challenge, we propose a novel approach, Neuroplastic Expansion (NE), inspired by cortical expansion in cognitive science. NE maintains learnability and adaptability throughout the entire training process by dynamically growing the network from a smaller initial size to its full dimension. Our method is designed with three key components: (1) elastic topology generation based on potential gradients, (2) dormant neuron pruning to optimize network expressivity, and (3) neuron consolidation via experience review to strike a balance in the plasticitystability dilemma. Extensive experiments demonstrate that NE effectively mitigates plasticity loss and outperforms state-of-the-art methods across various tasks in MuJoCo and DeepMind Control Suite environments. NE enables more adaptive learning in complex, dynamic environments, which represents a crucial step towards transitioning deep reinforcement learning from static, one-time training paradigms to more flexible, continually adapting models. We make our code publicly available.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper11
- Stable Gradients for Stable Learning at Scale in Deep Reinforcement LearningRoger Creus Castanyer, Johan S. Obando-Ceron, Lu Li, Pierre-Luc Bacon 等NeurIPS 2025 · 被引用 26 次
- XQC: Well-conditioned Optimization Accelerates Deep Reinforcement LearningDaniel Palenicek, Florian Vogt, Joe Watson, Ingmar Posner 等ICLR 2026 · 被引用 20 次
- Measure gradients, not activations! Enhancing neuronal activity in deep reinforcement learningJiashun Liu, Zihao Wu, Johan S. Obando-Ceron, Pablo Samuel Castro 等NeurIPS 2025 · 被引用 15 次
- Simplicial Embeddings Improve Sample Efficiency in Actor–Critic AgentsJohan Obando-Ceron, Walter Mayor, Samuel Lavoie, Scott Fujimoto 等ICLR 2026 · 被引用 12 次
- Asymmetric Proximal Policy Optimization: mini-critics boost LLM reasoningJiashun Liu, Johan S. Obando-Ceron, Han Lu, Yancheng He 等ICLR 2026 · 被引用 11 次
它引用的顶会 Paper26
- Rigging the Lottery: Making All Tickets WinnersUtku Evci, Trevor Gale, Jacob Menick, Pablo Samuel Castro 等ICML 2020 · 被引用 723 次
- Mastering Visual Continuous Control: Improved Data-Augmented Reinforcement LearningDenis Yarats, Rob Fergus, Alessandro Lazaric, Lerrel PintoICLR 2022 · 被引用 457 次
- On Warm-Starting Neural Network TrainingJordan T. Ash, Ryan P. AdamsNeurIPS 2020 · 被引用 288 次
- The Primacy Bias in Deep Reinforcement LearningEvgenii Nikishin, Max Schwarzer, Pierluca D'Oro, Pierre-Luc Bacon 等ICML 2022 · 被引用 269 次
- Understanding Plasticity in Neural NetworksClare Lyle, Zeyu Zheng, Evgenii Nikishin, Bernardo Ávila Pires 等ICML 2023 · 被引用 162 次
相关 Paper
- Balancing Plasticity and Stability with Fast and Slow Successor FeaturesRaymond Chua, Doina Precup, Blake RichardsICML 2026
- Rewiring Neurons in Non-Stationary EnvironmentsZhicheng Sun, Yadong MuNeurIPS 2023 · 被引用 4 次
- Mitigating Plasticity Loss in Continual Reinforcement Learning by Reducing ChurnHongyao Tang, Johan S. Obando-Ceron, Pablo Samuel Castro, Aaron C. Courville 等ICML 2025
- Meta-Reinforcement Learning with Self-Modifying NetworksMathieu Chalvidal, Thomas Serre, Rufin VanRullenNeurIPS 2022 · 被引用 13 次
- A Forget-and-Grow Strategy for Deep Reinforcement Learning Scaling in Continuous ControlZilin Kang, Chenyuan Hu, Yu Luo, Zhecheng Yuan 等ICML 2025
