Self-Normalized Resets for Plasticity in Continual Learning
Vivek F. Farias, Adam Daniel Jozefiak
Abstract
Plasticity Loss is an increasingly important phenomenon that refers to the empirical observation that as a neural network is continually trained on a sequence of changing tasks, its ability to adapt to a new task diminishes over time. We introduce Self-Normalized Resets (SNR), a simple adaptive algorithm that mitigates plasticity loss by resetting a neuron's weights when evidence suggests its firing rate has effectively dropped to zero. Across a battery of continual learning problems and network architectures, we demonstrate that SNR consistently attains superior performance compared to its competitor algorithms. We also demonstrate that SNR is robust to its sole hyperparameter, its rejection percentile threshold, while competitor algorithms show significant sensitivity. SNR's threshold-based reset mechanism is motivated by a simple hypothesis test that we derive. Seen through the lens of this hypothesis test, competing reset proposals yield suboptimal error rates in correctly detecting inactive neurons, potentially explaining our experimental observations. We also conduct a theoretical investigation of the optimization landscape for the problem of learning a single ReLU. We show that even when initialized adversarially, an idealized version of SNR learns the target ReLU, while regularization based approaches can fail to learn.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 4232aae2-6717-4bd4-9bd3-4168c4f69bf6Cited by top-tier papers7
- Measure gradients, not activations! Enhancing neuronal activity in deep reinforcement learningJiashun Liu, Zihao Wu, Johan S. Obando-Ceron, Pablo Samuel Castro et al.NeurIPS 2025 · 15 citations
- Forget Forgetting: Continual Learning in a World of Abundant MemoryDongkyu Cho, Taesup Moon, Rumi Chunara, Kyunghyun Cho et al.ICLR 2026 · 9 citations
- FIRE: Frobenius-Isometry Reinitialization for Balancing the Stability-Plasticity TradeoffIsaac Han, Sangyeon Park, Seungwon Oh, Donghu Kim et al.ICLR 2026 · 7 citations
- Activation Function Design Sustains Plasticity in Continual LearningLute Lillo, Nick CheneyICLR 2026 · 6 citations
- The Dual Nature of Plasticity Loss in Deep Continual Learning: Dissection and MitigationHaoyu Wang, Wei Dai, Jiawei Zhang, Jialun Ma et al.NeurIPS 2025 · 2 citations
Builds on3
- On Warm-Starting Neural Network TrainingJordan T. Ash, Ryan P. AdamsNeurIPS 2020 · 288 citations
- Implicit Under-Parameterization Inhibits Data-Efficient Deep Reinforcement LearningAviral Kumar, Rishabh Agarwal, Dibya Ghosh, Sergey LevineICLR 2021 · 155 citations
- Slow and Steady Wins the Race: Maintaining Plasticity with Hare and Tortoise NetworksHojoon Lee, Hyeonseo Cho, Hyunseung Kim, Donghu Kim et al.ICML 2024 · 36 citations
Related papers
- Learning Continually by Spectral RegularizationAlex Lewandowski, Michal Bortkiewicz, Saurabh Kumar, András György et al.ICLR 2025
- A Study of Plasticity Loss in On-Policy Deep Reinforcement LearningArthur Juliani, Jordan T. AshNeurIPS 2024 · 37 citations
- Activation by Interval-wise Dropout: A Simple Way to Prevent Neural Networks from Plasticity LossSangyeon Park, Isaac Han, Seungwon Oh, Kyung-Joong KimICML 2025
- Normalization and effective learning rates in reinforcement learningClare Lyle, Zeyu Zheng, Khimya Khetarpal, James Martens et al.NeurIPS 2024 · 69 citations
- Spectral Collapse Drives Loss of Plasticity in Deep Continual LearningArjun Prakash, Naicheng He, Kaicheng Guo, Saket Tiwari et al.ICML 2026
