Self-Normalized Resets for Plasticity in Continual Learning
Vivek F. Farias, Adam Daniel Jozefiak
摘要
Plasticity Loss is an increasingly important phenomenon that refers to the empirical observation that as a neural network is continually trained on a sequence of changing tasks, its ability to adapt to a new task diminishes over time. We introduce Self-Normalized Resets (SNR), a simple adaptive algorithm that mitigates plasticity loss by resetting a neuron's weights when evidence suggests its firing rate has effectively dropped to zero. Across a battery of continual learning problems and network architectures, we demonstrate that SNR consistently attains superior performance compared to its competitor algorithms. We also demonstrate that SNR is robust to its sole hyperparameter, its rejection percentile threshold, while competitor algorithms show significant sensitivity. SNR's threshold-based reset mechanism is motivated by a simple hypothesis test that we derive. Seen through the lens of this hypothesis test, competing reset proposals yield suboptimal error rates in correctly detecting inactive neurons, potentially explaining our experimental observations. We also conduct a theoretical investigation of the optimization landscape for the problem of learning a single ReLU. We show that even when initialized adversarially, an idealized version of SNR learns the target ReLU, while regularization based approaches can fail to learn.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper7
- Measure gradients, not activations! Enhancing neuronal activity in deep reinforcement learningJiashun Liu, Zihao Wu, Johan S. Obando-Ceron, Pablo Samuel Castro 等NeurIPS 2025 · 被引用 15 次
- Forget Forgetting: Continual Learning in a World of Abundant MemoryDongkyu Cho, Taesup Moon, Rumi Chunara, Kyunghyun Cho 等ICLR 2026 · 被引用 9 次
- FIRE: Frobenius-Isometry Reinitialization for Balancing the Stability-Plasticity TradeoffIsaac Han, Sangyeon Park, Seungwon Oh, Donghu Kim 等ICLR 2026 · 被引用 7 次
- Activation Function Design Sustains Plasticity in Continual LearningLute Lillo, Nick CheneyICLR 2026 · 被引用 6 次
- The Dual Nature of Plasticity Loss in Deep Continual Learning: Dissection and MitigationHaoyu Wang, Wei Dai, Jiawei Zhang, Jialun Ma 等NeurIPS 2025 · 被引用 2 次
它引用的顶会 Paper3
- On Warm-Starting Neural Network TrainingJordan T. Ash, Ryan P. AdamsNeurIPS 2020 · 被引用 288 次
- Implicit Under-Parameterization Inhibits Data-Efficient Deep Reinforcement LearningAviral Kumar, Rishabh Agarwal, Dibya Ghosh, Sergey LevineICLR 2021 · 被引用 155 次
- Slow and Steady Wins the Race: Maintaining Plasticity with Hare and Tortoise NetworksHojoon Lee, Hyeonseo Cho, Hyunseung Kim, Donghu Kim 等ICML 2024 · 被引用 36 次
相关 Paper
- Learning Continually by Spectral RegularizationAlex Lewandowski, Michal Bortkiewicz, Saurabh Kumar, András György 等ICLR 2025
- A Study of Plasticity Loss in On-Policy Deep Reinforcement LearningArthur Juliani, Jordan T. AshNeurIPS 2024 · 被引用 37 次
- Activation by Interval-wise Dropout: A Simple Way to Prevent Neural Networks from Plasticity LossSangyeon Park, Isaac Han, Seungwon Oh, Kyung-Joong KimICML 2025
- Normalization and effective learning rates in reinforcement learningClare Lyle, Zeyu Zheng, Khimya Khetarpal, James Martens 等NeurIPS 2024 · 被引用 69 次
- Spectral Collapse Drives Loss of Plasticity in Deep Continual LearningArjun Prakash, Naicheng He, Kaicheng Guo, Saket Tiwari 等ICML 2026
