Layerwise Proximal Replay: A Proximal Point Method for Online Continual Learning
Jinsoo Yoo, Yunpeng Liu, Frank Wood, Geoff Pleiss
Abstract
In online continual learning, a neural network incrementally learns from a non-i.i.d. data stream. Nearly all online continual learning methods employ experience replay to simultaneously prevent catastrophic forgetting and underfitting on past data. Our work demonstrates a limitation of this approach: neural networks trained with experience replay tend to have unstable optimization trajectories, impeding their overall accuracy. Surprisingly, these instabilities persist even when the replay buffer stores all previous training examples, suggesting that this issue is orthogonal to catastrophic forgetting. We minimize these instabilities through a simple modification of the optimization geometry. Our solution, Layerwise Proximal Replay (LPR), balances learning from new and replay data while only allowing for gradual changes in the hidden activation of past data. We demonstrate that LPR consistently improves replay-based online continual learning methods across multiple problem settings, regardless of the amount of available replay memory.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext bfdcc8d1-5e28-4e08-bd13-14cc6486a269Cited by top-tier papers3
- Region-Based Optimization in Continual Learning for Audio Deepfake DetectionYujie Chen, Jiangyan Yi, Cunhang Fan, Jianhua Tao et al.AAAI 2025 · 10 citations
- PROL: Rehearsal Free Continual Learning in Streaming Data via Prompt Online LearningM. Anwar Ma'sum, Mahardhika Pratama, Savitha Ramasamy, Lin Liu et al.ICCV 2025 · 1 citation
- Online Curvature-Aware Replay: Leveraging 2nd Order Information for Online Continual LearningEdoardo Urettini, Antonio CartaICML 2025
Builds on13
- Dark Experience for General Continual Learning: a Strong, Simple BaselinePietro Buzzega, Matteo Boschini, Angelo Porrello, Davide Abati et al.NeurIPS 2020 · 1,494 citations
- Gradient Projection Memory for Continual LearningGobinda Saha, Isha Garg, Kaushik RoyICLR 2021 · 409 citations
- On Warm-Starting Neural Network TrainingJordan T. Ash, Ryan P. AdamsNeurIPS 2020 · 288 citations
- New Insights on Reducing Abrupt Representation Change in Online Continual LearningLucas Caccia, Rahaf Aljundi, Nader Asadi, Tinne Tuytelaars et al.ICLR 2022 · 279 citations
- Flattening Sharpness for Dynamic Gradient Projection Memory Benefits Continual LearningDanruo Deng, Guangyong Chen, Jianye Hao, Qiong Wang et al.NeurIPS 2021 · 112 citations
Related papers
- Using Hindsight to Anchor Past Knowledge in Continual LearningArslan Chaudhry, Albert Gordo, Puneet K. Dokania, Philip H. S. Torr et al.AAAI 2021 · 279 citations
- Online Bias Correction for Task-Free Continual LearningAristotelis Chrysakis, Marie-Francine MoensICLR 2023
- Online Prototype Learning for Online Continual LearningYujie Wei, Jiaxin Ye, Zhizhong Huang, Junping Zhang et al.ICCV 2023 · 78 citations
- PCR: Proxy-Based Contrastive Replay for Online Class-Incremental Continual LearningHuiwei Lin, Baoquan Zhang, Shanshan Feng, Xutao Li et al.CVPR 2023
- Retrospective Adversarial Replay for Continual LearningLilly Kumari, Shengjie Wang, Tianyi Zhou, Jeff A. BilmesNeurIPS 2022 · 57 citations
