Gradient-Guided Epsilon Constraint Method for Online Continual Learning
Song Lai, Changyi Ma, Fei Zhu, Zhe Zhao, Xi Lin, Gaofeng Meng, Qingfu Zhang
Abstract
Online Continual Learning (OCL) requires models to learn sequentially from data streams with limited memory. Rehearsal-based methods, particularly Experience Replay (ER), are commonly used in OCL scenarios. This paper revisits ER through the lens of ϵ -constraint optimization, revealing that ER implicitly employs a soft constraint on past task performance, with its weighting parameter post-hoc defining a slack variable. While effective, ER’s implicit and fixed slack strategy has limitations: it can inadvertently lead to updates that negatively impact generalization, and its fixed trade-off between plasticity and stability may not optimally balance current streaming with memory retention, potentially over-fitting to the memory buffer. To address these shortcomings, we propose the G radient-Guided E psilon C onstraint ( GEC ) method for online continual learning. GEC explicitly formulates the OCL update as an ϵ -constraint optimization problem, which minimize the loss on the current task data and transform the stability objective as constraints and propose a gradient-guided method to dynamically adjusts the update direction based on whether the performance on memory samples violates a predefined slack tolerance ¯ ε : if forgetting exceeds this tolerance, GEC prioritizes constraint satisfaction; otherwise, it focuses on the current task while controlling the rate of increase in memory loss. Empirical evaluations on standard OCL benchmarks demonstrate GEC’s ability to achieve a superior trade-off, leading to improved overall performance. Code is available at https://github.com/laisong-22004009/GEC_OCL .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 1ecf4304-8008-4c7d-ac38-6df6909fd686Builds on10
- Dark Experience for General Continual Learning: a Strong, Simple BaselinePietro Buzzega, Matteo Boschini, Angelo Porrello, Davide Abati et al.NeurIPS 2020 · 1,494 citations
- Learning Fast, Learning Slow: A General Continual Learning Method based on Complementary Learning SystemElahe Arani, Fahad Sarfraz, Bahram ZonoozICLR 2022 · 168 citations
- Look-ahead Meta Learning for Continual LearningGunshi Gupta, Karmesh Yadav, Liam PaullNeurIPS 2020 · 74 citations
- Meta Continual Learning Revisited: Implicitly Enhancing Online Hessian Approximation via Variance ReductionYichen Wu, Long-Kai Huang, Renzhen Wang, Deyu Meng et al.ICLR 2024 · 42 citations
- A Unified and General Framework for Continual LearningZhenyi Wang, Yan Li, Li Shen, Heng HuangICLR 2024 · 42 citations
Related papers
- Layerwise Proximal Replay: A Proximal Point Method for Online Continual LearningJinsoo Yoo, Yunpeng Liu, Frank Wood, Geoff PleissICML 2024 · 13 citations
- GCR: Gradient Coreset based Replay Buffer Selection for Continual LearningRishabh Tiwari, KrishnaTeja Killamsetty, Rishabh K. Iyer, Pradeep ShenoyCVPR 2022 · 102 citations
- Using Hindsight to Anchor Past Knowledge in Continual LearningArslan Chaudhry, Albert Gordo, Puneet K. Dokania, Philip H. S. Torr et al.AAAI 2021 · 279 citations
- A simple but strong baseline for online continual learning: Repeated Augmented RehearsalYaqian Zhang, Bernhard Pfahringer, Eibe Frank, Albert Bifet et al.NeurIPS 2022 · 13 citations
- Navigating Memory Construction by Global Pseudo-Task Simulation for Continual LearningYejia Liu, Wang Zhu, Shaolei RenNeurIPS 2022 · 4 citations
