Continual evaluation for lifelong learning: Identifying the stability gap
Matthias De Lange, Gido M. van de Ven, Tinne Tuytelaars
Abstract
Time-dependent data-generating distributions have proven to be difficult for gradient-based training of neural networks, as the greedy updates result in catastrophic forgetting of previously learned knowledge. Despite the progress in the field of continual learning to overcome this forgetting, we show that a set of common state-of-the-art methods still suffers from substantial forgetting upon starting to learn new tasks, except that this forgetting is temporary and followed by a phase of performance recovery. We refer to this intriguing but potentially problematic phenomenon as the stability gap. The stability gap had likely remained under the radar due to standard practice in the field of evaluating continual learning models only after each task. Instead, we establish a framework for continual evaluation that uses per-iteration evaluation and we define a new set of metrics to quantify worst-case performance. Empirically we show that experience replay, constraint-based replay, knowledge-distillation, and parameter regularization methods are all prone to the stability gap; and that the stability gap can be observed in class-, task-, and domain-incremental learning benchmarks. Additionally, a controlled experiment shows that the stability gap increases when tasks are more dissimilar. Finally, by disentangling gradients into plasticity and stability components, we propose a conceptual explanation for the stability gap.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 2e75b170-2d73-43c6-b4f1-375beabc8a37Cited by top-tier papers11
- Addressing Loss of Plasticity and Catastrophic Forgetting in Continual LearningMohamed Elsayed, A. Rupam MahmoodICLR 2024 · 52 citations
- CLAP4CLIP: Continual Learning with Probabilistic Finetuning for Vision-Language ModelsSaurav Jha, Dong Gong, Lina YaoNeurIPS 2024 · 36 citations
- Evolving Parameterized Prompt Memory for Continual LearningMuhammad Rifki Kurniawan, Xiang Song, Zhiheng Ma, Yuhang He et al.AAAI 2024 · 31 citations
- Layerwise Proximal Replay: A Proximal Point Method for Online Continual LearningJinsoo Yoo, Yunpeng Liu, Frank Wood, Geoff PleissICML 2024 · 13 citations
- Maintaining Fairness in Logit-based Knowledge Distillation for Class-Incremental LearningZijian Gao, Shanhao Han, Xingxing Zhang, Kele Xu et al.AAAI 2025 · 10 citations
Builds on8
- Task2Vec: Task Embedding for Meta-LearningAlessandro Achille, Michael Lam, Rahul Tewari, Avinash Ravichandran et al.ICCV 2019 · 359 citations
- New Insights on Reducing Abrupt Representation Change in Online Continual LearningLucas Caccia, Rahaf Aljundi, Nader Asadi, Tinne Tuytelaars et al.ICLR 2022 · 279 citations
- Online Class-Incremental Continual Learning with Adversarial Shapley ValueDongsub Shim, Zheda Mai, Jihwan Jeong, Scott Sanner et al.AAAI 2021 · 262 citations
- Continual Prototype Evolution: Learning Online from Non-Stationary Data StreamsMatthias De Lange, Tinne TuytelaarsICCV 2021 · 251 citations
- Online Continual Learning from Imbalanced DataAristotelis Chrysakis, Marie-Francine MoensICML 2020 · 166 citations
Related papers
- Preserving Linear Separability in Continual Learning by Backward Feature ProjectionQiao Gu, Dongsub Shim, Florian ShkurtiCVPR 2023
- Dark Experience for General Continual Learning: a Strong, Simple BaselinePietro Buzzega, Matteo Boschini, Angelo Porrello, Davide Abati et al.NeurIPS 2020 · 1,494 citations
- STAR: Stability-Inducing Weight Perturbation for Continual LearningMasih Eskandar, Tooba Imtiaz, Davin Hill, Zifeng Wang et al.ICLR 2025
- Understanding the Role of Training Regimes in Continual LearningSeyed-Iman Mirzadeh, Mehrdad Farajtabar, Razvan Pascanu, Hassan GhasemzadehNeurIPS 2020 · 295 citations
- Temporal-Difference Variational Continual LearningLuckeciano Carvalho Melo, Alessandro Abate, Yarin GalNeurIPS 2025 · 1 citation
