Mitigating Catastrophic Forgetting in Online Continual Learning by Modeling Previous Task Interrelations via Pareto Optimization
Yichen Wu, Hong Wang, Peilin Zhao, Yefeng Zheng, Ying Wei, Long-Kai Huang
摘要
Catastrophic forgetting remains a core challenge in continual learning (CL), where the models struggle to retain previous knowledge when learning new tasks. While existing replay-based CL methods have been proposed to tackle this challenge by utilizing a memory buffer to store data from previous tasks, they generally overlook the interdependence between previously learned tasks and fail to encapsulate the optimally integrated knowledge in previous tasks, leading to sub-optimal performance of the previous tasks. Against this issue, we first reformulate replaybased CL methods as a unified hierarchical gradient aggregation framework. We then incorporate the Pareto optimization to capture the interrelationship among previously learned tasks and design a Pareto-Optimized CL algorithm (POCL), which effectively enhances the overall performance of past tasks while ensuring the performance of the current task. To further stabilize the gradients of different tasks, we carefully devise a hyper-gradient-based implementation manner for POCL. Comprehensive empirical results demonstrate that the proposed POCL outperforms current state-of-the-art CL methods across multiple datasets and different settings.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper10
- A Layer Selection Approach to Test Time AdaptationSabyasachi Sahoo, Mostafa ElAraby, Jonas Ngnawé, Yann Batiste Pequignot 等AAAI 2025 · 被引用 6 次
- Mitigating Catastrophic Forgetting in Large Language Models with Forgetting-aware PruningWei Huang, Anda Cheng, Yinggui WangEMNLP 2025 · 被引用 4 次
- Pareto Continual Learning: Preference-Conditioned Learning and Adaption for Dynamic Stability-Plasticity Trade-offSong Lai, Zhe Zhao, Fei Zhu, Xi Lin 等AAAI 2025 · 被引用 4 次
- Gradient-Guided Epsilon Constraint Method for Online Continual LearningSong Lai, Changyi Ma, Fei Zhu, Zhe Zhao 等NeurIPS 2025 · 被引用 3 次
- FedAGC: Federated Continual Learning with Asymmetric Gradient CorrectionChengchao Zhang, Fanhua Shang, Hongyin Liu, Liang Wan 等ICCV 2025 · 被引用 3 次
它引用的顶会 Paper10
- Dark Experience for General Continual Learning: a Strong, Simple BaselinePietro Buzzega, Matteo Boschini, Angelo Porrello, Davide Abati 等NeurIPS 2020 · 被引用 1,494 次
- New Insights on Reducing Abrupt Representation Change in Online Continual LearningLucas Caccia, Rahaf Aljundi, Nader Asadi, Tinne Tuytelaars 等ICLR 2022 · 被引用 279 次
- Learning Fast, Learning Slow: A General Continual Learning Method based on Complementary Learning SystemElahe Arani, Fahad Sarfraz, Bahram ZonoozICLR 2022 · 被引用 168 次
- Improved Schemes for Episodic Memory-based Lifelong LearningYunhui Guo, Mingrui Liu, Tianbao Yang, Tajana RosingNeurIPS 2020 · 被引用 105 次
- Look-ahead Meta Learning for Continual LearningGunshi Gupta, Karmesh Yadav, Liam PaullNeurIPS 2020 · 被引用 74 次
相关 Paper
- The Ideal Continual Learner: An Agent That Never ForgetsLiangzu Peng, Paris Giampouras, René VidalICML 2023 · 被引用 39 次
- Continual Learning with Recursive Gradient OptimizationHao Liu, Huaping LiuICLR 2022 · 被引用 52 次
- Dealing with Cross-Task Class Discrimination in Online Continual LearningYiduo Guo, Bing Liu, Dongyan ZhaoCVPR 2023
- GCR: Gradient Coreset based Replay Buffer Selection for Continual LearningRishabh Tiwari, KrishnaTeja Killamsetty, Rishabh K. Iyer, Pradeep ShenoyCVPR 2022 · 被引用 102 次
- Layerwise Optimization by Gradient Decomposition for Continual LearningShixiang Tang, Dapeng Chen, Jinguo Zhu, Shijie Yu 等CVPR 2021
