Rethinking the Stability-Plasticity Trade-off in Continual Learning from an Architectural Perspective
Aojun Lu, Hangjie Yuan, Tao Feng, Yanan Sun
摘要
The quest for Continual Learning (CL) seeks to empower neural networks with the ability to learn and adapt incrementally. Central to this pursuit is addressing the stability-plasticity dilemma, which involves striking a balance between two conflicting objectives: preserving previously learned knowledge and acquiring new knowledge. While numerous CL methods aim to achieve this tradeoff, they often overlook the impact of network architecture on stability and plasticity, restricting the trade-off to the parameter level. In this paper, we delve into the conflict between stability and plasticity at the architectural level. We reveal that under an equal parameter constraint, deeper networks exhibit better plasticity, while wider networks are characterized by superior stability. To address this architectural-level dilemma, we introduce a novel framework denoted Dual-Arch, which serves as a plug-in component for CL. This framework leverages the complementary strengths of two distinct and independent networks: one dedicated to plasticity and the other to stability. Each network is designed with a specialized and lightweight architecture, tailored to its respective objective. Extensive experiments demonstrate that Dual-Arch enhances the performance of existing CL methods while being up to 87% more compact in terms of parameters. Code: https: //github.com/byyx666/Dual-Arch .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper7
- Continual GUI AgentsZiwei Liu, Borui Kang, Hangjie Yuan, Zixiang Zhao 等ICML 2026 · 被引用 6 次
- FOREVER: Forgetting Curve-Inspired Memory Replay for Language Model Continual LearningYujie Feng, Hao Wang, Jian Li, Xu Chu 等ACL 2026 · 被引用 3 次
- Rep Deep & Machine Learning: Exemplar-Free Continual Video Action Recognition via Slow-Fast Collaborative LearningXueyi Zhang, Chengwei Zhang, Zheng Li, Xiyu Wang 等AAAI 2026 · 被引用 1 次
- ZeroFlow: Overcoming Catastrophic Forgetting is Easier than You ThinkTao Feng, Wei Li, Didi Zhu, Hangjie Yuan 等ICML 2025
- Plasticity Activation via Polar Operator: A Plug-in Method for Balancing Stability and PlasticityGuodong Zheng, Enneng Yang, Xiaoyan Wang, Yihan Chen 等ICML 2026
它引用的顶会 Paper19
- Dark Experience for General Continual Learning: a Strong, Simple BaselinePietro Buzzega, Matteo Boschini, Angelo Porrello, Davide Abati 等NeurIPS 2020 · 被引用 1,494 次
- Understanding the Role of Training Regimes in Continual LearningSeyed-Iman Mirzadeh, Mehrdad Farajtabar, Razvan Pascanu, Hassan GhasemzadehNeurIPS 2020 · 被引用 295 次
- DualNet: Continual Learning, Fast and SlowQuang Pham, Chenghao Liu, Steven C. H. HoiNeurIPS 2021 · 被引用 192 次
- Learning Fast, Learning Slow: A General Continual Learning Method based on Complementary Learning SystemElahe Arani, Fahad Sarfraz, Bahram ZonoozICLR 2022 · 被引用 168 次
- Overcoming Catastrophic Forgetting in Incremental Object Detection via Elastic Response DistillationTao Feng, Mang Wang, Hangjie YuanCVPR 2022 · 被引用 101 次
相关 Paper
- Recall-Oriented Continual Learning with Generative Adversarial Meta-ModelHaneol Kang, Dong-Wan ChoiAAAI 2024 · 被引用 3 次
- Adapt Before Continual LearningAojun Lu, Tao Feng, Hangjie Yuan, Chunhui Ding 等AAAI 2026
- Achieving a Better Stability-Plasticity Trade-off via Auxiliary Networks in Continual LearningSanghwan Kim, Lorenzo Noci, Antonio Orvieto, Thomas HofmannCVPR 2023
- Pareto Continual Learning: Preference-Conditioned Learning and Adaption for Dynamic Stability-Plasticity Trade-offSong Lai, Zhe Zhao, Fei Zhu, Xi Lin 等AAAI 2025 · 被引用 4 次
- Adaptive Aggregation Networks for Class-Incremental LearningYaoyao Liu, Bernt Schiele, Qianru SunCVPR 2021
