Learning without Isolation: Pathway Protection for Continual Learning
Zhikang Chen, Abudukelimu Wuerkaixi, Sen Cui, Haoxuan Li, Ding Li, Jingfeng Zhang, Bo Han, Gang Niu, Houfang Liu, Yi Yang, Sifan Yang, Changshui Zhang, Tianling Ren
Abstract
Deep networks are prone to catastrophic forgetting during sequential task learning, i.e., losing the knowledge about old tasks upon learning new tasks. To this end, continual learning (CL) has emerged, whose existing methods focus mostly on regulating or protecting the parameters associated with the previous tasks. However, parameter protection is often impractical, since the size of parameters for storing the old-task knowledge increases linearly with the number of tasks, otherwise it is hard to preserve the parameters related to the old-task knowledge. In this work, we bring a dual opinion from neuroscience and physics to CL: in the whole networks, the pathways matter more than the parameters when concerning the knowledge acquired from the old tasks. Following this opinion, we propose a novel CL framework, learning without isolation (LwI), where model fusion is formulated as graph matching and the pathways occupied by the old tasks are protected without being isolated. Thanks to the sparsity of activation channels in a deep network, LwI can adaptively allocate available pathways for a new task, realizing pathway protection and addressing catastrophic forgetting in a parameterefficient manner. Experiments on popular benchmark datasets demonstrate the superiority of the proposed LwI.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 3fd132ca-fe97-4bf5-bdd7-35a12f207e1bCited by top-tier papers2
- Decentralized Dynamic Cooperation of Personalized Models for Federated Continual LearningDanni Yang, Zhikang Chen, Sen Cui, Mengyue Yang et al.NeurIPS 2025 · 2 citations
- HAD: Heterogeneity-Aware Distillation for Lifelong Heterogeneous LearningXuerui Zhang, Xuehao Wang, Zhan Zhuang, Linglan Zhao et al.CVPR 2026
Builds on16
- LoRA: Low-Rank Adaptation of Large Language ModelsEdward J. Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu et al.ICLR 2022 · 18,833 citations
- Gradient Projection Memory for Continual LearningGobinda Saha, Isha Garg, Kaushik RoyICLR 2021 · 409 citations
- Supermasks in SuperpositionMitchell Wortsman, Vivek Ramanujan, Rosanne Liu, Aniruddha Kembhavi et al.NeurIPS 2020 · 364 citations
- Model Fusion via Optimal TransportSidak Pal Singh, Martin JaggiNeurIPS 2020 · 330 citations
- Class-Incremental Learning by Knowledge Distillation with Adaptive Feature ConsolidationMinsoo Kang, Jaeyoo Park, Bohyung HanCVPR 2022 · 189 citations
Related papers
- Growing a Brain with Sparsity-Inducing Generation for Continual LearningHyundong Jin, Gyeong-Hyeon Kim, Chanho Ahn, Eunwoo KimICCV 2023 · 7 citations
- Mitigating Forgetting in Online Continual Learning via Instance-Aware ParameterizationHung-Jen Chen, An-Chieh Cheng, Da-Cheng Juan, Wei Wei et al.NeurIPS 2020 · 50 citations
- Split-and-Bridge: Adaptable Class Incremental Learning within a Single Neural NetworkJong-Yeong Kim, Dong-Wan ChoiAAAI 2021 · 28 citations
- Residual Continual LearningJanghyeon Lee, Donggyu Joo, Hyeong Gwon Hong, Junmo KimAAAI 2020 · 25 citations
- Continual Learning on Dynamic Graphs via Parameter IsolationPeiyan Zhang, Yuchen Yan, Chaozhuo Li, Senzhang Wang et al.SIGIR 2023 · 45 citations
