Task-aware Orthogonal Sparse Network for Exploring Shared Knowledge in Continual Learning
Yusong Hu, De Cheng, Dingwen Zhang, Nannan Wang, Tongliang Liu, Xinbo Gao
摘要
Continual learning (CL) aims to learn from sequentially arriving tasks without catastrophic forgetting (CF). By partitioning the network into two parts based on the Lottery Ticket Hypothesisone for holding the knowledge of the old tasks while the other for learning the knowledge of the new task-the recent progress has achieved forget-free CL. Although addressing the CF issue well, such methods would encounter serious under-fitting in long-term CL, in which the learning process will continue for a long time and the number of new tasks involved will be much higher. To solve this problem, this paper partitions the network into three parts-with a new part for exploring the knowledge sharing between the old and new tasks. With the shared knowledge, this part of network can be learnt to simultaneously consolidate the old tasks and fit to the new task. To achieve this goal, we propose a task-aware Orthogonal Sparse Network (OSN), which contains shared knowledge induced network partition and sharpness-aware orthogonal sparse network learning. The former partitions the network to select shared parameters, while the latter guides the exploration of shared knowledge through shared parameters. Qualitative and quantitative analyses, show that the proposed OSN induces minimum to no interference with past tasks, i.e., approximately no forgetting, while greatly improves the model plasticity and capacity, and finally achieves the state-of-the-art performances.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- MOS: Model Surgery for Pre-Trained Model-Based Class-Incremental LearningHai-Long Sun, Da-Wei Zhou, Hanbin Zhao, Le Gan 等AAAI 2025 · 被引用 31 次
- Training Consistent Mixture-of-Experts-Based Prompt Generator for Continual LearningYue Lu, Shizhou Zhang, De Cheng, Guoqiang Liang 等AAAI 2025 · 被引用 8 次
- Dual Consolidation for Pre-Trained Model-Based Domain-Incremental LearningDa-Wei Zhou, Zi-Wen Cai, Han-Jia Ye, Lijun Zhang 等CVPR 2025
- Probabilistic Group Mask Guided Discrete Optimization for Incremental LearningFengqiang Wan, Yang YangICML 2025
它引用的顶会 Paper9
- Gradient Projection Memory for Continual LearningGobinda Saha, Isha Garg, Kaushik RoyICLR 2021 · 被引用 409 次
- Overcoming Catastrophic Forgetting With Unlabeled Data in the WildKibok Lee, Kimin Lee, Jinwoo Shin, Honglak LeeICCV 2019 · 被引用 231 次
- Forget-free Continual Learning with Winning SubnetworksHaeyong Kang, Rusty John Lloyd Mina, Sultan Rizky Hikmawan Madjid, Jaehong Yoon 等ICML 2022 · 被引用 159 次
- Flattening Sharpness for Dynamic Gradient Projection Memory Benefits Continual LearningDanruo Deng, Guangyong Chen, Jianye Hao, Qiong Wang 等NeurIPS 2021 · 被引用 112 次
- TRGP: Trust Region Gradient Projection for Continual LearningSen Lin, Li Yang, Deliang Fan, Junshan ZhangICLR 2022 · 被引用 107 次
相关 Paper
- Parameter-Level Soft-Masking for Continual LearningTatsuya Konishi, Mori Kurokawa, Chihiro Ono, Zixuan Ke 等ICML 2023 · 被引用 63 次
- Enhancing Knowledge Transfer for Task Incremental Learning with Data-free SubnetworkQiang Gao, Xiaojun Shan, Yuchen Zhang, Fan ZhouNeurIPS 2023 · 被引用 13 次
- BNS: Building Network Structures Dynamically for Continual LearningQi Qin, Wenpeng Hu, Han Peng, Dongyan Zhao 等NeurIPS 2021 · 被引用 54 次
- Learning Bayesian Sparse Networks with Full Experience Replay for Continual LearningQingsen Yan, Dong Gong, Yuhang Liu, Anton van den Hengel 等CVPR 2022 · 被引用 38 次
- Long Live the Lottery: The Existence of Winning Tickets in Lifelong LearningTianlong Chen, Zhenyu Zhang, Sijia Liu, Shiyu Chang 等ICLR 2021 · 被引用 23 次
