Task Difficulty Aware Parameter Allocation & Regularization for Lifelong Learning
Wenjin Wang, Yunqing Hu, Qianglong Chen, Yin Zhang
Abstract
Parameter regularization or allocation methods are effective in overcoming catastrophic forgetting in lifelong learning. However, they solve all tasks in a sequence uniformly and ignore the differences in the learning difficulty of different tasks. So parameter regularization methods face significant forgetting when learning a new task very different from learned tasks, and parameter allocation methods face unnecessary parameter overhead when learning simple tasks. In this paper, we propose the Parameter Allocation & Regularization (PAR), which adaptively select an appropriate strategy for each task from parameter allocation and regularization based on its learning difficulty. A task is easy for a model that has learned tasks related to it and vice versa. We propose a divergence estimation method based on the Nearest-Prototype distance to measure the task relatedness using only features of the new task. Moreover, we propose a time-efficient relatedness-aware sampling-based architecture search strategy to reduce the parameter overhead for allocation. Experimental results on multiple benchmarks demonstrate that, compared with SOTAs, our method is scalable and significantly reduces the model's redundancy while improving the model's performance. Further qualitative analysis indicates that PAR obtains reasonable task-relatedness.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers5
- Mitigating Catastrophic Forgetting in Target Language Adaptation of LLMs via Source-Shielded UpdatesAtsuki Yamaguchi, Terufumi Morishita, Aline Villavicencio, Nikolaos AletrasACL 2026 · 3 citations
- Not Just Object, But State: Compositional Incremental Learning without ForgettingYanyi Zhang, Binglin Qiu, Qi Jia, Yu Liu et al.NeurIPS 2024 · 2 citations
- Hybrid Re-matching for Continual Learning with Parameter-Efficient TuningWeicheng Wang, Guoli Jia, Xialei Liu, Liang Lin et al.NeurIPS 2025
- Knowledge Swapping via Learning and UnlearningMingyu Xing, Lechao Cheng, Shengeng Tang, Yaxiong Wang et al.ICML 2025
- Learning Conditional Space-Time Prompt Distributions for Video Class-Incremental LearningXiaohan Zou, Wenchao Ma, Shu ZhaoCVPR 2025
Builds on18
- Dark Experience for General Continual Learning: a Strong, Simple BaselinePietro Buzzega, Matteo Boschini, Angelo Porrello, Davide Abati et al.NeurIPS 2020 · 1,494 citations
- BatchEnsemble: an Alternative Approach to Efficient Ensemble and Lifelong LearningYeming Wen, Dustin Tran, Jimmy BaICLR 2020 · 569 citations
- Gradient Projection Memory for Continual LearningGobinda Saha, Isha Garg, Kaushik RoyICLR 2021 · 409 citations
- Continual Learning of a Mixed Sequence of Similar and Dissimilar TasksZixuan Ke, Bing Liu, Xingchang HuangNeurIPS 2020 · 173 citations
- Overcoming Catastrophic Forgetting in Graph Neural NetworksHuihui Liu, Yiding Yang, Xinchao WangAAAI 2021 · 166 citations
Related papers
- Parameter-efficient Continual Learning for Enhancing Plasticity without Forgetting under Limited Model CapacityYitian Chen, Shigeng Zhang, Xuan Liu, Mingming Lu et al.CVPR 2026
- Adaptive Plasticity Improvement for Continual LearningYan-Shuo Liang, Wu-Jun LiCVPR 2023
- Sharing Less is More: Lifelong Learning in Deep Networks with Selective Layer TransferSeungwon Lee, Sima Behpour, Eric EatonICML 2021 · 21 citations
- Is Forgetting Less a Good Inductive Bias for Forward Transfer?Jiefeng Chen, Timothy Nguyen, Dilan Görür, Arslan ChaudhryICLR 2023 · 1 citation
- Multi-Domain Multi-Task Rehearsal for Lifelong LearningFan Lyu, Shuai Wang, Wei Feng, Zihan Ye et al.AAAI 2021 · 34 citations
