Task Difficulty Aware Parameter Allocation & Regularization for Lifelong Learning
Wenjin Wang, Yunqing Hu, Qianglong Chen, Yin Zhang
摘要
Parameter regularization or allocation methods are effective in overcoming catastrophic forgetting in lifelong learning. However, they solve all tasks in a sequence uniformly and ignore the differences in the learning difficulty of different tasks. So parameter regularization methods face significant forgetting when learning a new task very different from learned tasks, and parameter allocation methods face unnecessary parameter overhead when learning simple tasks. In this paper, we propose the Parameter Allocation & Regularization (PAR), which adaptively select an appropriate strategy for each task from parameter allocation and regularization based on its learning difficulty. A task is easy for a model that has learned tasks related to it and vice versa. We propose a divergence estimation method based on the Nearest-Prototype distance to measure the task relatedness using only features of the new task. Moreover, we propose a time-efficient relatedness-aware sampling-based architecture search strategy to reduce the parameter overhead for allocation. Experimental results on multiple benchmarks demonstrate that, compared with SOTAs, our method is scalable and significantly reduces the model's redundancy while improving the model's performance. Further qualitative analysis indicates that PAR obtains reasonable task-relatedness.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- Mitigating Catastrophic Forgetting in Target Language Adaptation of LLMs via Source-Shielded UpdatesAtsuki Yamaguchi, Terufumi Morishita, Aline Villavicencio, Nikolaos AletrasACL 2026 · 被引用 3 次
- Not Just Object, But State: Compositional Incremental Learning without ForgettingYanyi Zhang, Binglin Qiu, Qi Jia, Yu Liu 等NeurIPS 2024 · 被引用 2 次
- Hybrid Re-matching for Continual Learning with Parameter-Efficient TuningWeicheng Wang, Guoli Jia, Xialei Liu, Liang Lin 等NeurIPS 2025
- Knowledge Swapping via Learning and UnlearningMingyu Xing, Lechao Cheng, Shengeng Tang, Yaxiong Wang 等ICML 2025
- Learning Conditional Space-Time Prompt Distributions for Video Class-Incremental LearningXiaohan Zou, Wenchao Ma, Shu ZhaoCVPR 2025
它引用的顶会 Paper18
- Dark Experience for General Continual Learning: a Strong, Simple BaselinePietro Buzzega, Matteo Boschini, Angelo Porrello, Davide Abati 等NeurIPS 2020 · 被引用 1,494 次
- BatchEnsemble: an Alternative Approach to Efficient Ensemble and Lifelong LearningYeming Wen, Dustin Tran, Jimmy BaICLR 2020 · 被引用 569 次
- Gradient Projection Memory for Continual LearningGobinda Saha, Isha Garg, Kaushik RoyICLR 2021 · 被引用 409 次
- Continual Learning of a Mixed Sequence of Similar and Dissimilar TasksZixuan Ke, Bing Liu, Xingchang HuangNeurIPS 2020 · 被引用 173 次
- Overcoming Catastrophic Forgetting in Graph Neural NetworksHuihui Liu, Yiding Yang, Xinchao WangAAAI 2021 · 被引用 166 次
相关 Paper
- Parameter-efficient Continual Learning for Enhancing Plasticity without Forgetting under Limited Model CapacityYitian Chen, Shigeng Zhang, Xuan Liu, Mingming Lu 等CVPR 2026
- Adaptive Plasticity Improvement for Continual LearningYan-Shuo Liang, Wu-Jun LiCVPR 2023
- Sharing Less is More: Lifelong Learning in Deep Networks with Selective Layer TransferSeungwon Lee, Sima Behpour, Eric EatonICML 2021 · 被引用 21 次
- Is Forgetting Less a Good Inductive Bias for Forward Transfer?Jiefeng Chen, Timothy Nguyen, Dilan Görür, Arslan ChaudhryICLR 2023 · 被引用 1 次
- Multi-Domain Multi-Task Rehearsal for Lifelong LearningFan Lyu, Shuai Wang, Wei Feng, Zihan Ye 等AAAI 2021 · 被引用 34 次
