Sharing Less is More: Lifelong Learning in Deep Networks with Selective Layer Transfer
Seungwon Lee, Sima Behpour, Eric Eaton
Abstract
Effective lifelong learning across diverse tasks requires the transfer of diverse knowledge, yet transferring irrelevant knowledge may lead to interference and catastrophic forgetting. In deep networks, transferring the appropriate granularity of knowledge is as important as the transfer mechanism, and must be driven by the relationships among tasks. We first show that the lifelong learning performance of several current deep learning architectures can be significantly improved by transfer at the appropriate layers. We then develop an expectation-maximization (EM) method to automatically select the appropriate transfer configuration and optimize the task network weights. This EM-based selective transfer is highly effective, balancing transfer performance on all tasks with avoiding catastrophic forgetting, as demonstrated on three algorithms in several lifelong object classification scenarios.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers2
- Continual Object Detection via Prototypical Task Correlation Guided Gating MechanismBinbin Yang, Xinchi Deng, Han Shi, Changlin Li et al.CVPR 2022 · 39 citations
- Heterogeneous Continual LearningDivyam Madaan, Hongxu Yin, Wonmin Byeon, Jan Kautz et al.CVPR 2023
Builds on4
- Scalable and Order-robust Continual Learning with Additive Parameter DecompositionJaehong Yoon, Saehoon Kim, Eunho Yang, Sung Ju HwangICLR 2020 · 206 citations
- Lifelong Learning of Compositional StructuresJorge A. Mendez, Eric EatonICLR 2021 · 50 citations
- Incremental Multi-Domain Learning with Network Latent Tensor FactorizationAdrian Bulat, Jean Kossaifi, Georgios Tzimiropoulos, Maja PanticAAAI 2020 · 33 citations
- Conditional Channel Gated Networks for Task-Aware Continual LearningDavide Abati, Jakub M. Tomczak, Tijmen Blankevoort, Simone Calderara et al.CVPR 2020
Related papers
- Continual Learning with Adaptive Weights (CLAW)Tameem Adel, Han Zhao, Richard E. TurnerICLR 2020 · 79 citations
- Continual Learning in the Teacher-Student Setup: Impact of Task SimilaritySebastian Lee, Sebastian Goldt, Andrew M. SaxeICML 2021 · 98 citations
- Layerwise Optimization by Gradient Decomposition for Continual LearningShixiang Tang, Dapeng Chen, Jinguo Zhu, Shijie Yu et al.CVPR 2021
- Lifelong GAN: Continual Learning for Conditional Image GenerationMengyao Zhai, Lei Chen, Frederick Tung, Jiawei He et al.ICCV 2019 · 204 citations
- Detect, Decide, Unlearn: A Transfer-Aware Framework for Continual LearningYiwen Wang, Diana Benavides-Prado, Yun Sing KohICLR 2026
