Efficient Feature Transformations for Discriminative and Generative Continual Learning
Vinay Kumar Verma, Kevin J. Liang, Nikhil Mehta, Piyush Rai, Lawrence Carin
摘要
As neural networks are increasingly being applied to real-world applications, mechanisms to address distributional shift and sequential task learning without forgetting are critical. Methods incorporating network expansion have shown promise by naturally adding model capacity for learning new tasks while simultaneously avoiding catastrophic forgetting. However, the growth in the number of additional parameters of many of these types of methods can be computationally expensive at larger scales, at times prohibitively so. Instead, we propose a simple task-specific feature map transformation strategy for continual learning, which we call Efficient Feature Transformations (EFTs). These EFTs provide powerful flexibility for learning new tasks, achieved with minimal parameters added to the base architecture. We further propose a feature distance maximization strategy, which significantly improves task prediction in class incremental settings, without needing expensive generative models. We demonstrate the efficacy and efficiency of our method with an extensive set of experiments in discriminative (CIFAR-100 and ImageNet-1K) and generative (LSUN, CUB-200, Cats) sequences of tasks. Even with low single-digit parameter growth rates, EFTs can outperform many other continual learning methods in a wide range of settings.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper23
- Representation Compensation Networks for Continual Semantic SegmentationChang-Bin Zhang, Jia-Wen Xiao, Xialei Liu, Ying-Cong Chen 等CVPR 2022 · 被引用 102 次
- Posterior Meta-Replay for Continual LearningChristian Henning, Maria R. Cervera, Francesco D'Angelo, Johannes von Oswald 等NeurIPS 2021 · 被引用 78 次
- General Incremental Learning with Domain-aware Categorical RepresentationsJiangwei Xie, Shipeng Yan, Xuming HeCVPR 2022 · 被引用 37 次
- Discrete Key-Value BottleneckFrederik Träuble, Anirudh Goyal, Nasim Rahaman, Michael Curtis Mozer 等ICML 2023 · 被引用 25 次
- CAM-GAN: Continual Adaptation Modules for Generative Adversarial NetworksSakshi Varshney, Vinay Kumar Verma, P. K. Srijith, Lawrence Carin 等NeurIPS 2021 · 被引用 25 次
它引用的顶会 Paper10
- BatchEnsemble: an Alternative Approach to Efficient Ensemble and Lifelong LearningYeming Wen, Dustin Tran, Jimmy BaICLR 2020 · 被引用 569 次
- Continual learning with hypernetworksJohannes von Oswald, Christian Henning, João Sacramento, Benjamin F. GreweICLR 2020 · 被引用 412 次
- Supermasks in SuperpositionMitchell Wortsman, Vivek Ramanujan, Rosanne Liu, Aniruddha Kembhavi 等NeurIPS 2020 · 被引用 364 次
- A Neural Dirichlet Process Mixture Model for Task-Free Continual LearningSoochan Lee, Junsoo Ha, Dongsu Zhang, Gunhee KimICLR 2020 · 被引用 238 次
- Scalable and Order-robust Continual Learning with Additive Parameter DecompositionJaehong Yoon, Saehoon Kim, Eunho Yang, Sung Ju HwangICLR 2020 · 被引用 206 次
相关 Paper
- On Generalizing Beyond Domains in Cross-Domain Continual LearningChristian Simon, Masoud Faraki, Yi-Hsuan Tsai, Xiang Yu 等CVPR 2022 · 被引用 34 次
- Incremental Learning Using Conditional Adversarial NetworksYe Xiang, Ying Fu, Pan Ji, Hua HuangICCV 2019 · 被引用 188 次
- Effect of scale on catastrophic forgetting in neural networksVinay Venkatesh Ramasesh, Aitor Lewkowycz, Ethan DyerICLR 2022 · 被引用 212 次
- Contextual Transformation Networks for Online Continual LearningQuang Pham, Chenghao Liu, Doyen Sahoo, Steven C. H. HoiICLR 2021 · 被引用 60 次
- DyTox: Transformers for Continual Learning with DYnamic TOken eXpansionArthur Douillard, Alexandre Ramé, Guillaume Couairon, Matthieu CordCVPR 2022 · 被引用 315 次
