Rectification-Based Knowledge Retention for Continual Learning
Pravendra Singh, Pratik Mazumder, Piyush Rai, Vinay P. Namboodiri
Abstract
Deep learning models suffer from catastrophic forgetting when trained in an incremental learning setting. In this work, we propose a novel approach to address the task incremental learning problem, which involves training a model on new tasks that arrive in an incremental manner. The task incremental learning problem becomes even more challenging when the test set contains classes that are not part of the train set, i.e., a task incremental generalized zero-shot learning problem. Our approach can be used in both the zero-shot and non zero-shot task incremental learning settings. Our proposed method uses weight rectifications and affine transformations in order to adapt the model to different tasks that arrive sequentially. Specifically, we adapt the network weights to work for new tasks by "rectifying" the weights learned from the previous task. We learn these weight rectifications using very few parameters. We additionally learn affine transformations on the outputs generated by the network in order to better adapt them for the new task. We perform experiments on several datasets in both zero-shot and non zero-shot task incremental learning settings and empirically show that our approach achieves state-of-the-art results. Specifically, our approach outperforms the state-of-the-art non zero-shot task incremental learning method by over 5% on the CIFAR-100 dataset. Our approach also significantly outperforms the state-of-the-art task incremental generalized zero-shot learning method by absolute margins of 6.91% and 6.33% for the AWA1 and CUB datasets, respectively. We validate our approach using various ablation studies.
We propose a novel approach called Rectification-based Knowledge Retention (RKR) for the task incremental learning problem in the zero-shot and non zero-shot setting. Our approach (RKR) learns weight rectifications to adapt the network weights for a new task. After learning these weight rectifications, we can quickly adapt the network to work for images from that task by simply applying these weight rectifications to the network weights. We utilize an efficient technique for learning these weight rectifications to limit the model size. We also learn affine transformations (scaling factors) for all the intermediate outputs of the network that allow better adaptation of the network to the respective 1
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 5aa2e49c-b6c7-460e-86aa-c89335342f46Cited by top-tier papers10
- Representation Compensation Networks for Continual Semantic SegmentationChang-Bin Zhang, Jia-Wen Xiao, Xialei Liu, Ying-Cong Chen et al.CVPR 2022 · 102 citations
- Continual Learning with Lifelong Vision TransformerZhen Wang, Liu Liu, Yiqun Duan, Yajing Kong et al.CVPR 2022 · 63 citations
- DFIL: Deepfake Incremental Learning by Exploiting Domain-invariant Forgery CluesKun Pan, Yifang Yin, Yao Wei, Feng Lin et al.ACM MM 2023 · 35 citations
- Preparing the Future for Continual Semantic SegmentationZihan Lin, Zilei Wang, Yixin ZhangICCV 2023 · 8 citations
- DevFD : Developmental Face Forgery Detection by Learning Shared and Orthogonal LoRA SubspacesTianshuo Zhang, Li Gao, Siran Peng, Xiangyu Zhu et al.NeurIPS 2025 · 4 citations
Builds on4
- Continual learning with hypernetworksJohannes von Oswald, Christian Henning, João Sacramento, Benjamin F. GreweICLR 2020 · 412 citations
- Scalable and Order-robust Continual Learning with Additive Parameter DecompositionJaehong Yoon, Saehoon Kim, Eunho Yang, Sung Ju HwangICLR 2020 · 206 citations
- Calibrating CNNs for Lifelong LearningPravendra Singh, Vinay Kumar Verma, Pratik Mazumder, Lawrence Carin et al.NeurIPS 2020 · 78 citations
- Semantic Drift Compensation for Class-Incremental LearningLu Yu, Bartlomiej Twardowski, Xialei Liu, Luis Herranz et al.CVPR 2020
Related papers
- Class-Incremental Learning by Knowledge Distillation with Adaptive Feature ConsolidationMinsoo Kang, Jaeyoo Park, Bohyung HanCVPR 2022 · 189 citations
- Incremental Embedding Learning via Zero-Shot TranslationKun Wei, Cheng Deng, Xu Yang, Maosen LiAAAI 2021 · 24 citations
- iTAML: An Incremental Task-Agnostic Meta-learning ApproachJathushan Rajasegaran, Salman H. Khan, Munawar Hayat, Fahad Shahbaz Khan et al.CVPR 2020
- Prototype Augmentation and Self-Supervision for Incremental LearningFei Zhu, Xu-Yao Zhang, Chuang Wang, Fei Yin et al.CVPR 2021
- DKT: Diverse Knowledge Transfer Transformer for Class Incremental LearningXinyuan Gao, Yuhang He, Songlin Dong, Jie Cheng et al.CVPR 2023
