Online Continual Learning through Mutual Information Maximization
Yiduo Guo, Bing Liu, Dongyan Zhao
Abstract
This paper proposes a new online continual learning technique called OCM based on mutual information maximization. It achieves two objectives that are critical in dealing with catastrophic forgetting (CF). (1) It reduces feature bias caused by cross entropy (CE) as CE learns only discriminative features for each task, but these features may not be discriminative for another task. To learn a new task well, the network parameters learned before have to be modified, which causes CF. The new approach encourages the learning of each task to make use of holistic representations or the full features of the task training data. (2) It encourages preservation of the previously learned knowledge when training a new batch of incrementally arriving data. Empirical evaluation shows that OCM substantially outperforms the online CL baselines. For example, for CIFAR10, OCM improves the accuracy of the best baseline by 13.1% from 64.1% (baseline) to 77.2% (OCM). The code is publicly available at https://github.com/gydpku/OCM .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext e06b6473-dd21-4e49-8d0d-2621a695b4b1Cited by top-tier papers57
- A Theoretical Study on Solving Continual LearningGyuhak Kim, Changnan Xiao, Tatsuya Konishi, Zixuan Ke et al.NeurIPS 2022 · 119 citations
- Online Prototype Learning for Online Continual LearningYujie Wei, Jiaxin Ye, Zhizhong Huang, Junping Zhang et al.ICCV 2023 · 78 citations
- A Unified Replay-Based Continuous Learning Framework for Spatio-Temporal Prediction on Streaming DataHao Miao, Yan Zhao, Chenjuan Guo, Bin Yang et al.ICDE 2024 · 64 citations
- A Unified Approach to Domain Incremental Learning with Memory: Theory and AlgorithmHaizhou Shi, Hao WangNeurIPS 2023 · 60 citations
- Meta Continual Learning Revisited: Implicitly Enhancing Online Hessian Approximation via Variance ReductionYichen Wu, Long-Kai Huang, Renzhen Wang, Deyu Meng et al.ICLR 2024 · 42 citations
Builds on20
- CutMix: Regularization Strategy to Train Strong Classifiers With Localizable FeaturesSangdoo Yun, Dongyoon Han, Sanghyuk Chun, Seong Joon Oh et al.ICCV 2019 · 5,843 citations
- Continual learning with hypernetworksJohannes von Oswald, Christian Henning, João Sacramento, Benjamin F. GreweICLR 2020 · 412 citations
- Gradient Projection Memory for Continual LearningGobinda Saha, Isha Garg, Kaushik RoyICLR 2021 · 409 citations
- Co2L: Contrastive Continual LearningHyuntak Cha, Jaeho Lee, Jinwoo ShinICCV 2021 · 391 citations
- Supermasks in SuperpositionMitchell Wortsman, Vivek Ramanujan, Rosanne Liu, Aniruddha Kembhavi et al.NeurIPS 2020 · 364 citations
Related papers
- Continual Learning by Using Information of Each Class HolisticallyWenpeng Hu, Qi Qin, Mengyu Wang, Jinwen Ma et al.AAAI 2021 · 64 citations
- Continual Normalization: Rethinking Batch Normalization for Online Continual LearningQuang Pham, Chenghao Liu, Steven C. H. HoiICLR 2022 · 72 citations
- Not Just Selection, but Exploration: Online Class-Incremental Continual Learning via Dual View ConsistencyYanan Gu, Xu Yang, Kun Wei, Cheng DengCVPR 2022 · 69 citations
- Parameter-Level Soft-Masking for Continual LearningTatsuya Konishi, Mori Kurokawa, Chihiro Ono, Zixuan Ke et al.ICML 2023 · 63 citations
- Improving Plasticity in Online Continual Learning via Collaborative LearningMaorong Wang, Nicolas Michel, Ling Xiao, Toshihiko YamasakiCVPR 2024 · 7 citations
