Continual Normalization: Rethinking Batch Normalization for Online Continual Learning
Quang Pham, Chenghao Liu, Steven C. H. Hoi
摘要
Existing continual learning methods use Batch Normalization (BN) to facilitate training and improve generalization across tasks. However, the non-i.i.d and non-stationary nature of continual learning data, especially in the online setting, amplify the discrepancy between training and testing in BN and hinder the performance of older tasks. In this work, we study the cross-task normalization effect of BN in online continual learning where BN normalizes the testing data using moments biased towards the current task, resulting in higher catastrophic forgetting. This limitation motivates us to propose a simple yet effective method that we call Continual Normalization (CN) to facilitate training similar to BN while mitigating its negative effect. Extensive experiments on different continual learning algorithms and online scenarios show that CN is a direct replacement for BN and can provide substantial performance improvements. Our implementation is available at https://github.com/phquang/Continual-Normalization.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper15
- CrossQ: Batch Normalization in Deep Reinforcement Learning for Greater Sample Efficiency and SimplicityAditya Bhatt, Daniel Palenicek, Boris Belousov, Max Argus 等ICLR 2024 · 被引用 106 次
- MOS: Model Surgery for Pre-Trained Model-Based Class-Incremental LearningHai-Long Sun, Da-Wei Zhou, Hanbin Zhao, Le Gan 等AAAI 2025 · 被引用 31 次
- Summarizing Stream Data for Memory-Constrained Online Continual LearningJianyang Gu, Kai Wang, Wei Jiang, Yang YouAAAI 2024 · 被引用 30 次
- Does Continual Learning Equally Forget All Parameters?Haiyan Zhao, Tianyi Zhou, Guodong Long, Jing Jiang 等ICML 2023 · 被引用 21 次
- Overcoming Recency Bias of Normalization Statistics in Continual Learning: Balance and AdaptationYilin Lyu, Liyuan Wang, Xingxing Zhang, Zicheng Sun 等NeurIPS 2023 · 被引用 17 次
它引用的顶会 Paper7
- Dark Experience for General Continual Learning: a Strong, Simple BaselinePietro Buzzega, Matteo Boschini, Angelo Porrello, Davide Abati 等NeurIPS 2020 · 被引用 1,494 次
- Continual learning with hypernetworksJohannes von Oswald, Christian Henning, João Sacramento, Benjamin F. GreweICLR 2020 · 被引用 412 次
- Anatomy of Catastrophic Forgetting: Hidden Representations and Task SemanticsVinay Venkatesh Ramasesh, Ethan Dyer, Maithra RaghuICLR 2021 · 被引用 207 次
- DualNet: Continual Learning, Fast and SlowQuang Pham, Chenghao Liu, Steven C. H. HoiNeurIPS 2021 · 被引用 192 次
- TaskNorm: Rethinking Batch Normalization for Meta-LearningJohn Bronskill, Jonathan Gordon, James Requeima, Sebastian Nowozin 等ICML 2020 · 被引用 93 次
相关 Paper
- Rebalancing Batch Normalization for Exemplar-Based Class-Incremental LearningSungmin Cha, Sungjun Cho, Dasol Hwang, Sunwon Hong 等CVPR 2023
- Online Continual Learning through Mutual Information MaximizationYiduo Guo, Bing Liu, Dongyan ZhaoICML 2022 · 被引用 139 次
- Residual Continual LearningJanghyeon Lee, Donggyu Joo, Hyeong Gwon Hong, Junmo KimAAAI 2020 · 被引用 25 次
- CBA: Improving Online Continual Learning via Continual Bias AdaptorQuanziang Wang, Renzhen Wang, Yichen Wu, Xixi Jia 等ICCV 2023 · 被引用 25 次
- Mitigating Forgetting in Online Continual Learning with Neuron CalibrationHaiyan Yin, Peng Yang, Ping LiNeurIPS 2021 · 被引用 39 次
