Elastic Weight Consolidation Done Right for Continual Learning
Xuan Liu, Xiaobin Chang
摘要
Weight regularization methods in continual learning (CL) alleviate catastrophic forgetting by assessing and penalizing changes to important model weights. Elastic Weight Consolidation (EWC) is a foundational and widely used approach within this framework that estimates weight importance based on gradients.However, it has consistently shown suboptimal performance.In this paper, we conduct a systematic analysis of importance estimation in EWC from a gradient-based perspective.For the first time, we find that EWC’s reliance on the Fisher Information Matrix (FIM) results in gradient vanishing and inaccurate importance estimation in certain scenarios.Our analysis also reveals that Memory Aware Synapses (MAS), a variant of EWC, imposes unnecessary constraints on parameters irrelevant to prior tasks, termed the redundant protection.Consequently, both EWC and its variant exhibit fundamental misalignments in estimating the importance of weights, leading to inferior performance.To tackle these issues, we propose the Logits Reversal (LR) operation, a simple yet effective modification that rectifies the importance estimation of EWC.Specifically, reversing the logit values during the calculation of the FIM can effectively prevent both the gradient vanishing and the redundant protection.Extensive experiments across various CL tasks and datasets show that the proposed method significantly outperforms existing EWC and its variants. Therefore, we refer to it as EWC Done Right (EWC-DR).
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper21
- ViLT: Vision-and-Language Transformer Without Convolution or Region SupervisionWonjae Kim, Bokyung Son, Ildoo KimICML 2021 · 被引用 2,258 次
- Gradient Projection Memory for Continual LearningGobinda Saha, Isha Garg, Kaushik RoyICLR 2021 · 被引用 409 次
- Class-Incremental Learning via Dual AugmentationFei Zhu, Zhen Cheng, Xu-Yao Zhang, Cheng-Lin LiuNeurIPS 2021 · 被引用 256 次
- Always Be Dreaming: A New Approach for Data-Free Class-Incremental LearningJames Seale Smith, Yen-Chang Hsu, Jonathan C. Balloch, Yilin Shen 等ICCV 2021 · 被引用 208 次
- Self-Sustaining Representation Expansion for Non-Exemplar Class-Incremental LearningKai Zhu, Wei Zhai, Yang Cao, Jiebo Luo 等CVPR 2022 · 被引用 155 次
相关 Paper
- Continual learning in recurrent neural networksBenjamin Ehret, Christian Henning, Maria R. Cervera, Alexander Meulemans 等ICLR 2021 · 被引用 4,433 次
- Revisiting Weight Regularization for Low-Rank Continual LearningYaoyue Zheng, Yin Zhang, Joost van de Weijer, Gido van de Ven 等ICLR 2026 · 被引用 7 次
- Artificial Neuronal Ensembles with Learned Context Dependent GatingMatthew J. Tilley, Michelle Miller, David FreedmanICLR 2023 · 被引用 2 次
- Sliced Cramer Synaptic Consolidation for Preserving Deeply Learned RepresentationsSoheil Kolouri, Nicholas A. Ketz, Andrea Soltoggio, Praveen K. PillyICLR 2020 · 被引用 42 次
- Avoid Catastrophic Forgetting with Rank-1 Fisher from Diffusion ModelsZekun Wang, Anant Gupta, Zihan Dong, Christopher J. MacLellanICLR 2026 · 被引用 4 次
