Towards Continual Learning Desiderata via HSIC-Bottleneck Orthogonalization and Equiangular Embedding
Depeng Li, Tianqi Wang, Junwei Chen, Qining Ren, Kenji Kawaguchi, Zhigang Zeng
摘要
Deep neural networks are susceptible to catastrophic forgetting when trained on sequential tasks. Various continual learning (CL) methods often rely on exemplar buffers or/and network expansion for balancing model stability and plasticity, which, however, compromises their practical value due to privacy and memory concerns. Instead, this paper considers a strict yet realistic setting, where the training data from previous tasks is unavailable and the model size remains relatively constant during sequential training. To achieve such desiderata, we propose a conceptually simple yet effective method that attributes forgetting to layer-wise parameter overwriting and the resulting decision boundary distortion. This is achieved by the synergy between two key components: HSIC-Bottleneck Orthogonalization (HBO) implements non-overwritten parameter updates mediated by Hilbert-Schmidt independence criterion in an orthogonal space and EquiAngular Embedding (EAE) enhances decision boundary adaptation between old and new tasks with predefined basis vectors. Extensive experiments demonstrate that our method achieves competitive accuracy performance, even with absolute superiority of zero exemplar buffer and 1.02x the base model.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- What Matters in Graph Class Incremental Learning? An Information Preservation PerspectiveJialu Li, Yu Wang, Pengfei Zhu, Wanyu Lin 等NeurIPS 2024 · 被引用 14 次
- Training Consistent Mixture-of-Experts-Based Prompt Generator for Continual LearningYue Lu, Shizhou Zhang, De Cheng, Guoqiang Liang 等AAAI 2025 · 被引用 8 次
- Harnessing Neural Unit Dynamics for Effective and Scalable Class-Incremental LearningDepeng Li, Tianqi Wang, Junwei Chen, Wei Dai 等ICML 2024 · 被引用 5 次
- Multi-Synaptic Cooperation: A Bio-Inspired Framework for Robust and Scalable Continual LearningPenghui Li, Zhuang Ma, Yunliang Zang, Qiang YuICLR 2026
- Dynamic Integration of Task-Specific Adapters for Class Incremental LearningJiashuo Li, Shaokun Wang, Bo Qian, Yuhang He 等CVPR 2025
它引用的顶会 Paper34
- The Many Faces of Robustness: A Critical Analysis of Out-of-Distribution GeneralizationDan Hendrycks, Steven Basart, Norman Mu, Saurav Kadavath 等ICCV 2021 · 被引用 2,294 次
- Dark Experience for General Continual Learning: a Strong, Simple BaselinePietro Buzzega, Matteo Boschini, Angelo Porrello, Davide Abati 等NeurIPS 2020 · 被引用 1,494 次
- Learning to Prompt for Continual LearningZifeng Wang, Zizhao Zhang, Chen-Yu Lee, Han Zhang 等CVPR 2022 · 被引用 635 次
- Gradient Projection Memory for Continual LearningGobinda Saha, Isha Garg, Kaushik RoyICLR 2021 · 被引用 409 次
- Co2L: Contrastive Continual LearningHyuntak Cha, Jaeho Lee, Jinwoo ShinICCV 2021 · 被引用 391 次
相关 Paper
- DualHSIC: HSIC-Bottleneck and Alignment for Continual LearningZifeng Wang, Zheng Zhan, Yifan Gong, Yucai Shao 等ICML 2023 · 被引用 21 次
- Prototype-Sample Relation Distillation: Towards Replay-Free Continual LearningNader Asadi, MohammadReza Davari, Sudhir P. Mudur, Rahaf Aljundi 等ICML 2023 · 被引用 61 次
- Discrete Key-Value BottleneckFrederik Träuble, Anirudh Goyal, Nasim Rahaman, Michael Curtis Mozer 等ICML 2023 · 被引用 25 次
- STAR: Stability-Inducing Weight Perturbation for Continual LearningMasih Eskandar, Tooba Imtiaz, Davin Hill, Zifeng Wang 等ICLR 2025
- Heads collapse, features stay: Why Replay needs big buffersGiulia Lanzillotta, Damiano Meier, Thomas HofmannICLR 2026 · 被引用 3 次
