Towards Continual Learning Desiderata via HSIC-Bottleneck Orthogonalization and Equiangular Embedding
Depeng Li, Tianqi Wang, Junwei Chen, Qining Ren, Kenji Kawaguchi, Zhigang Zeng
Abstract
Deep neural networks are susceptible to catastrophic forgetting when trained on sequential tasks. Various continual learning (CL) methods often rely on exemplar buffers or/and network expansion for balancing model stability and plasticity, which, however, compromises their practical value due to privacy and memory concerns. Instead, this paper considers a strict yet realistic setting, where the training data from previous tasks is unavailable and the model size remains relatively constant during sequential training. To achieve such desiderata, we propose a conceptually simple yet effective method that attributes forgetting to layer-wise parameter overwriting and the resulting decision boundary distortion. This is achieved by the synergy between two key components: HSIC-Bottleneck Orthogonalization (HBO) implements non-overwritten parameter updates mediated by Hilbert-Schmidt independence criterion in an orthogonal space and EquiAngular Embedding (EAE) enhances decision boundary adaptation between old and new tasks with predefined basis vectors. Extensive experiments demonstrate that our method achieves competitive accuracy performance, even with absolute superiority of zero exemplar buffer and 1.02x the base model.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext bfd47543-3f70-4e9a-8735-10f96c549803Cited by top-tier papers5
- What Matters in Graph Class Incremental Learning? An Information Preservation PerspectiveJialu Li, Yu Wang, Pengfei Zhu, Wanyu Lin et al.NeurIPS 2024 · 14 citations
- Training Consistent Mixture-of-Experts-Based Prompt Generator for Continual LearningYue Lu, Shizhou Zhang, De Cheng, Guoqiang Liang et al.AAAI 2025 · 8 citations
- Harnessing Neural Unit Dynamics for Effective and Scalable Class-Incremental LearningDepeng Li, Tianqi Wang, Junwei Chen, Wei Dai et al.ICML 2024 · 5 citations
- Multi-Synaptic Cooperation: A Bio-Inspired Framework for Robust and Scalable Continual LearningPenghui Li, Zhuang Ma, Yunliang Zang, Qiang YuICLR 2026
- Dynamic Integration of Task-Specific Adapters for Class Incremental LearningJiashuo Li, Shaokun Wang, Bo Qian, Yuhang He et al.CVPR 2025
Builds on34
- The Many Faces of Robustness: A Critical Analysis of Out-of-Distribution GeneralizationDan Hendrycks, Steven Basart, Norman Mu, Saurav Kadavath et al.ICCV 2021 · 2,294 citations
- Dark Experience for General Continual Learning: a Strong, Simple BaselinePietro Buzzega, Matteo Boschini, Angelo Porrello, Davide Abati et al.NeurIPS 2020 · 1,494 citations
- Learning to Prompt for Continual LearningZifeng Wang, Zizhao Zhang, Chen-Yu Lee, Han Zhang et al.CVPR 2022 · 635 citations
- Gradient Projection Memory for Continual LearningGobinda Saha, Isha Garg, Kaushik RoyICLR 2021 · 409 citations
- Co2L: Contrastive Continual LearningHyuntak Cha, Jaeho Lee, Jinwoo ShinICCV 2021 · 391 citations
Related papers
- DualHSIC: HSIC-Bottleneck and Alignment for Continual LearningZifeng Wang, Zheng Zhan, Yifan Gong, Yucai Shao et al.ICML 2023 · 21 citations
- Prototype-Sample Relation Distillation: Towards Replay-Free Continual LearningNader Asadi, MohammadReza Davari, Sudhir P. Mudur, Rahaf Aljundi et al.ICML 2023 · 61 citations
- Discrete Key-Value BottleneckFrederik Träuble, Anirudh Goyal, Nasim Rahaman, Michael Curtis Mozer et al.ICML 2023 · 25 citations
- STAR: Stability-Inducing Weight Perturbation for Continual LearningMasih Eskandar, Tooba Imtiaz, Davin Hill, Zifeng Wang et al.ICLR 2025
- Heads collapse, features stay: Why Replay needs big buffersGiulia Lanzillotta, Damiano Meier, Thomas HofmannICLR 2026 · 3 citations
