Continual Learning in Low-rank Orthogonal Subspaces
Arslan Chaudhry, Naeemullah Khan, Puneet K. Dokania, Philip H. S. Torr
Abstract
In continual learning (CL), a learner is faced with a sequence of tasks, arriving one after the other, and the goal is to remember all the tasks once the continual learning experience is finished. The prior art in CL uses episodic memory, parameter regularization or extensible network structures to reduce interference among tasks, but in the end, all the approaches learn different tasks in a joint vector space. We believe this invariably leads to interference among different tasks. We propose to learn tasks in different (low-rank) vector subspaces that are kept orthogonal to each other in order to minimize interference. Further, to keep the gradients of different tasks coming from these subspaces orthogonal to each other, we learn isometric mappings by posing network training as an optimization problem over the Stiefel manifold. To the best of our understanding, we report, for the first time, strong results over experience-replay baseline with and without memory on standard classification benchmarks in continual learning. The code is made publicly available.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext fb700acd-bece-4dbd-948e-d1edffce8548Cited by top-tier papers47
- GaLore: Memory-Efficient LLM Training by Gradient Low-Rank ProjectionJiawei Zhao, Zhenyu Zhang, Beidi Chen, Zhangyang Wang et al.ICML 2024 · 433 citations
- Federated Continual Learning with Weighted Inter-client TransferJaehong Yoon, Wonyong Jeong, Giwoong Lee, Eunho Yang et al.ICML 2021 · 303 citations
- Preservation of the Global Knowledge by Not-True Distillation in Federated LearningGihun Lee, Minchan Jeong, Yongjin Shin, Sangmin Bae et al.NeurIPS 2022 · 235 citations
- Forget-free Continual Learning with Winning SubnetworksHaeyong Kang, Rusty John Lloyd Mina, Sultan Rizky Hikmawan Madjid, Jaehong Yoon et al.ICML 2022 · 159 citations
- Online Continual Learning through Mutual Information MaximizationYiduo Guo, Bing Liu, Dongyan ZhaoICML 2022 · 139 citations
Builds on2
- Using Hindsight to Anchor Past Knowledge in Continual LearningArslan Chaudhry, Albert Gordo, Puneet K. Dokania, Philip H. S. Torr et al.AAAI 2021 · 279 citations
- Efficient Riemannian Optimization on the Stiefel Manifold via the Cayley TransformJun Li, Fuxin Li, Sinisa TodorovicICLR 2020 · 139 citations
Related papers
- Expandable and Differentiable Dual Memories with Orthogonal Regularization for Exemplar-free Continual LearningHyung-Jun Moon, Sung-Bae ChoAAAI 2026
- SplitLoRA: Balancing Stability and Plasticity in Continual Learning Through Gradient Space SplittingHaomiao Qiu, Miao Zhang, Ziyue Qiao, Weili Guan et al.ICLR 2026 · 10 citations
- Continual Learning with Recursive Gradient OptimizationHao Liu, Huaping LiuICLR 2022 · 52 citations
- Continual Learning with Scaled Gradient ProjectionGobinda Saha, Kaushik RoyAAAI 2023 · 44 citations
- Adaptive Orthogonal Projection for Batch and Online Continual LearningYiduo Guo, Wenpeng Hu, Dongyan Zhao, Bing LiuAAAI 2022 · 56 citations
