SparCL: Sparse Continual Learning on the Edge
Zifeng Wang, Zheng Zhan, Yifan Gong, Geng Yuan, Wei Niu, Tong Jian, Bin Ren, Stratis Ioannidis, Yanzhi Wang, Jennifer G. Dy
Abstract
Existing work in continual learning (CL) focuses on mitigating catastrophic forgetting, i.e., model performance deterioration on past tasks when learning a new task. However, the training efficiency of a CL system is under-investigated, which limits the real-world application of CL systems under resource-limited scenarios. In this work, we propose a novel framework called Sparse Continual Learning (SparCL), which is the first study that leverages sparsity to enable cost-effective continual learning on edge devices. SparCL achieves both training acceleration and accuracy preservation through the synergy of three aspects: weight sparsity, data efficiency, and gradient sparsity. Specifically, we propose task-aware dynamic masking (TDM) to learn a sparse network throughout the entire CL process, dynamic data removal (DDR) to remove less informative training data, and dynamic gradient masking (DGM) to sparsify the gradient updates. Each of them not only improves efficiency, but also further mitigates catastrophic forgetting. SparCL consistently improves the training efficiency of existing state-of-the-art (SOTA) CL methods by at most 23× less training FLOPs, and, surprisingly, further improves the SOTA accuracy by at most 1.7%. SparCL also outperforms competitive baselines obtained from adapting SOTA sparse training methods to the CL setting in both efficiency and accuracy. We also evaluate the effectiveness of SparCL on a real mobile phone, further indicating the practical potential of our method. Source code will be released.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers14
- CLAP4CLIP: Continual Learning with Probabilistic Finetuning for Vision-Language ModelsSaurav Jha, Dong Gong, Lina YaoNeurIPS 2024 · 36 citations
- Continual Learning in the Frequency DomainRuiqi Liu, Boyu Diao, Libo Huang, Zijia An et al.NeurIPS 2024 · 26 citations
- Cost-effective On-device Continual Learning over Memory Hierarchy with MiroXinyue Ma, Suyeon Jeong, Minjia Zhang, Di Wang et al.MobiCom 2023 · 21 citations
- MEMOIR: Lifelong Model Editing with Minimal Overwrite and Informed Retention for LLMsKe Wang, Yiming Qin, Nikolaos Dimitriadis, Alessandro Favero et al.NeurIPS 2025 · 15 citations
- Enabling Real-Time Inference in Online Continual Learning via Device-Cloud CollaborationHaibo Liu, Chen Gong, Zhenzhe Zheng, Shengzhong Liu et al.WWW 2025 · 10 citations
Builds on19
- Dark Experience for General Continual Learning: a Strong, Simple BaselinePietro Buzzega, Matteo Boschini, Angelo Porrello, Davide Abati et al.NeurIPS 2020 · 1,494 citations
- Pruning neural networks without any data by iteratively conserving synaptic flowHidenori Tanaka, Daniel Kunin, Daniel L. K. Yamins, Surya GanguliNeurIPS 2020 · 884 citations
- Picking Winning Tickets Before Training by Preserving Gradient FlowChaoqi Wang, Guodong Zhang, Roger B. GrosseICLR 2020 · 743 citations
- Rigging the Lottery: Making All Tickets WinnersUtku Evci, Trevor Gale, Jacob Menick, Pablo Samuel Castro et al.ICML 2020 · 723 citations
- Learning to Prompt for Continual LearningZifeng Wang, Zizhao Zhang, Chen-Yu Lee, Han Zhang et al.CVPR 2022 · 635 citations
Related papers
- Lethe: Plasticity-aware Active Forgetting for Resource-Efficient On-Device Continual LearningHaibo Liu, Chenxin Mao, Zhenzhe Zheng, Fan Wu et al.KDD 2026
- Critical Patch-Aware Sparse Prompting with Decoupled Training for Continual Learning on the EdgeWonseon Lim, Jaesung Lee, Dae-Won KimCVPR 2026 · 1 citation
- Delta: A Cloud-assisted Data Enrichment Framework for On-Device Continual LearningChen Gong, Zhenzhe Zheng, Fan Wu, Xiaofeng Jia et al.MobiCom 2024 · 6 citations
- CarM: hierarchical episodic memory for continual learningSoobee Lee, Minindu Weerakoon, Jonghyun Choi, Minjia Zhang et al.DAC 2022 · 27 citations
- FedKNOW: Federated Continual Learning with Signature Task Knowledge Integration at EdgeYaxin Luopan, Rui Han, Qinglong Zhang, Chi Harold Liu et al.ICDE 2023 · 31 citations
