Learning Bayesian Sparse Networks with Full Experience Replay for Continual Learning
Qingsen Yan, Dong Gong, Yuhang Liu, Anton van den Hengel, Javen Qinfeng Shi
Abstract
Continual Learning (CL) methods aim to enable machine learning models to learn new tasks without catastrophic forgetting of those that have been previously mastered. Existing CL approaches often keep a buffer of previously-seen samples, perform knowledge distillation, or use regularization techniques towards this goal. Despite their performance, they still suffer from interference across tasks which leads to catastrophic forgetting. To ameliorate this problem, we propose to only activate and select sparse neurons for learning current and past tasks at any stage. More parameters space and model capacity can thus be reserved for the future tasks. This minimizes the interference between parameters for different tasks. To do so, we propose a Sparse neural Network for Continual Learning (SNCL), which employs variational Bayesian sparsity priors on the activations of the neurons in all layers. Full Experience Replay (FER) provides effective supervision in learning the sparse activations of the neurons in different layers. A loss-aware reservoir-sampling strategy is developed to maintain the memory buffer. The proposed method is agnostic as to the network structures and the task boundaries. Experiments on different datasets show that SNCL achieves state-of-the-art result for mitigating forgetting. * † indicates equal contribution. D. Gong
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 4a5ce84f-592d-4066-a9a0-0eef26e7bdffCited by top-tier papers9
- Task-Free Dynamic Sparse Vision Transformer for Continual LearningFei Ye, Adrian G. BorsAAAI 2024 · 7 citations
- COPAL: Continual Pruning in Large Language Generative ModelsSrikanth Malla, Joon Hee Choi, Chiho ChoiICML 2024 · 6 citations
- Model Inversion with Layer-Specific Modeling and Alignment for Data-Free Continual LearningRuilin Tong, Haodong Lu, Yuhang Liu, Dong GongNeurIPS 2025 · 6 citations
- Socialized Learning: Making Each Other Better Through Multi-Agent CollaborationXinjie Yao, Yu Wang, Pengfei Zhu, Wanyu Lin et al.ICML 2024 · 5 citations
- Lifelong Compression Mixture Model via Knowledge Relationship GraphFei Ye, Adrian G. BorsAAAI 2023 · 2 citations
Builds on4
- Dark Experience for General Continual Learning: a Strong, Simple BaselinePietro Buzzega, Matteo Boschini, Angelo Porrello, Davide Abati et al.NeurIPS 2020 · 1,494 citations
- Using Hindsight to Anchor Past Knowledge in Continual LearningArslan Chaudhry, Albert Gordo, Puneet K. Dokania, Philip H. S. Torr et al.AAAI 2021 · 279 citations
- Conditional Channel Gated Networks for Task-Aware Continual LearningDavide Abati, Jakub M. Tomczak, Tijmen Blankevoort, Simone Calderara et al.CVPR 2020
- Semantic Drift Compensation for Class-Incremental LearningLu Yu, Bartlomiej Twardowski, Xialei Liu, Luis Herranz et al.CVPR 2020
Related papers
- Task-aware Orthogonal Sparse Network for Exploring Shared Knowledge in Continual LearningYusong Hu, De Cheng, Dingwen Zhang, Nannan Wang et al.ICML 2024 · 14 citations
- Accurate Forgetting for Heterogeneous Federated Continual LearningAbudukelimu Wuerkaixi, Sen Cui, Jingfeng Zhang, Kunda Yan et al.ICLR 2024 · 25 citations
- Residual Continual LearningJanghyeon Lee, Donggyu Joo, Hyeong Gwon Hong, Junmo KimAAAI 2020 · 25 citations
- Sketch-Based Replay Projection for Continual LearningJack Julian, Yun Sing Koh, Albert BifetKDD 2024 · 2 citations
- TriRE: A Multi-Mechanism Learning Paradigm for Continual Knowledge Retention and PromotionPreetha Vijayan, Prashant Shivaram Bhat, Bahram Zonooz, Elahe AraniNeurIPS 2023 · 8 citations
