Learning Bayesian Sparse Networks with Full Experience Replay for Continual Learning
Qingsen Yan, Dong Gong, Yuhang Liu, Anton van den Hengel, Javen Qinfeng Shi
摘要
Continual Learning (CL) methods aim to enable machine learning models to learn new tasks without catastrophic forgetting of those that have been previously mastered. Existing CL approaches often keep a buffer of previously-seen samples, perform knowledge distillation, or use regularization techniques towards this goal. Despite their performance, they still suffer from interference across tasks which leads to catastrophic forgetting. To ameliorate this problem, we propose to only activate and select sparse neurons for learning current and past tasks at any stage. More parameters space and model capacity can thus be reserved for the future tasks. This minimizes the interference between parameters for different tasks. To do so, we propose a Sparse neural Network for Continual Learning (SNCL), which employs variational Bayesian sparsity priors on the activations of the neurons in all layers. Full Experience Replay (FER) provides effective supervision in learning the sparse activations of the neurons in different layers. A loss-aware reservoir-sampling strategy is developed to maintain the memory buffer. The proposed method is agnostic as to the network structures and the task boundaries. Experiments on different datasets show that SNCL achieves state-of-the-art result for mitigating forgetting. * † indicates equal contribution. D. Gong
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper9
- Task-Free Dynamic Sparse Vision Transformer for Continual LearningFei Ye, Adrian G. BorsAAAI 2024 · 被引用 7 次
- COPAL: Continual Pruning in Large Language Generative ModelsSrikanth Malla, Joon Hee Choi, Chiho ChoiICML 2024 · 被引用 6 次
- Model Inversion with Layer-Specific Modeling and Alignment for Data-Free Continual LearningRuilin Tong, Haodong Lu, Yuhang Liu, Dong GongNeurIPS 2025 · 被引用 6 次
- Socialized Learning: Making Each Other Better Through Multi-Agent CollaborationXinjie Yao, Yu Wang, Pengfei Zhu, Wanyu Lin 等ICML 2024 · 被引用 5 次
- Lifelong Compression Mixture Model via Knowledge Relationship GraphFei Ye, Adrian G. BorsAAAI 2023 · 被引用 2 次
它引用的顶会 Paper4
- Dark Experience for General Continual Learning: a Strong, Simple BaselinePietro Buzzega, Matteo Boschini, Angelo Porrello, Davide Abati 等NeurIPS 2020 · 被引用 1,494 次
- Using Hindsight to Anchor Past Knowledge in Continual LearningArslan Chaudhry, Albert Gordo, Puneet K. Dokania, Philip H. S. Torr 等AAAI 2021 · 被引用 279 次
- Conditional Channel Gated Networks for Task-Aware Continual LearningDavide Abati, Jakub M. Tomczak, Tijmen Blankevoort, Simone Calderara 等CVPR 2020
- Semantic Drift Compensation for Class-Incremental LearningLu Yu, Bartlomiej Twardowski, Xialei Liu, Luis Herranz 等CVPR 2020
相关 Paper
- Task-aware Orthogonal Sparse Network for Exploring Shared Knowledge in Continual LearningYusong Hu, De Cheng, Dingwen Zhang, Nannan Wang 等ICML 2024 · 被引用 14 次
- Accurate Forgetting for Heterogeneous Federated Continual LearningAbudukelimu Wuerkaixi, Sen Cui, Jingfeng Zhang, Kunda Yan 等ICLR 2024 · 被引用 25 次
- Residual Continual LearningJanghyeon Lee, Donggyu Joo, Hyeong Gwon Hong, Junmo KimAAAI 2020 · 被引用 25 次
- Sketch-Based Replay Projection for Continual LearningJack Julian, Yun Sing Koh, Albert BifetKDD 2024 · 被引用 2 次
- TriRE: A Multi-Mechanism Learning Paradigm for Continual Knowledge Retention and PromotionPreetha Vijayan, Prashant Shivaram Bhat, Bahram Zonooz, Elahe AraniNeurIPS 2023 · 被引用 8 次
