Optimizing Reusable Knowledge for Continual Learning via Metalearning
Julio Hurtado, Alain Raymond-Saez, Alvaro Soto
摘要
When learning tasks over time, artificial neural networks suffer from a problem known as Catastrophic Forgetting (CF). This happens when the weights of a network are overwritten during the training of a new task causing forgetting of old information. To address this issue, we propose MetA Reusable Knowledge or MARK, a new method that fosters weight reusability instead of overwriting when learning a new task. Specifically, MARK keeps a set of shared weights among tasks. We envision these shared weights as a common Knowledge Base (KB) that is not only used to learn new tasks, but also enriched with new knowledge as the model learns new tasks. Key components behind MARK are two-fold. On the one hand, a metalearning approach provides the key mechanism to incrementally enrich the KB with new knowledge and to foster weight reusability among tasks. On the other hand, a set of trainable masks provides the key mechanism to selectively choose from the KB relevant weights to solve each task. By using MARK, we achieve state of the art results in several popular benchmarks, surpassing the best performing methods in terms of average accuracy by over 10% on the 20-Split-MiniImageNet dataset, while achieving almost zero forgetfulness using 55% of the number of parameters. Furthermore, an ablation study provides evidence that, indeed, MARK is learning reusable knowledge that is selectively used by each task.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper9
- Divide and not forget: Ensemble of selectively trained experts in Continual LearningGrzegorz Rypesc, Sebastian Cygert, Valeriya Khan, Tomasz Trzcinski 等ICLR 2024 · 被引用 52 次
- Self-Evolved Dynamic Expansion Model for Task-Free Continual LearningFei Ye, Adrian G. BorsICCV 2023 · 被引用 28 次
- Wasserstein Expansible Variational Autoencoder for Discriminative and Generative Continual LearningFei Ye, Adrian G. BorsICCV 2023 · 被引用 6 次
- Minimax Forward and Backward Learning of Evolving Tasks with Performance GuaranteesVerónica Álvarez, Santiago Mazuelas, José Antonio LozanoNeurIPS 2023 · 被引用 3 次
- Continual Unsupervised Generative Modelling via Online Optimal TransportFei Ye, Adrian G. Bors, Kun ZhangAAAI 2025
它引用的顶会 Paper10
- Continual learning with hypernetworksJohannes von Oswald, Christian Henning, João Sacramento, Benjamin F. GreweICLR 2020 · 被引用 412 次
- Gradient Projection Memory for Continual LearningGobinda Saha, Isha Garg, Kaushik RoyICLR 2021 · 被引用 409 次
- Supermasks in SuperpositionMitchell Wortsman, Vivek Ramanujan, Rosanne Liu, Aniruddha Kembhavi 等NeurIPS 2020 · 被引用 364 次
- Continual Prototype Evolution: Learning Online from Non-Stationary Data StreamsMatthias De Lange, Tinne TuytelaarsICCV 2021 · 被引用 251 次
- Online Learned Continual Compression with Adaptive Quantization ModulesLucas Caccia, Eugene Belilovsky, Massimo Caccia, Joelle PineauICML 2020 · 被引用 95 次
相关 Paper
- CLR: Channel-wise Lightweight Reprogramming for Continual LearningYunhao Ge, Yuecheng Li, Shuo Ni, Jiaping Zhao 等ICCV 2023 · 被引用 16 次
- One Person, One Model, One World: Learning Continual User Representation without ForgettingFajie Yuan, Guoxiao Zhang, Alexandros Karatzoglou, Joemon M. Jose 等SIGIR 2021 · 被引用 52 次
- Calibrating CNNs for Lifelong LearningPravendra Singh, Vinay Kumar Verma, Pratik Mazumder, Lawrence Carin 等NeurIPS 2020 · 被引用 78 次
- A Minimalistic Unified Framework for Incremental Learning across Image Restoration TasksXiaoxuan Gong, Jie MaNeurIPS 2025 · 被引用 2 次
- Layerwise Optimization by Gradient Decomposition for Continual LearningShixiang Tang, Dapeng Chen, Jinguo Zhu, Shijie Yu 等CVPR 2021
