Efficient Continual Learning with Modular Networks and Task-Driven Priors
Tom Veniat, Ludovic Denoyer, Marc'Aurelio Ranzato
摘要
Existing literature in Continual Learning (CL) has focused on overcoming catastrophic forgetting, the inability of the learner to recall how to perform tasks observed in the past. There are however other desirable properties of a CL system, such as the ability to transfer knowledge from previous tasks and to scale memory and compute sub-linearly with the number of tasks. Since most current benchmarks focus only on forgetting using short streams of tasks, we first propose a new suite of benchmarks to probe CL algorithms across these new axes. Finally, we introduce a new modular architecture, whose modules represent atomic skills that can be composed to perform a certain task. Learning a task reduces to figuring out which past modules to re-use, and which new modules to instantiate to solve the current task. Our learning algorithm leverages a task-driven prior over the exponential search space of all possible ways to combine modules, enabling efficient learning on long streams of tasks. Our experiments show that this modular architecture and learning algorithm perform competitively on widely used CL benchmarks while yielding superior performance on the more challenging benchmarks we introduce in this work.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper29
- Learning to Prompt for Continual LearningZifeng Wang, Zizhao Zhang, Chen-Yu Lee, Han Zhang 等CVPR 2022 · 被引用 635 次
- Remember the Past: Distilling Datasets into Addressable Memories for Neural NetworksZhiwei Deng, Olga RussakovskyNeurIPS 2022 · 被引用 140 次
- RMM: Reinforced Memory Management for Class-Incremental LearningYaoyao Liu, Bernt Schiele, Qianru SunNeurIPS 2021 · 被引用 125 次
- TRGP: Trust Region Gradient Projection for Continual LearningSen Lin, Li Yang, Deliang Fan, Junshan ZhangICLR 2022 · 被引用 107 次
- Continual Learning via Local Module CompositionOleksiy Ostapenko, Pau Rodríguez, Massimo Caccia, Laurent CharlinNeurIPS 2021 · 被引用 98 次
它引用的顶会 Paper1
相关 Paper
- A Probabilistic Framework for Modular Continual LearningLazar Valkov, Akash Srivastava, Swarat Chaudhuri, Charles SuttonICLR 2024 · 被引用 5 次
- A Combinatorial Perspective on Transfer LearningJianan Wang, Eren Sezener, David Budden, Marcus Hutter 等NeurIPS 2020 · 被引用 9 次
- Continual World: A Robotic Benchmark For Continual Reinforcement LearningMaciej Wolczyk, Michal Zajac, Razvan Pascanu, Lukasz Kucinski 等NeurIPS 2021 · 被引用 152 次
- Is Forgetting Less a Good Inductive Bias for Forward Transfer?Jiefeng Chen, Timothy Nguyen, Dilan Görür, Arslan ChaudhryICLR 2023 · 被引用 1 次
- Progressive Prompts: Continual Learning for Language ModelsAnastasia Razdaibiedina, Yuning Mao, Rui Hou, Madian Khabsa 等ICLR 2023 · 被引用 15 次
