Continual Learning via Local Module Composition
Oleksiy Ostapenko, Pau Rodríguez, Massimo Caccia, Laurent Charlin
摘要
Modularity is a compelling solution to continual learning (CL), the problem of modeling sequences of related tasks. Learning and then composing modules to solve different tasks provides an abstraction to address the principal challenges of CL including catastrophic forgetting, backward and forward transfer across tasks, and sub-linear model growth. We introduce local module composition (LMC), an approach to modular CL where each module is provided a local structural component that estimates a module's relevance to the input. Dynamic module composition is performed layer-wise based on local relevance scores. We demonstrate that agnosticity to task identities (IDs) arises from (local) structural learning that is module-specific as opposed to the task- and/or model-specific as in previous works, making LMC applicable to more CL settings compared to previous works. In addition, LMC also tracks statistics about the input distribution and adds new modules when outlier samples are detected. In the first set of experiments, LMC performs favorably compared to existing methods on the recent Continual Transfer-learning Benchmark without requiring task identities. In another study, we show that the locality of structural learning allows LMC to interpolate to related but unseen tasks (OOD), as well as to compose modular networks trained independently on different task sequences into a third modular network without any fine-tuning. Finally, in search for limitations of LMC we study it on more challenging sequences of 30 and 100 tasks, demonstrating that local module selection becomes much more challenging in presence of a large number of candidate modules. In this setting best performing LMC spawns much fewer modules compared to an oracle based baseline, however, it reaches a lower overall accuracy. The codebase is available under https://github.com/oleksost/LMC.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper24
- New Insights on Reducing Abrupt Representation Change in Online Continual LearningLucas Caccia, Rahaf Aljundi, Nader Asadi, Tinne Tuytelaars 等ICLR 2022 · 被引用 279 次
- Towards Modular LLMs by Building and Reusing a Library of LoRAsOleksiy Ostapenko, Zhan Su, Edoardo M. Ponti, Laurent Charlin 等ICML 2024 · 被引用 70 次
- Dynamically Expandable Graph Convolution for Streaming RecommendationBowei He, Xu He, Yingxue Zhang, Ruiming Tang 等WWW 2023 · 被引用 60 次
- A Model or 603 Exemplars: Towards Memory-Efficient Class-Incremental LearningDa-Wei Zhou, Qi-Wei Wang, Han-Jia Ye, De-Chuan ZhanICLR 2023 · 被引用 46 次
- Multi-Head Adapter Routing for Cross-Task GeneralizationLucas Page-Caccia, Edoardo Maria Ponti, Zhan Su, Matheus Pereira 等NeurIPS 2023 · 被引用 38 次
它引用的顶会 Paper10
- Gradient Starvation: A Learning Proclivity in Neural NetworksMohammad Pezeshki, Sékou-Oumar Kaba, Yoshua Bengio, Aaron C. Courville 等NeurIPS 2021 · 被引用 378 次
- A Neural Dirichlet Process Mixture Model for Task-Free Continual LearningSoochan Lee, Junsoo Ha, Dongsu Zhang, Gunhee KimICLR 2020 · 被引用 238 次
- Efficient Continual Learning with Modular Networks and Task-Driven PriorsTom Veniat, Ludovic Denoyer, Marc'Aurelio RanzatoICLR 2021 · 被引用 110 次
- Continuous Meta-Learning without TasksJames Harrison, Apoorva Sharma, Chelsea Finn, Marco PavoneNeurIPS 2020 · 被引用 86 次
- Learning where to learn: Gradient sparsity in meta and continual learningJohannes von Oswald, Dominic Zhao, Seijin Kobayashi, Simon Schug 等NeurIPS 2021 · 被引用 61 次
相关 Paper
- A Probabilistic Framework for Modular Continual LearningLazar Valkov, Akash Srivastava, Swarat Chaudhuri, Charles SuttonICLR 2024 · 被引用 5 次
- A Combinatorial Perspective on Transfer LearningJianan Wang, Eren Sezener, David Budden, Marcus Hutter 等NeurIPS 2020 · 被引用 9 次
- Continual Sequence Generation with Adaptive Compositional ModulesYanzhe Zhang, Xuezhi Wang, Diyi YangACL 2022 · 被引用 53 次
- Self-Expansion of Pre-trained Models with Mixture of Adapters for Continual LearningHuiyi Wang, Haodong Lu, Lina Yao, Dong GongCVPR 2025
- Self-Composing Policies for Scalable Continual Reinforcement LearningMikel Malagón, Josu Ceberio, José Antonio LozanoICML 2024 · 被引用 13 次
