Growing a Brain with Sparsity-Inducing Generation for Continual Learning
Hyundong Jin, Gyeong-Hyeon Kim, Chanho Ahn, Eunwoo Kim
摘要
Deep neural networks suffer from catastrophic forgetting in continual learning, where they tend to lose information about previously learned tasks when optimizing a new incoming task. Recent strategies isolate the important parameters for previous tasks to retain old knowledge while learning the new task. However, using the fixed old knowledge might act as an obstacle to capturing novel representations. To overcome this limitation, we propose a framework that evolves the previously allocated parameters by absorbing the knowledge of the new task. The approach performs under two different networks. The base network learns knowledge of sequential tasks, and the sparsity-inducing hyper-network generates parameters for each time step for evolving old knowledge. The generated parameters transform old parameters of the base network to reflect the new knowledge. We design the hypernetwork to generate sparse parameters conditional to the task-specific information and the structural information of the base network. We evaluate the proposed approach on class-incremental and task-incremental learning scenarios for image classification and video action recognition tasks. Experimental results show that the proposed method consistently outperforms a large variety of continual learning approaches for those scenarios by evolving old knowledge.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- Model Inversion with Layer-Specific Modeling and Alignment for Data-Free Continual LearningRuilin Tong, Haodong Lu, Yuhang Liu, Dong GongNeurIPS 2025 · 被引用 6 次
- RainbowPrompt: Diversity-Enhanced Prompt-Evolving for Continual LearningKiseong Hong, Gyeong-Hyeon Kim, Eunwoo KimICCV 2025 · 被引用 3 次
- Instruction-Grounded Visual Projectors for Continual Learning of Generative Vision-Language ModelsHyundong Jin, Hyung Jin Chang, Eunwoo KimICCV 2025 · 被引用 2 次
- Coreset Selection via Reducible Loss in Continual LearningRuilin Tong, Yuhang Liu, Javen Qinfeng Shi, Dong GongICLR 2025
- Parameter-efficient Continual Learning for Enhancing Plasticity without Forgetting under Limited Model CapacityYitian Chen, Shigeng Zhang, Xuan Liu, Mingming Lu 等CVPR 2026
它引用的顶会 Paper14
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- Continual learning in recurrent neural networksBenjamin Ehret, Christian Henning, Maria R. Cervera, Alexander Meulemans 等ICLR 2021 · 被引用 4,433 次
- Continual learning with hypernetworksJohannes von Oswald, Christian Henning, João Sacramento, Benjamin F. GreweICLR 2020 · 被引用 412 次
- AdaShare: Learning What To Share For Efficient Deep Multi-Task LearningXimeng Sun, Rameswar Panda, Rogério Feris, Kate SaenkoNeurIPS 2020 · 被引用 337 次
- DyTox: Transformers for Continual Learning with DYnamic TOken eXpansionArthur Douillard, Alexandre Ramé, Guillaume Couairon, Matthieu CordCVPR 2022 · 被引用 315 次
相关 Paper
- Conditional Channel Gated Networks for Task-Aware Continual LearningDavide Abati, Jakub M. Tomczak, Tijmen Blankevoort, Simone Calderara 等CVPR 2020
- Layerwise Optimization by Gradient Decomposition for Continual LearningShixiang Tang, Dapeng Chen, Jinguo Zhu, Shijie Yu 等CVPR 2021
- Continual Semantic Segmentation via Repulsion-Attraction of Sparse and Disentangled Latent RepresentationsUmberto Michieli, Pietro ZanuttighCVPR 2021
- Hypercorrelation Evolution for Video Class-Incremental LearningSen Liang, Kai Zhu, Wei Zhai, Zhiheng Liu 等AAAI 2024 · 被引用 4 次
- Learning without Isolation: Pathway Protection for Continual LearningZhikang Chen, Abudukelimu Wuerkaixi, Sen Cui, Haoxuan Li 等ICML 2025
