Dense Network Expansion for Class Incremental Learning
Zhiyuan Hu, Yunsheng Li, Jiancheng Lyu, Dashan Gao, Nuno Vasconcelos
摘要
The problem of class incremental learning (CIL) is considered. State-of-the-art approaches use a dynamic architecture based on network expansion (NE), in which a task expert is added per task. While effective from a computational standpoint, these methods lead to models that grow quickly with the number of tasks. A new NE method, dense network expansion (DNE), is proposed to achieve a better trade-off between accuracy and model complexity. This is accomplished by the introduction of dense connections between the intermediate layers of the task expert networks, that enable the transfer of knowledge from old to new tasks via feature sharing and reusing. This sharing is implemented with a cross-task attention mechanism, based on a new task attention block (TAB), that fuses information across tasks. Unlike traditional attention mechanisms, TAB operates at the level of the feature mixing and is decoupled with spatial attentions. This is shown more effective than a joint spatial-and-task attention for CIL. The proposed DNE approach can strictly maintain the feature space of old classes while growing the network and feature scale at a much slower rate than previous methods. In result, it outperforms the previous SOTA methods by a margin of 4% in terms of accuracy, with similar or even smaller model scale.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper33
- Boosting Continual Learning of Vision-Language Models via Mixture-of-Experts AdaptersJiazuo Yu, Yunzhi Zhuge, Lu Zhang, Ping Hu 等CVPR 2024 · 被引用 80 次
- Mixture of Noise for Pre-Trained Model-Based Class-Incremental LearningKai Jiang, Zhengyan Shi, Dell Zhang, Hongyuan Zhang 等NeurIPS 2025 · 被引用 38 次
- MOS: Model Surgery for Pre-Trained Model-Based Class-Incremental LearningHai-Long Sun, Da-Wei Zhou, Hanbin Zhao, Le Gan 等AAAI 2025 · 被引用 31 次
- LLMs Can Evolve Continually on Modality for X-Modal ReasoningJiazuo Yu, Haomiao Xiong, Lu Zhang, Haiwen Diao 等NeurIPS 2024 · 被引用 13 次
- CAPrompt: Cyclic Prompt Aggregation for Pre-Trained Model Based Class Incremental LearningQiwei Li, Jiahuan ZhouAAAI 2025 · 被引用 10 次
它引用的顶会 Paper11
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- Per-Pixel Classification is Not All You Need for Semantic SegmentationBowen Cheng, Alexander G. Schwing, Alexander KirillovNeurIPS 2021 · 被引用 2,196 次
- Frozen in Time: A Joint Video and Image Encoder for End-to-End RetrievalMax Bain, Arsha Nagrani, Gül Varol, Andrew ZissermanICCV 2021 · 被引用 1,550 次
- Decoupling Representation and Classifier for Long-Tailed RecognitionBingyi Kang, Saining Xie, Marcus Rohrbach, Zhicheng Yan 等ICLR 2020 · 被引用 1,496 次
相关 Paper
- Resolving Task Confusion in Dynamic Expansion Architectures for Class Incremental LearningBingchen Huang, Zhineng Chen, Peng Zhou, Jiayin Chen 等AAAI 2023 · 被引用 31 次
- Task-Agnostic Guided Feature Expansion for Class-Incremental LearningBowen Zheng, Da-Wei Zhou, Han-Jia Ye, De-Chuan ZhanCVPR 2025
- DKT: Diverse Knowledge Transfer Transformer for Class Incremental LearningXinyuan Gao, Yuhang He, Songlin Dong, Jie Cheng 等CVPR 2023
- Prototype Reminiscence and Augmented Asymmetric Knowledge Aggregation for Non-Exemplar Class-Incremental LearningWuxuan Shi, Mang YeICCV 2023 · 被引用 49 次
- Harnessing Neural Unit Dynamics for Effective and Scalable Class-Incremental LearningDepeng Li, Tianqi Wang, Junwei Chen, Wei Dai 等ICML 2024 · 被引用 5 次
