MOS: Model Surgery for Pre-Trained Model-Based Class-Incremental Learning
Hai-Long Sun, Da-Wei Zhou, Hanbin Zhao, Le Gan, De-Chuan Zhan, Han-Jia Ye
摘要
Class-Incremental Learning (CIL) requires models to continually acquire knowledge of new classes without forgetting old ones. Despite Pre-trained Models (PTMs) have shown excellent performance in CIL, catastrophic forgetting still occurs as the model learns new concepts. Existing work seeks to utilize lightweight components to adjust the PTM, while the forgetting phenomenon still comes from parameter and retrieval levels. Specifically, iterative updates of the model result in parameter drift, while mistakenly retrieving irrelevant modules leads to the mismatch during inference. To this end, we propose MOdel Surgery (MOS) to rescue the model from forgetting previous knowledge. By training task-specific adapters, we continually adjust the PTM to downstream tasks. To mitigate parameter-level forgetting, we present an adapter merging approach to learn task-specific adapters, which aims to bridge the gap between different components while reserve task-specific information. Besides, to address retrieval-level forgetting, we introduce a training-free self-refined adapter retrieval mechanism during inference, which leverages the model's inherent ability for better adapter retrieval. By jointly rectifying the model with those steps, MOS can robustly resist catastrophic forgetting in the learning process. Extensive experiments on seven benchmark datasets validate MOS's state-of-the-art performance. Code is available at: https://github.com/sun-hailong/AAAI25-MOS
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper25
- Mixture of Noise for Pre-Trained Model-Based Class-Incremental LearningKai Jiang, Zhengyan Shi, Dell Zhang, Hongyuan Zhang 等NeurIPS 2025 · 被引用 38 次
- Enhancing Multimodal Large Language Models Complex Reason via Similarity ComputationXiaofeng Zhang, Fanshuo Zeng, Yihao Quan, Zheng Hui 等AAAI 2025 · 被引用 36 次
- Mitigating Visual Forgetting via Take-along Visual Conditioning for Multi-modal Long CoT ReasoningHai-Long Sun, Zhun Sun, Houwen Peng, Han-Jia YeACL 2025 · 被引用 23 次
- Fly-CL: A Fly-Inspired Framework for Enhancing Efficient Decorrelation and Reduced Training Time in Pre-trained Model-based Continual Representation LearningHeming Zou, Yunliang Zang, Wutong Xu, Xiangyang JiICLR 2026 · 被引用 13 次
- External Knowledge Injection for CLIP-Based Class-Incremental LearningDa-Wei Zhou, Kai-Wen Li, Jingyi Ning, Han-Jia Ye 等ICCV 2025 · 被引用 13 次
它引用的顶会 Paper22
- LoRA: Low-Rank Adaptation of Large Language ModelsEdward J. Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu 等ICLR 2022 · 被引用 18,833 次
- AdaptFormer: Adapting Vision Transformers for Scalable Visual RecognitionShoufa Chen, Chongjian Ge, Zhan Tong, Jiangliu Wang 等NeurIPS 2022 · 被引用 1,291 次
- Scaling & Shifting Your Features: A New Baseline for Efficient Model TuningDongze Lian, Daquan Zhou, Jiashi Feng, Xinchao WangNeurIPS 2022 · 被引用 415 次
- RanPAC: Random Projections and Pre-trained Models for Continual LearningMark D. McDonnell, Dong Gong, Amin Parvaneh, Ehsan Abbasnejad 等NeurIPS 2023 · 被引用 245 次
- SLCA: Slow Learner with Classifier Alignment for Continual Learning on a Pre-trained ModelGengwei Zhang, Liyuan Wang, Guoliang Kang, Ling Chen 等ICCV 2023 · 被引用 196 次
相关 Paper
- Integrating Task-Specific and Universal Adapters for Pre-Trained Model-Based Class-Incremental LearningYan Wang, Da-Wei Zhou, Han-Jia YeICCV 2025 · 被引用 7 次
- Random Amalgamation of Adapters for Flatter Loss Landscapes: Towards Class-Incremental Learning with Better StabilityYao Deng, Xiang Xiang, Jiaqi GuiAAAI 2026
- AnaCP: Toward Upper-Bound Continual Learning via Analytic Contrastive ProjectionSaleh Momeni, Changnan Xiao, Bing LiuNeurIPS 2025 · 被引用 8 次
- Knowledge Memorization and Rumination for Pre-trained Model-based Class-Incremental LearningZijian Gao, Wangwang Jia, Xingxing Zhang, Dulan Zhou 等CVPR 2025
- CL-LoRA: Continual Low-Rank Adaptation for Rehearsal-Free Class-Incremental LearningJiangpeng He, Zhihao Duan, Fengqing ZhuCVPR 2025
