Dynamic Multi-Layer Null Space Projection for Vision-Language Continual Learning
Borui Kang, Lei Wang, Zhiping Wu, Tao Feng, Yawen Li, Yang Gao, Wenbin Li
摘要
Vision-Language Models (VLM) have emerged as a highly promising approach for Continual Learning (CL) due to their powerful generalizable features. While adapter-based VLM can exploit both task-specific and task-agnostic features, current CL methods have largely overlooked the distinct and evolving parameter distributions in visual and language modalities, which are found crucial for effectively mitigating catastrophic forgetting. In this study, we find that the visual modality experiences a broader parameter distribution and greater variance during class increments than the textual modality, leading to higher vulnerability to forgetting. Consequently, we handle the branches of the two modalities asymmetrically. Specifically, we propose a Dynamic Multi-layer Null Space Projection (DMNSP) strategy and apply it only to the visual modality branch, while optimizing the language branch according to the original optimizer. DMNSP can restrict the update of visual parameters within the common subspace of multiple null spaces, further limiting the impact of non-zero residual terms. Simultaneously, combined with a dynamic projection coefficient, we can precisely control the magnitude of gradient projection to the null space, endowing the model with a good balance of stability and plasticity. Extensive experiments on TinyImageNet, CIFAR100 and ImageNet-R demonstrate that our method outperforms current approaches in accuracy and knowledge retention, setting a new standard for state-of-the-art performance in class incremental learning. Our code is available at https://github.com/RL-VIG/DMNSP.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- Affordance-First Decomposition for Continual Learning in Video–Language UnderstandingMengzhu xu, Hanzhi Liu, Ningkang Peng, qianyu Chen 等CVPR 2026 · 被引用 7 次
- Learning from Itself: Mining Internal Knowledge from Vision Language Models for Continual LearningYizheng Gong, Siyue Yu, Waleed Al-Nuaimy, Jimin XiaoCVPR 2026
- Branch, or Layer? Zeroth-Order Optimization for Continual Learning of Vision-Language ModelsZiwei Liu, Borui Kang, Wei Li, Hangjie Yuan 等AAAI 2026
- Don't Forget Why You Started: Tackling Dual Forgetting in Vision-Language Continual LearningBorui Kang, Jinrui Gu, Tao Feng, Qi Fan 等ICML 2026
它引用的顶会 Paper26
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- LoRA: Low-Rank Adaptation of Large Language ModelsEdward J. Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu 等ICLR 2022 · 被引用 18,833 次
- AdaptFormer: Adapting Vision Transformers for Scalable Visual RecognitionShoufa Chen, Chongjian Ge, Zhan Tong, Jiangliu Wang 等NeurIPS 2022 · 被引用 1,291 次
- Learning to Prompt for Continual LearningZifeng Wang, Zizhao Zhang, Chen-Yu Lee, Han Zhang 等CVPR 2022 · 被引用 635 次
- Gradient Projection Memory for Continual LearningGobinda Saha, Isha Garg, Kaushik RoyICLR 2021 · 被引用 409 次
相关 Paper
- Memory-Free Continual Learning with Null Space Adaptation for Zero-Shot Vision-Language ModelsYujin Jo, Taesup KimICLR 2026 · 被引用 2 次
- Boosting Continual Learning of Vision-Language Models via Mixture-of-Experts AdaptersJiazuo Yu, Yunzhi Zhuge, Lu Zhang, Ping Hu 等CVPR 2024 · 被引用 80 次
- Vision-language Incremental Learning with Dual Class-individual MemoryFuhai Chen, Feng Zhang, Xiaoguang Ma, Yiyi Zhou 等AAAI 2026
- Subspace Alignment for CLIP-based Continual Learning via Canonical Correlation AnalysisHuan Zhang, Shuyu Dong, Yujin Zheng, Dingwen Wang 等CVPR 2026
- Training Networks in Null Space of Feature Covariance for Continual LearningShipeng Wang, Xiaorong Li, Jian Sun, Zongben XuCVPR 2021
