Overcoming Generic Knowledge Loss with Selective Parameter Update
Wenxuan Zhang, Paul Janson, Rahaf Aljundi, Mohamed Elhoseiny
摘要
Foundation models encompass an extensive knowledge base and offer remarkable transferability. However, this knowledge becomes outdated or insufficient over time. The challenge lies in continuously updating foundation models to accommodate novel information while retaining their original capabilities. Leveraging the fact that foundation models have initial knowledge on various tasks and domains, we propose a novel approach that, instead of updating all parame-ters equally, localizes the updates to a sparse set of parame-ters relevant to the task being learned. We strike a balance between efficiency and new task performance, while main-taining the transferability and generalizability of foundation models. We extensively evaluate our method on foundational vision-language models with a diverse spectrum of continual learning tasks. Our method achieves improvements on the accuracy of the newly learned tasks up to 7% while preserving the pretraining knowledge with a negligible decrease of 0.9% on a representative control set accuracy. Code is avail-able here: https://github.com/wx-zhang/spu
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper10
- Ask and Remember: A Questions-Only Replay Strategy for Continual Visual Question AnsweringImad Eddine Marouf, Enzo Tartaglione, Stéphane Lathuilière, Joost van de WeijerICCV 2025 · 被引用 4 次
- Spectral Imbalance Causes Forgetting in Low-Rank Continual AdaptationHao Gu, Mao-Lin Luo, Zi-Hao Zhou, Han-Chen Zhang 等ICML 2026 · 被引用 3 次
- Little By Little: Continual Learning via Incremental Mixture of Rank-1 Associative Memory ExpertsHaodong Lu, Chongyang Zhao, Minhui Xue, Lina Yao 等ICML 2026 · 被引用 2 次
- Learn from Downstream and Be Yourself in Multimodal Large Language Models Fine-TuningWenke Huang, Jian Liang, Zekun Shi, Didi Zhu 等ICML 2025 · 被引用 1 次
- L2-LoRA: Improving Low-Rank Adaptation with Layer-Specific RegularizationXiang Zhang, Rui Xie, Shikun ZhangAAAI 2026
它引用的顶会 Paper32
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- LoRA: Low-Rank Adaptation of Large Language ModelsEdward J. Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu 等ICLR 2022 · 被引用 18,833 次
- BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language ModelsJunnan Li, Dongxu Li, Silvio Savarese, Steven C. H. HoiICML 2023 · 被引用 7,873 次
- Flamingo: a Visual Language Model for Few-Shot LearningJean-Baptiste Alayrac, Jeff Donahue, Pauline Luc, Antoine Miech 等NeurIPS 2022 · 被引用 6,707 次
- PaLM-E: An Embodied Multimodal Language ModelDanny Driess, Fei Xia, Mehdi S. M. Sajjadi, Corey Lynch 等ICML 2023 · 被引用 2,601 次
相关 Paper
- Instruction-Grounded Visual Projectors for Continual Learning of Generative Vision-Language ModelsHyundong Jin, Hyung Jin Chang, Eunwoo KimICCV 2025 · 被引用 2 次
- SD-LoRA: Scalable Decoupled Low-Rank Adaptation for Class Incremental LearningYichen Wu, Hongming Piao, Long-Kai Huang, Renzhen Wang 等ICLR 2025
- Sparse Tuning Enhances Plasticity in PTM-based Continual LearningHuan Zhang, Shenghua Fan, Shuyu Dong, Yujin Zheng 等AAAI 2026
- KeepLoRA: Continual Learning with Residual Gradient AdaptationMao-Lin Luo, Zi-Hao Zhou, Yi-Lin Zhang, Yuanyu Wan 等ICLR 2026 · 被引用 23 次
- Progressive Prompts: Continual Learning for Language ModelsAnastasia Razdaibiedina, Yuning Mao, Rui Hou, Madian Khabsa 等ICLR 2023 · 被引用 15 次
