Lifelong Knowledge Editing for Vision Language Models with Low-Rank Mixture-of-Experts
Qizhou Chen, Chengyu Wang, Dakan Wang, Taolin Zhang, Wangyue Li, Xiaofeng He
摘要
Model editing aims to correct inaccurate knowledge, update outdated information, and incorporate new data into Large Language Models (LLMs) without the need for retraining. This task poses challenges in lifelong scenarios where edits must be continuously applied for real-world applications. While some editors demonstrate strong robustness for lifelong editing in pure LLMs, Vision LLMs (VLLMs), which incorporate an additional vision modality, are not directly adaptable to existing LLM editors. In this paper, we propose LiveEdit, a Lifelong vision language model Edit to bridge the gap between lifelong LLM editing and VLLMs. We begin by training an editing expert generator to independently produce low-rank experts for each editing instance, with the goal of correcting the relevant responses of the VLLM. A hard filtering mechanism is developed to utilize visual semantic knowledge, thereby coarsely eliminating visually irrelevant experts for input queries during the inference stage of the post-edited model. Finally, to integrate visually relevant experts, we introduce a soft routing mechanism based on textual semantic relevance to achieve multi-expert fusion. For evaluation, we establish a benchmark for lifelong VLLM editing. Extensive experiments demonstrate that LiveEdit offers significant advantages in lifelong VLLM editing scenarios. Further experiments validate the rationality and effectiveness of each module design in LiveEdit. 1
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper11
- When Large Multimodal Models Confront Evolving Knowledge: Challenges and ExplorationsKailin Jiang, Yuntao Du, Yukai Ding, Yuchen Ren 等ICLR 2026 · 被引用 7 次
- DSCA: Dynamic Subspace Concept Alignment for Lifelong VLM EditingGyanendra Das, Sai Satyam JenaCVPR 2026 · 被引用 2 次
- Reliable Lifelong Multimodal Editing: Conflict-Aware Retrieval Meets Multi-Level GuidanceQiang Zhang, Fanrui Zhang, Jiawei Liu, Ming Hu 等NeurIPS 2025 · 被引用 1 次
- DeFT-LoRA: Decoupled and Fused Tuning with LoRA Experts for Universal Cross-Domain RetrievalKe Xu, Xiaozheng Shen, Shanshan Wang, Mengzhu Wang 等AAAI 2026
- ReasonEdit: Editing Vision--Language Models using Human ReasoningJiaxing Qiu, Kaihua Hou, Roxana Daneshjou, Ahmed Alaa 等ICML 2026
它引用的顶会 Paper40
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Visual Instruction TuningHaotian Liu, Chunyuan Li, Qingyang Wu, Yong Jae LeeNeurIPS 2023 · 被引用 11,349 次
- BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language ModelsJunnan Li, Dongxu Li, Silvio Savarese, Steven C. H. HoiICML 2023 · 被引用 7,873 次
相关 Paper
- Reinforced Lifelong Editing for Language ModelsZherui Li, Houcheng Jiang, Hao Chen, Baolong Bi 等ICML 2025
- Attribution Analysis Meets Model Editing: Advancing Knowledge Correction in Vision Language Models with VisEditQizhou Chen, Taolin Zhang, Chengyu Wang, Xiaofeng He 等AAAI 2025 · 被引用 9 次
- Towards Scalable Lifelong Knowledge Editing with Selective Knowledge SuppressionDahyun Jung, Jaewook Lee, Heuiseok LimACL 2026
- Visual-Oriented Fine-Grained Knowledge Editing for MultiModal Large Language ModelsZhen Zeng, Leijiang Gu, Xun Yang, Zhangling Duan 等ICCV 2025 · 被引用 3 次
- MMKE-Bench: A Multimodal Editing Benchmark for Diverse Visual KnowledgeYuntao Du, Kailin Jiang, Zhi Gao, Chenrui Shi 等ICLR 2025
