In-Context Editing: Learning Knowledge from Self-Induced Distributions
Siyuan Qi, Bangcheng Yang, Kailin Jiang, Xiaobo Wang, Jiaqi Li, Yifan Zhong, Yaodong Yang, Zilong Zheng
Abstract
In scenarios where language models must incorporate new information efficiently without extensive retraining, traditional fine-tuning methods are prone to overfitting, degraded generalization, and unnatural language generation. To address these limitations, we introduce Consistent In-Context Editing (ICE), a novel approach leveraging the model's in-context learning capability to optimize towards a contextual distribution rather than a one-hot target. ICE introduces a simple yet effective optimization framework for the model to internalize new knowledge by aligning its output distributions with and without additional context. This method enhances the robustness and effectiveness of gradient-based tuning methods, preventing overfitting and preserving the model's integrity. We analyze ICE across four critical aspects of knowledge editing: accuracy, locality, generalization, and linguistic quality, demonstrating its advantages. Experimental results confirm the effectiveness of ICE and demonstrate its potential for continual editing, ensuring that the integrity of the model is preserved while updating information.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 2b29897a-3277-4483-b395-ac9e73fc0c43Cited by top-tier papers9
- Self-Distilled Reasoner: On-Policy Self-Distillation for Large Language ModelsSiyan Zhao, Zhihui Xie, Mengchen Liu, Jing Huang et al.ICML 2026 · 245 citations
- Doc-to-LoRA: Learning to Instantly Internalize ContextsRujikorn Charakorn, Edoardo Cetin, Shinnosuke Uesaka, Robert LangeICML 2026 · 27 citations
- CamEdit: Continuous Camera Parameter Control for Photorealistic Image EditingXinran Qin, Zhixin Wang, Fan Li, Haoyu Chen et al.NeurIPS 2025 · 17 citations
- When Large Multimodal Models Confront Evolving Knowledge: Challenges and ExplorationsKailin Jiang, Yuntao Du, Yukai Ding, Yuchen Ren et al.ICLR 2026 · 7 citations
- MMKU-Bench: A Multimodal Update Benchmark for Diverse Visual KnowledgeBaochen Fu, Yuntao Du, Cheng Chang, Baihao Jin et al.ICML 2026 · 7 citations
Builds on31
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah et al.NeurIPS 2020 · 64,255 citations
- Chain-of-Thought Prompting Elicits Reasoning in Large Language ModelsJason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma et al.NeurIPS 2022 · 22,562 citations
- Locating and Editing Factual Associations in GPTKevin Meng, David Bau, Alex Andonian, Yonatan BelinkovNeurIPS 2022 · 3,415 citations
- Fast Model Editing at ScaleEric Mitchell, Charles Lin, Antoine Bosselut, Chelsea Finn et al.ICLR 2022 · 527 citations
- Memory-Based Model Editing at ScaleEric Mitchell, Charles Lin, Antoine Bosselut, Christopher D. Manning et al.ICML 2022 · 520 citations
Related papers
- Can We Edit Factual Knowledge by In-Context Learning?Ce Zheng, Lei Li, Qingxiu Dong, Yuxuan Fan et al.EMNLP 2023 · 40 citations
- Knowledge Editing through Chain-of-ThoughtChangyue Wang, Weihang Su, Qingyao Ai, Yichen Tang et al.EMNLP 2025 · 2 citations
- Interpretability-based Tailored Knowledge Editing in TransformersYihuai Hong, Aldo LipaniEMNLP 2024
- Should We Really Edit Language Models? On the Evaluation of Edited Language ModelsQi Li, Xiang Liu, Zhenheng Tang, Peijie Dong et al.NeurIPS 2024 · 25 citations
- Decoding by Contrasting Knowledge: Enhancing Large Language Model Confidence on Edited FactsBaolong Bi, Shenghua Liu, Lingrui Mei, Yiwei Wang et al.ACL 2025
