Plug-and-Play Knowledge Injection for Pre-trained Language Models
Zhengyan Zhang, Zhiyuan Zeng, Yankai Lin, Huadong Wang, Deming Ye, Chaojun Xiao, Xu Han, Zhiyuan Liu, Peng Li, Maosong Sun, Jie Zhou
Abstract
Injecting external knowledge can improve the performance of pre-trained language models (PLMs) on various downstream NLP tasks. However, massive retraining is required to deploy new knowledge injection methods or knowledge bases for downstream tasks. In this work, we are the first to study how to improve the flexibility and efficiency of knowledge injection by reusing existing downstream models. To this end, we explore a new paradigm plugand-play knowledge injection, where knowledge bases are injected into frozen existing downstream models by a knowledge plugin. Correspondingly, we propose a plug-and-play injection method map-tuning, which trains a mapping of knowledge embeddings to enrich model inputs with mapped embeddings while keeping model parameters frozen. Experimental results on three knowledge-driven NLP tasks show that existing injection methods are not suitable for the new paradigm, while maptuning effectively improves the performance of downstream models. Moreover, we show that a frozen downstream model can be well adapted to different domains with different mapping networks of domain knowledge. Our code and models are available at https://github.com/ THUNLP/Knowledge-Plugin .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext b8ca993f-0eb0-4465-bf94-38c9d12e5653Cited by top-tier papers9
- Cartridges: Lightweight and general-purpose long context representations via self-studySabri Eyuboglu, Ryan Ehrlich, Simran Arora, Neel Guha et al.ICLR 2026 · 74 citations
- Augmentation-Adapted Retriever Improves Generalization of Language Models as Generic Plug-InZichun Yu, Chenyan Xiong, Shi Yu, Zhiyuan LiuACL 2023 · 15 citations
- What Will My Model Forget? Forecasting Forgotten Examples in Language Model RefinementXisen Jin, Xiang RenICML 2024 · 8 citations
- Bidirectional LMs are Better Knowledge Memorizers? A Benchmark for Real-world Knowledge InjectionYuwei Zhang, Wenhao Yu, Shangbin Feng, Yifan Zhu et al.ACL 2026 · 7 citations
- Plug-and-Play Document Modules for Pre-trained ModelsChaojun Xiao, Zhengyan Zhang, Xu Han, Chi-Min Chan et al.ACL 2023 · 5 citations
Builds on14
- Retrieval-Augmented Generation for Knowledge-Intensive NLP TasksPatrick Lewis, Ethan Perez, Aleksandra Piktus, Fabio Petroni et al.NeurIPS 2020 · 19,162 citations
- LoRA: Low-Rank Adaptation of Large Language ModelsEdward J. Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu et al.ICLR 2022 · 18,833 citations
- Flamingo: a Visual Language Model for Few-Shot LearningJean-Baptiste Alayrac, Jeff Donahue, Pauline Luc, Antoine Miech et al.NeurIPS 2022 · 6,707 citations
- Plug and Play Language Models: A Simple Approach to Controlled Text GenerationSumanth Dathathri, Andrea Madotto, Janice Lan, Jane Hung et al.ICLR 2020 · 1,166 citations
- LUKE: Deep Contextualized Entity Representations with Entity-aware Self-attentionIkuya Yamada, Akari Asai, Hiroyuki Shindo, Hideaki Takeda et al.EMNLP 2020 · 562 citations
Related papers
- Knowledge Rumination for Pre-trained Language ModelsYunzhi Yao, Peng Wang, Shengyu Mao, Chuanqi Tan et al.EMNLP 2023 · 2 citations
- KILM: Knowledge Injection into Encoder-Decoder Language ModelsYan Xu, Mahdi Namazifar, Devamanyu Hazarika, Aishwarya Padmakumar et al.ACL 2023 · 16 citations
- Unfreeze with Care: Space-Efficient Fine-Tuning of Semantic Parsing ModelsWeiqi Sun, Haidar Khan, Nicolas Guenon des Mesnards, Melanie Rubino et al.WWW 2022 · 5 citations
- Bridging Subword Gaps in Pretrain-Finetune Paradigm for Natural Language GenerationXin Liu, Baosong Yang, Dayiheng Liu, Haibo Zhang et al.ACL 2021
- Injecting Domain Knowledge in Language Models for Task-oriented Dialogue SystemsDenis Emelin, Daniele Bonadiman, Sawsan Alqahtani, Yi Zhang et al.EMNLP 2022 · 9 citations
