Injecting Domain Knowledge in Language Models for Task-oriented Dialogue Systems
Denis Emelin, Daniele Bonadiman, Sawsan Alqahtani, Yi Zhang, Saab Mansour
Abstract
Pre-trained language models (PLM) have advanced the state-of-the-art across NLP applications, but lack domain-specific knowledge that does not naturally occur in pre-training data. Previous studies augmented PLMs with symbolic knowledge for different downstream NLP tasks. However, knowledge bases (KBs) utilized in these studies are usually large-scale and static, in contrast to small, domain-specific, and modifiable knowledge bases that are prominent in real-world task-oriented dialogue (TOD) systems. In this paper, we showcase the advantages of injecting domain-specific knowledge prior to fine-tuning on TOD tasks. To this end, we utilize light-weight adapters that can be easily integrated with PLMs and serve as a repository for facts learned from different KBs. To measure the efficacy of proposed knowledge injection methods, we introduce Knowledge Probing using Response Selection (KPRS) -a probe designed specifically for TOD models. Experiments 1 on KPRS and the response generation task show improvements of knowledge injection with adapters over strong baselines.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 8dc0f020-40fa-4635-ae23-ffd18f4e38e9Cited by top-tier papers3
- StructGPT: A General Framework for Large Language Model to Reason over Structured DataJinhao Jiang, Kun Zhou, Zican Dong, Keming Ye et al.EMNLP 2023 · 173 citations
- Killing Two Birds with One Stone: Cross-modal Reinforced Prompting for Graph and Language TasksWenyuan Jiang, Wenwei Wu, Le Zhang, Zixuan Yuan et al.KDD 2024 · 4 citations
- RA2FD: Distilling Faithfulness into Efficient Dialogue SystemsZhiyuan Zhu, Yusheng Liao, Chenxin Xu, Yunfeng Guan et al.EMNLP 2024 · 3 citations
Builds on4
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah et al.NeurIPS 2020 · 64,255 citations
- BART: Denoising Sequence-to-Sequence Pre-training for Natural Language Generation, Translation, and ComprehensionMike Lewis, Yinhan Liu, Naman Goyal, Marjan Ghazvininejad et al.ACL 2020 · 1,224 citations
- A Simple Language Model for Task-Oriented DialogueEhsan Hosseini-Asl, Bryan McCann, Chien-Sheng Wu, Semih Yavuz et al.NeurIPS 2020 · 590 citations
- Dialogue State Tracking with a Language Model using Schema-Driven PromptingChia-Hsuan Lee, Hao Cheng, Mari OstendorfEMNLP 2021 · 87 citations
Related papers
- ADPL: Adversarial Prompt-based Domain Adaptation for Dialogue Summarization with Knowledge DisentanglementLulu Zhao, Fujia Zheng, Weihao Zeng, Keqing He et al.SIGIR 2022 · 6 citations
- Probing Linguistic Information for Logical Inference in Pre-trained Language ModelsZeming Chen, Qiyue GaoAAAI 2022 · 11 citations
- Plug-and-Play Knowledge Injection for Pre-trained Language ModelsZhengyan Zhang, Zhiyuan Zeng, Yankai Lin, Huadong Wang et al.ACL 2023 · 10 citations
- Mixture-of-Domain-Adapters: Decoupling and Injecting Domain Knowledge to Pre-trained Language Models' MemoriesShizhe Diao, Tianyang Xu, Ruijia Xu, Jiawei Wang et al.ACL 2023 · 17 citations
- Q-TOD: A Query-driven Task-oriented Dialogue SystemXin Tian, Yingzhan Lin, Mengfei Song, Siqi Bao et al.EMNLP 2022 · 13 citations
