Knowledge Enhanced Fine-Tuning for Better Handling Unseen Entities in Dialogue Generation
Leyang Cui, Yu Wu, Shujie Liu, Yue Zhang
摘要
Although pre-training models have achieved great success in dialogue generation, their performance drops dramatically when the input contains an entity that does not appear in pretraining and fine-tuning datasets (unseen entity). To address this issue, existing methods leverage an external knowledge base to generate appropriate responses. In real-world scenario, the entity may not be included by the knowledge base or suffer from the precision of knowledge retrieval. To deal with this problem, instead of introducing knowledge base as the input, we force the model to learn a better semantic representation by predicting the information in the knowledge base, only based on the input context. Specifically, with the help of a knowledge base, we introduce two auxiliary training objectives: 1) Interpret Masked Word, which conjectures the meaning of the masked entity given the context; 2) Hypernym Generation, which predicts the hypernym of the entity based on the context. Experiment results on two dialogue corpus verify the effectiveness of our methods under both knowledge available and unavailable settings.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- Contrastive Learning Reduces Hallucination in ConversationsWeiwei Sun, Zhengliang Shi, Shen Gao, Pengjie Ren 等AAAI 2023 · 被引用 92 次
- Learning to Sample and Aggregate: Few-shot Reasoning over Temporal Knowledge GraphsRuijie Wang, Zheng Li, Dachun Sun, Shengzhong Liu 等NeurIPS 2022 · 被引用 61 次
- A Synthetic Data Generation Framework for Grounded DialoguesJianzhu Bao, Rui Wang, Yasheng Wang, Aixin Sun 等ACL 2023 · 被引用 11 次
它引用的顶会 Paper8
- BART: Denoising Sequence-to-Sequence Pre-training for Natural Language Generation, Translation, and ComprehensionMike Lewis, Yinhan Liu, Naman Goyal, Marjan Ghazvininejad 等ACL 2020 · 被引用 1,224 次
- LUKE: Deep Contextualized Entity Representations with Entity-aware Self-attentionIkuya Yamada, Akari Asai, Hiroyuki Shindo, Hideaki Takeda 等EMNLP 2020 · 被引用 562 次
- Sequential Latent Knowledge Selection for Knowledge-Grounded DialogueByeongchang Kim, Jaewoo Ahn, Gunhee KimICLR 2020 · 被引用 179 次
- Knowledge-Grounded Dialogue Generation with Pre-trained Language ModelsXueliang Zhao, Wei Wu, Can Xu, Chongyang Tao 等EMNLP 2020 · 被引用 153 次
- Bridging the Gap between Prior and Posterior Knowledge Selection for Knowledge-Grounded Dialogue GenerationXiuyi Chen, Fandong Meng, Peng Li, Feilong Chen 等EMNLP 2020 · 被引用 78 次
相关 Paper
- Eliciting Knowledge from Large Pre-Trained Models for Unsupervised Knowledge-Grounded ConversationYanyang Li, Jianqiao Zhao, Michael R. Lyu, Liwei WangEMNLP 2022 · 被引用 11 次
- KPT: Keyword-Guided Pre-training for Grounded Dialog GenerationQi Zhu, Fei Mi, Zheng Zhang, Yasheng Wang 等AAAI 2023 · 被引用 5 次
- Contextualize Knowledge Bases with Transformer for End-to-end Task-Oriented Dialogue SystemsYanjie Gou, Yinjie Lei, Lingqiao Liu, Yong Dai 等EMNLP 2021 · 被引用 11 次
- Masking Orchestration: Multi-Task Pretraining for Multi-Role Dialogue Representation LearningTianyi Wang, Yating Zhang, Xiaozhong Liu, Changlong Sun 等AAAI 2020 · 被引用 8 次
- Show, Interpret and Tell: Entity-Aware Contextualised Image Captioning in WikipediaKhanh Nguyen, Ali Furkan Biten, Andrés Mafla, Lluís Gómez 等AAAI 2023 · 被引用 15 次
