Amalgamating Knowledge from Two Teachers for Task-oriented Dialogue System with Adversarial Training
Wanwei He, Min Yang, Rui Yan, Chengming Li, Ying Shen, Ruifeng Xu
Abstract
The challenge of both achieving task completion by querying the knowledge base and generating human-like responses for task-oriented dialogue systems is attracting increasing research attention. In this paper, we propose a "Two-Teacher One-Student" learning framework (TTOS) for task-oriented dialogue, with the goal of retrieving accurate KB entities and generating human-like responses simultaneously. TTOS amalgamates knowledge from two teacher networks that together provide comprehensive guidance to build a highquality task-oriented dialogue system (student network). Each teacher network is trained via reinforcement learning with a goal-specific reward, which can be viewed as an expert towards the goal and transfers the professional characteristic to the student network. Instead of adopting the classic student-teacher learning of forcing the output of a student network to exactly mimic the soft targets produced by the teacher networks, we introduce two discriminators as in generative adversarial network (GAN) to transfer knowledge from two teachers to the student. The usage of discriminators relaxes the rigid coupling between the student and teachers. Extensive experiments on two benchmark datasets (i.e., CamRest and In-Car Assistant) demonstrate that TTOS significantly outperforms baseline methods. For reproducibility, we release the code and data at https://github.com/siat-nlp/TTOS .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext a12fb390-7921-4bec-a32b-bd247dcc9a3eCited by top-tier papers6
- GALAXY: A Generative Pre-trained Model for Task-Oriented Dialog with Semi-supervised Learning and Explicit Policy InjectionWanwei He, Yinpei Dai, Yinhe Zheng, Yuchuan Wu et al.AAAI 2022 · 181 citations
- Unified Dialog Model Pre-training for Task-Oriented Dialog Understanding and GenerationWanwei He, Yinpei Dai, Min Yang, Jian Sun et al.SIGIR 2022 · 41 citations
- End-to-end Task-oriented Dialogue: A Survey of Tasks, Methods, and Future DirectionsLibo Qin, Wenbo Pan, Qiguang Chen, Lizi Liao et al.EMNLP 2023 · 12 citations
- Exploring Auxiliary Reasoning Tasks for Task-oriented Dialog Systems with Meta Cooperative LearningBowen Qin, Min Yang, Lidong Bing, Qingshan Jiang et al.AAAI 2021 · 9 citations
- From Retrieval to Generation: A Simple and Unified Generative Model for End-to-End Task-Oriented DialogueZeyuan Ding, Zhihao Yang, Ling Luo, Yuanyuan Sun et al.AAAI 2024 · 6 citations
Related papers
- Dual-Feedback Knowledge Retrieval for Task-Oriented Dialogue SystemsTianyuan Shi, Liangzhi Li, Zijian Lin, Tao Yang et al.EMNLP 2023 · 9 citations
- Relevance Is a Guiding Light: Relevance-aware Adaptive Learning for End-to-end Task-oriented Dialogue SystemZhanpeng Chen, Zhihong Zhu, Wanshi Xu, Xianwei Zhuang et al.EMNLP 2024 · 5 citations
- Multi-Grained Knowledge Retrieval for End-to-End Task-Oriented DialogFanqi Wan, Weizhou Shen, Ke Yang, Xiaojun Quan et al.ACL 2023 · 14 citations
- GraphDialog: Integrating Graph Knowledge into End-to-End Task-Oriented Dialogue SystemsShiquan Yang, Rui Zhang, Sarah M. ErfaniEMNLP 2020 · 46 citations
- Q-TOD: A Query-driven Task-oriented Dialogue SystemXin Tian, Yingzhan Lin, Mengfei Song, Siqi Bao et al.EMNLP 2022 · 13 citations
