Amalgamating Knowledge from Two Teachers for Task-oriented Dialogue System with Adversarial Training
Wanwei He, Min Yang, Rui Yan, Chengming Li, Ying Shen, Ruifeng Xu
摘要
The challenge of both achieving task completion by querying the knowledge base and generating human-like responses for task-oriented dialogue systems is attracting increasing research attention. In this paper, we propose a "Two-Teacher One-Student" learning framework (TTOS) for task-oriented dialogue, with the goal of retrieving accurate KB entities and generating human-like responses simultaneously. TTOS amalgamates knowledge from two teacher networks that together provide comprehensive guidance to build a highquality task-oriented dialogue system (student network). Each teacher network is trained via reinforcement learning with a goal-specific reward, which can be viewed as an expert towards the goal and transfers the professional characteristic to the student network. Instead of adopting the classic student-teacher learning of forcing the output of a student network to exactly mimic the soft targets produced by the teacher networks, we introduce two discriminators as in generative adversarial network (GAN) to transfer knowledge from two teachers to the student. The usage of discriminators relaxes the rigid coupling between the student and teachers. Extensive experiments on two benchmark datasets (i.e., CamRest and In-Car Assistant) demonstrate that TTOS significantly outperforms baseline methods. For reproducibility, we release the code and data at https://github.com/siat-nlp/TTOS .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- GALAXY: A Generative Pre-trained Model for Task-Oriented Dialog with Semi-supervised Learning and Explicit Policy InjectionWanwei He, Yinpei Dai, Yinhe Zheng, Yuchuan Wu 等AAAI 2022 · 被引用 181 次
- Unified Dialog Model Pre-training for Task-Oriented Dialog Understanding and GenerationWanwei He, Yinpei Dai, Min Yang, Jian Sun 等SIGIR 2022 · 被引用 41 次
- End-to-end Task-oriented Dialogue: A Survey of Tasks, Methods, and Future DirectionsLibo Qin, Wenbo Pan, Qiguang Chen, Lizi Liao 等EMNLP 2023 · 被引用 12 次
- Exploring Auxiliary Reasoning Tasks for Task-oriented Dialog Systems with Meta Cooperative LearningBowen Qin, Min Yang, Lidong Bing, Qingshan Jiang 等AAAI 2021 · 被引用 9 次
- From Retrieval to Generation: A Simple and Unified Generative Model for End-to-End Task-Oriented DialogueZeyuan Ding, Zhihao Yang, Ling Luo, Yuanyuan Sun 等AAAI 2024 · 被引用 6 次
相关 Paper
- Dual-Feedback Knowledge Retrieval for Task-Oriented Dialogue SystemsTianyuan Shi, Liangzhi Li, Zijian Lin, Tao Yang 等EMNLP 2023 · 被引用 9 次
- Relevance Is a Guiding Light: Relevance-aware Adaptive Learning for End-to-end Task-oriented Dialogue SystemZhanpeng Chen, Zhihong Zhu, Wanshi Xu, Xianwei Zhuang 等EMNLP 2024 · 被引用 5 次
- Multi-Grained Knowledge Retrieval for End-to-End Task-Oriented DialogFanqi Wan, Weizhou Shen, Ke Yang, Xiaojun Quan 等ACL 2023 · 被引用 14 次
- GraphDialog: Integrating Graph Knowledge into End-to-End Task-Oriented Dialogue SystemsShiquan Yang, Rui Zhang, Sarah M. ErfaniEMNLP 2020 · 被引用 46 次
- Q-TOD: A Query-driven Task-oriented Dialogue SystemXin Tian, Yingzhan Lin, Mengfei Song, Siqi Bao 等EMNLP 2022 · 被引用 13 次
