Transferable Dialogue Systems and User Simulators
Bo-Hsiang Tseng, Yinpei Dai, Florian Kreyssig, Bill Byrne
Abstract
One of the difficulties in training dialogue systems is the lack of training data. We explore the possibility of creating dialogue data through the interaction between a dialogue system and a user simulator. Our goal is to develop a modelling framework that can incorporate new dialogue scenarios through self-play between the two agents. In this framework, we first pre-train the two agents on a collection of source domain dialogues, which equips the agents to converse with each other via natural language. With further fine-tuning on a small amount of target domain data, the agents continue to interact with the aim of improving their behaviors using reinforcement learning with structured reward functions. In experiments on the MultiWOZ dataset, two practical transfer learning problems are investigated: 1) domain adaptation and 2) single-to-multiple domain transfer. We demonstrate that the proposed framework is highly effective in bootstrapping the performance of the two agents in transfer learning. We also show that our method leads to improvements in dialogue system performance on complete datasets. Pre-training the Dialogue System and User Simulator In our joint learning framework, we first pre-train the DS and US using supervised learning so that two models are able to interact via natural language. This section presents the architectures of
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext a01ff621-e70c-4956-aed8-9c1418a5331fCited by top-tier papers11
- GALAXY: A Generative Pre-trained Model for Task-Oriented Dialog with Semi-supervised Learning and Explicit Policy InjectionWanwei He, Yinpei Dai, Yinhe Zheng, Yuchuan Wu et al.AAAI 2022 · 181 citations
- Flipping the Dialogue: Training and Evaluating User Language ModelsTarek Naous, Philippe Laban, Wei Xu, Jennifer NevilleICLR 2026 · 56 citations
- Unified Dialog Model Pre-training for Task-Oriented Dialog Understanding and GenerationWanwei He, Yinpei Dai, Min Yang, Jian Sun et al.SIGIR 2022 · 41 citations
- One Cannot Stand for Everyone! Leveraging Multiple User Simulators to train Task-oriented Dialogue SystemsYajiao Liu, Xin Jiang, Yichun Yin, Yasheng Wang et al.ACL 2023 · 5 citations
- DFA-RAG: Conversational Semantic Router for Large Language Model with Definite Finite AutomatonYiyou Sun, Junjie Hu, Wei Cheng, Haifeng ChenICML 2024 · 4 citations
Builds on5
- Task-Oriented Dialog Systems That Consider Multiple Appropriate Responses under the Same ContextYichi Zhang, Zhijian Ou, Zhou YuAAAI 2020 · 198 citations
- MinTL: Minimalist Transfer Learning for Task-Oriented Dialogue SystemsZhaojiang Lin, Andrea Madotto, Genta Indra Winata, Pascale FungEMNLP 2020 · 138 citations
- Paraphrase Augmented Task-Oriented Dialog GenerationSilin Gao, Yichi Zhang, Zhijian Ou, Zhou YuACL 2020 · 78 citations
- Multi-Agent Task-Oriented Dialog Policy Learning with Role-Aware Reward DecompositionRyuichi Takanobu, Runze Liang, Minlie HuangACL 2020 · 47 citations
- Multi-Domain Dialogue Acts and Response Co-GenerationKai Wang, Junfeng Tian, Rui Wang, Xiaojun Quan et al.ACL 2020 · 46 citations
Related papers
- NeuralWOZ: Learning to Collect Task-Oriented Dialogue via Model-Based SimulationSungdong Kim, Minsuk Chang, Sang-Woo LeeACL 2021
- Learning Efficient Dialogue Policy from Demonstrations through ShapingHuimin Wang, Baolin Peng, Kam-Fai WongACL 2020 · 18 citations
- Zero-Shot Transfer Learning with Synthesized Data for Multi-Domain Dialogue State TrackingGiovanni Campagna, Agata Foryciarz, Mehrad Moradshahi, Monica S. LamACL 2020 · 5 citations
- [CASPI] Causal-aware Safe Policy Improvement for Task-oriented DialogueGovardana Sachithanandam Ramachandran, Kazuma Hashimoto, Caiming XiongACL 2022 · 12 citations
- Learning Dialog Policies from Weak DemonstrationsGabriel Gordon-Hall, Philip John Gorinski, Shay B. CohenACL 2020 · 5 citations
