A Probabilistic End-To-End Task-Oriented Dialog Model with Latent Belief States towards Semi-Supervised Learning
Yichi Zhang, Zhijian Ou, Min Hu, Junlan Feng
Abstract
Structured belief states are crucial for user goal tracking and database query in task-oriented dialog systems. However, training belief trackers often requires expensive turn-level annotations of every user utterance. In this paper we aim at alleviating the reliance on belief state labels in building end-to-end dialog systems, by leveraging unlabeled dialog data towards semi-supervised learning. We propose a probabilistic dialog model, called the LAtent BElief State (LABES) model, where belief states are represented as discrete latent variables and jointly modeled with system responses given user inputs. Such latent variable modeling enables us to develop semi-supervised learning under the principled variational learning framework. Furthermore, we introduce LABES-S2S, which is a copyaugmented Seq2Seq model instantiation of LABES 1 . In supervised experiments, LABES-S2S obtains strong results on three benchmark datasets of different scales. In utilizing unlabeled dialog data, semi-supervised LABES-S2S significantly outperforms both supervisedonly and semi-supervised baselines. Remarkably, we can reduce the annotation demands to 50% without performance loss on MultiWOZ.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 55b9783f-0b01-448e-a1bf-9108418ab8f6Cited by top-tier papers10
- Multi-Task Pre-Training for Plug-and-Play Task-Oriented Dialogue SystemYixuan Su, Lei Shu, Elman Mansimov, Arshit Gupta et al.ACL 2022 · 218 citations
- GALAXY: A Generative Pre-trained Model for Task-Oriented Dialog with Semi-supervised Learning and Explicit Policy InjectionWanwei He, Yinpei Dai, Yinhe Zheng, Yuchuan Wu et al.AAAI 2022 · 181 citations
- Latent Execution for Neural Program Synthesis Beyond Domain-Specific LanguagesXinyun Chen, Dawn Song, Yuandong TianNeurIPS 2021 · 56 citations
- Semi-Supervised Variational Reasoning for Medical Dialogue GenerationDongdong Li, Zhaochun Ren, Pengjie Ren, Zhumin Chen et al.SIGIR 2021 · 45 citations
- Unified Dialog Model Pre-training for Task-Oriented Dialog Understanding and GenerationWanwei He, Yinpei Dai, Min Yang, Jian Sun et al.SIGIR 2022 · 41 citations
Builds on7
- A Simple Language Model for Task-Oriented DialogueEhsan Hosseini-Asl, Bryan McCann, Chien-Sheng Wu, Semih Yavuz et al.NeurIPS 2020 · 590 citations
- Task-Oriented Dialog Systems That Consider Multiple Appropriate Responses under the Same ContextYichi Zhang, Zhijian Ou, Zhou YuAAAI 2020 · 198 citations
- Sequential Latent Knowledge Selection for Knowledge-Grounded DialogueByeongchang Kim, Jaewoo Ahn, Gunhee KimICLR 2020 · 179 citations
- Paraphrase Augmented Task-Oriented Dialog GenerationSilin Gao, Yichi Zhang, Zhijian Ou, Zhou YuACL 2020 · 78 citations
- MOSS: End-to-End Dialog System Framework with Modular SupervisionWeixin Liang, Youzhi Tian, Chengcai Chen, Zhou YuAAAI 2020 · 55 citations
Related papers
- UBAR: Towards Fully End-to-End Task-Oriented Dialog System with GPT-2Yunyi Yang, Yunhao Li, Xiaojun QuanAAAI 2021 · 217 citations
- MinTL: Minimalist Transfer Learning for Task-Oriented Dialogue SystemsZhaojiang Lin, Andrea Madotto, Genta Indra Winata, Pascale FungEMNLP 2020 · 138 citations
- Unsupervised End-to-End Task-Oriented Dialogue with LLMs: The Power of the Noisy ChannelBrendan King, Jeffrey FlaniganEMNLP 2024 · 3 citations
- Semi-Supervised Dialogue Policy Learning via Stochastic Reward EstimationXinting Huang, Jianzhong Qi, Yu Sun, Rui ZhangACL 2020 · 19 citations
- Unsupervised Learning of Deterministic Dialogue Structure with Edge-Enhanced Graph Auto-EncoderYajing Sun, Yong Shan, Chengguang Tang, Yue Hu et al.AAAI 2021 · 17 citations
