Meta-Reinforced Multi-Domain State Generator for Dialogue Systems
Yi Huang, Junlan Feng, Min Hu, Xiaoting Wu, Xiaoyu Du, Shuo Ma
Abstract
A Dialogue State Tracker (DST) is a core component of a modular task-oriented dialogue system. Tremendous progress has been made in recent years. However, the major challenges remain. The state-of-the-art accuracy for DST is below 50% for a multi-domain dialogue task. A learnable DST for any new domain requires a large amount of labeled indomain data and training from scratch. In this paper, we propose a Meta-Reinforced Multi-Domain State Generator (MERET). Our first contribution is to improve the DST accuracy. We enhance a neural model based DST generator with a reward manager, which is built on policy gradient reinforcement learning (R-L) to fine-tune the generator. With this change, we are able to improve the joint accuracy of DST from 48.79% to 50.91% on the Multi-WOZ corpus. Second, we explore to train a DST meta-learning model with a few domains as source domains and a new domain as target domain. We apply the model-agnostic metalearning (MAML) algorithm to DST and the obtained meta-learning model is used for new domain adaptation. Our experimental results show this solution is able to outperform the traditional training approach with extremely less training data in target domain.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext b53906ca-c1e2-444d-805a-95a5a75fa858Cited by top-tier papers6
- Slot Self-Attentive Dialogue State TrackingFanghua Ye, Jarana Manotumruksa, Qiang Zhang, Shenghui Li et al.WWW 2021 · 68 citations
- Parallel Interactive Networks for Multi-Domain Dialogue State GenerationJunfan Chen, Richong Zhang, Yongyi Mao, Jie XuEMNLP 2020 · 22 citations
- MetaASSIST: Robust Dialogue State Tracking with Meta LearningFanghua Ye, Xi Wang, Jie Huang, Shenghui Li et al.EMNLP 2022 · 10 citations
- kFolden: k-Fold Ensemble for Out-Of-Distribution DetectionXiaoya Li, Jiwei Li, Xiaofei Sun, Chun Fan et al.EMNLP 2021 · 6 citations
- Prompter: Zero-shot Adaptive Prefixes for Dialogue State Tracking Domain AdaptationIbrahim Taha Aksu, Min-Yen Kan, Nancy F. ChenACL 2023 · 3 citations
Related papers
- A Student-Teacher Architecture for Dialog Domain Adaptation Under the Meta-Learning SettingKun Qian, Wei Wei, Zhou YuAAAI 2021 · 8 citations
- Zero-Shot Transfer Learning with Synthesized Data for Multi-Domain Dialogue State TrackingGiovanni Campagna, Agata Foryciarz, Mehrad Moradshahi, Monica S. LamACL 2020 · 5 citations
- Dialog State Tracking with Reinforced Data AugmentationYichun Yin, Lifeng Shang, Xin Jiang, Xiao Chen et al.AAAI 2020 · 25 citations
- Non-Autoregressive Dialog State TrackingHung Le, Richard Socher, Steven C. H. HoiICLR 2020 · 54 citations
- MA-DST: Multi-Attention-Based Scalable Dialog State TrackingAdarsh Kumar, Peter Ku, Anuj Kumar Goyal, Angeliki Metallinou et al.AAAI 2020 · 61 citations
