Zero-Shot Dialogue State Tracking via Cross-Task Transfer
Zhaojiang Lin, Bing Liu, Andrea Madotto, Seungwhan Moon, Zhenpeng Zhou, Paul A. Crook, Zhiguang Wang, Zhou Yu, Eunjoon Cho, Rajen Subba, Pascale Fung
Abstract
Zero-shot transfer learning for dialogue state tracking (DST) enables us to handle a variety of task-oriented dialogue domains without the expense of collecting in-domain data. In this work, we propose to transfer the crosstask knowledge from general question answering (QA) corpora for the zero-shot DST task. Specifically, we propose TransferQA, a transferable generative QA model that seamlessly combines extractive QA and multichoice QA via a text-to-text transformer framework, and tracks both categorical slots and non-categorical slots in DST. In addition, we introduce two effective ways to construct unanswerable questions, namely, negative question sampling and context truncation, which enable our model to handle "none" value slots in the zero-shot DST setting. The extensive experiments show that our approaches substantially improve the existing zero-shot and few-shot results on MultiWoz. Moreover, compared to the fully trained baseline on the Schema-Guided Dialogue dataset, our approach shows better generalization ability in unseen domains.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers7
- Leveraging Large Language Models to Power Chatbots for Collecting User Self-Reported DataJing Wei, Sungdong Kim, Hyunhoon Jung, Young-Ho KimCSCW 2024 · 82 citations
- A Dual Prompt Learning Framework for Few-Shot Dialogue State TrackingYuting Yang, Wenqiang Lei, Pei Huang, Juan Cao et al.WWW 2023 · 20 citations
- Large Language Models as Zero-shot Dialogue State Tracker through Function CallingZekun Li, Zhiyu Chen, Mike Ross, Patrick Huber et al.ACL 2024 · 9 citations
- Prompter: Zero-shot Adaptive Prefixes for Dialogue State Tracking Domain AdaptationIbrahim Taha Aksu, Min-Yen Kan, Nancy F. ChenACL 2023 · 3 citations
- Turn-Level Active Learning for Dialogue State TrackingZihan Zhang, Meng Fang, Fanghua Ye, Ling Chen et al.EMNLP 2023 · 3 citations
Builds on13
- Towards Scalable Multi-Domain Conversational Agents: The Schema-Guided Dialogue DatasetAbhinav Rastogi, Xiaoxue Zang, Srinivas Sunkara, Raghav Gupta et al.AAAI 2020 · 707 citations
- A Simple Language Model for Task-Oriented DialogueEhsan Hosseini-Asl, Bryan McCann, Chien-Sheng Wu, Semih Yavuz et al.NeurIPS 2020 · 590 citations
- TOD-BERT: Pre-trained Natural Language Understanding for Task-Oriented DialogueChien-Sheng Wu, Steven C. H. Hoi, Richard Socher, Caiming XiongEMNLP 2020 · 210 citations
- Task-Oriented Dialog Systems That Consider Multiple Appropriate Responses under the Same ContextYichi Zhang, Zhijian Ou, Zhou YuAAAI 2020 · 198 citations
- Efficient Dialogue State Tracking by Selectively Overwriting MemorySungdong Kim, Sohee Yang, Gyuwan Kim, Sang-Woo LeeACL 2020 · 189 citations
Related papers
- Divide, Conquer, and Combine: Mixture of Semantic-Independent Experts for Zero-Shot Dialogue State TrackingQingyue Wang, Liang Ding, Yanan Cao, Yibing Zhan et al.ACL 2023 · 5 citations
- Zero-Shot Transfer Learning with Synthesized Data for Multi-Domain Dialogue State TrackingGiovanni Campagna, Agata Foryciarz, Mehrad Moradshahi, Monica S. LamACL 2020 · 5 citations
- Similarity-based Multi-Domain Dialogue State Tracking with Copy Mechanisms for Task-based Virtual Personal AssistantsJarana Manotumruksa, Jeffrey Dalton, Edgar Meij, Emine YilmazWWW 2022 · 5 citations
- MA-DST: Multi-Attention-Based Scalable Dialog State TrackingAdarsh Kumar, Peter Ku, Anuj Kumar Goyal, Angeliki Metallinou et al.AAAI 2020 · 61 citations
- Non-Autoregressive Dialog State TrackingHung Le, Richard Socher, Steven C. H. HoiICLR 2020 · 54 citations
