Linguistically-Enriched and Context-AwareZero-shot Slot Filling
A. B. Siddique, Fuad T. Jamour, Vagelis Hristidis
摘要
Slot filling is identifying contiguous spans of words in an utterance that correspond to certain parameters (i.e., slots) of a user request/query. Slot filling is one of the most important challenges in modern task-oriented dialog systems. Supervised learning approaches have proven effective at tackling this challenge, but they need a significant amount of labeled training data in a given domain. However, new domains (i.e., unseen in training) may emerge after deployment. Thus, it is imperative that these models seamlessly adapt and fill slots from both seen and unseen domains -unseen domains contain unseen slot types with no training data, and even seen slots in unseen domains are typically presented in different contexts. This setting is commonly referred to as zero-shot slot filling. Little work has focused on this setting, with limited experimental evaluation. Existing models that mainly rely on contextindependent embedding-based similarity measures fail to detect slot values in unseen domains or do so only partially. We propose a new zero-shot slot filling neural model, LEONA, which works in three steps. Step one acquires domain-oblivious, context-aware representations of the utterance word by exploiting (a) linguistic features such as part-of-speech; (b) named entity recognition cues; and (c) contextual embeddings from pre-trained language models. Step two fine-tunes these rich representations and produces slotindependent tags for each word. Step three exploits generalizable context-aware utterance-slot similarity features at the word level, uses slot-independent tags, and contextualizes them to produce slot-specific predictions for each word. Our thorough evaluation on four diverse public datasets demonstrates that our approach consistently outperforms the state-of-the-art models by 17.52%, 22.15%, 17.42%, and 17.95% on average for unseen domains on SNIPS, ATIS, MultiWOZ, and SGD datasets, respectively.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper3
- Towards Scalable Multi-Domain Conversational Agents: The Schema-Guided Dialogue DatasetAbhinav Rastogi, Xiaoxue Zang, Srinivas Sunkara, Raghav Gupta 等AAAI 2020 · 被引用 707 次
- Few-shot Slot Tagging with Collapsed Dependency Transfer and Label-enhanced Task-adaptive Projection NetworkYutai Hou, Wanxiang Che, Yongkui Lai, Zhihan Zhou 等ACL 2020 · 被引用 186 次
- Neural Snowball for Few-Shot Relation LearningTianyu Gao, Xu Han, Ruobing Xie, Zhiyuan Liu 等AAAI 2020 · 被引用 85 次
相关 Paper
- Zero-Shot Slot Filling with Slot-Prefix Prompting and Attention Relationship DescriptorQiaoyang Luo, Lingqiao LiuAAAI 2023 · 被引用 10 次
- Novel Slot Detection: A Benchmark for Discovering Unknown Slot Types in the Task-Oriented Dialogue SystemYanan Wu, Zhiyuan Zeng, Keqing He, Hong Xu 等ACL 2021
- Robust Retrieval Augmented Generation for Zero-shot Slot FillingMichael R. Glass, Gaetano Rossiello, Md. Faisal Mahbub Chowdhury, Alfio GliozzoEMNLP 2021 · 被引用 18 次
- Understanding Medical Conversations with Scattered Keyword Attention and Weak Supervision from ResponsesXiaoming Shi, Haifeng Hu, Wanxiang Che, Zhongqian Sun 等AAAI 2020 · 被引用 36 次
- Adaptive End-to-End Metric Learning for Zero-Shot Cross-Domain Slot FillingYuanjun Shi, Linzhi Wu, Minglai ShaoEMNLP 2023 · 被引用 5 次
