Learning to Imagine: Distillation-Based Interactive Context Exploitation for Dialogue State Tracking
Jinyu Guo, Kai Shuang, Kaihang Zhang, Yixuan Liu, Jijie Li, Zihan Wang
摘要
In dialogue state tracking (DST), the exploitation of dialogue history is a crucial research direction, and the existing DST models can be divided into two categories: fullhistory models and partial-history models. Since the "select first, use later" mechanism explicitly filters the distracting information being passed to the downstream state prediction, the partial-history models have recently achieved a performance advantage over the full-history models. However, besides the redundant information, some critical dialogue context information was inevitably filtered out by the partialhistory models simultaneously. To reconcile the contextual consideration with avoiding the introduction of redundant information, we propose DICE-DST, a model-agnostic module widely applicable to the partial-history DST models, which aims to strengthen the ability of context exploitation for the encoder of each DST model. Specifically, we first construct a teacher encoder and devise two contextual reasoning tasks to train it to acquire extensive dialogue contextual knowledge. Then we transfer the contextual knowledge from the teacher encoder to the student encoder via a novel turn-level attention-alignment distillation. Experimental results show that our approach extensively improves the performance of partial-history DST models and thereby achieves new stateof-the-art performance on multiple mainstream datasets while keeping high efficiency.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Self-Supervised Continual Graph Learning in Adaptive Riemannian SpacesLi Sun, Junda Ye, Hao Peng, Feiyang Wang 等AAAI 2023 · 被引用 49 次
- Aligning Composed Query with Image via Discriminative Perception from Negative CorrespondencesYifan Wang, Wuliang Huang, Chun YuanAAAI 2025 · 被引用 3 次
它引用的顶会 Paper12
- ALBERT: A Lite BERT for Self-supervised Learning of Language RepresentationsZhenzhong Lan, Mingda Chen, Sebastian Goodman, Kevin Gimpel 等ICLR 2020 · 被引用 7,418 次
- Towards Scalable Multi-Domain Conversational Agents: The Schema-Guided Dialogue DatasetAbhinav Rastogi, Xiaoxue Zang, Srinivas Sunkara, Raghav Gupta 等AAAI 2020 · 被引用 707 次
- MobileBERT: a Compact Task-Agnostic BERT for Resource-Limited DevicesZhiqing Sun, Hongkun Yu, Xiaodan Song, Renjie Liu 等ACL 2020 · 被引用 660 次
- A Simple Language Model for Task-Oriented DialogueEhsan Hosseini-Asl, Bryan McCann, Chien-Sheng Wu, Semih Yavuz 等NeurIPS 2020 · 被引用 590 次
- Efficient Dialogue State Tracking by Selectively Overwriting MemorySungdong Kim, Sohee Yang, Gyuwan Kim, Sang-Woo LeeACL 2020 · 被引用 189 次
相关 Paper
- Beyond the Granularity: Multi-Perspective Dialogue Collaborative Selection for Dialogue State TrackingJinyu Guo, Kai Shuang, Jijie Li, Zihan Wang 等ACL 2022
- Dialogue State Distillation Network with Inter-slot Contrastive Learning for Dialogue State TrackingJing Xu, Dandan Song, Chong Liu, Siu Cheung Hui 等AAAI 2023 · 被引用 8 次
- Dual Slot Selector via Local Reliability Verification for Dialogue State TrackingJinyu Guo, Kai Shuang, Jijie Li, Zihan WangACL 2021
- Parallel Interactive Networks for Multi-Domain Dialogue State GenerationJunfan Chen, Richong Zhang, Yongyi Mao, Jie XuEMNLP 2020 · 被引用 22 次
- A Contextual Hierarchical Attention Network with Adaptive Objective for Dialogue State TrackingYong Shan, Zekang Li, Jinchao Zhang, Fandong Meng 等ACL 2020 · 被引用 58 次
