Zero-shot Cross-domain Dialogue State Tracking via Context-aware Auto-prompting and Instruction-following Contrastive Decoding
Xiaoyu Dong, Yujie Feng, Zexin Lu, Guangyuan Shi, Xiao-Ming Wu
摘要
Zero-shot cross-domain dialogue state tracking (DST) enables us to manage task-oriented dialogues in new, unseen domains without the cost of collecting in-domain data. Previous studies have implemented slot-based input improvements, such as schema-driven descriptions and question-answering formats, but still suffer from negative transfer for seen slots and inefficient transfer for unseen slots due to the significant source-target domain gap. To address these issues, we introduce a novel framework called Context-aware Auto-prompting and Instruction-following Contrastive Decoding (CAPID). This framework generates dynamic, context-aware slot queries, effectively improving the model's transferability. Our context-aware auto-prompting approach tailors slot queries to the current dialogue context, increasing flexibility and reducing ambiguities. Additionally, an instruction-following contrastive decoding strategy helps reduce errors related to off-topic slots by penalizing deviations from the provided instructions. Extensive experiments on two datasets, with varying model sizes (from 60M to 7B), demonstrate the superior performance of CAPID. The source code 1 is provided for reproducibility. * Equal contribution † Corresponding author. 1 https://github.com/dong7313/CAPID_ Dialogues Value of the slot <Restaurant-Name> for different slot formats Conventional Pre-defined Slot: <Restaurant-Name> Schema-driven Prompting: name of the restaurant Question-answering: What is the name of the restaurant? Their result: Bridge Guest House
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- HM3: Hierarchical Multi-Objective Model Merging for Pretrained ModelsYu Zhou, Xingyu Wu, Jibin Wu, Liang Feng 等NeurIPS 2025 · 被引用 14 次
- FOREVER: Forgetting Curve-Inspired Memory Replay for Language Model Continual LearningYujie Feng, Hao Wang, Jian Li, Xu Chu 等ACL 2026 · 被引用 3 次
- AIMMerging: Adaptive Iterative Model Merging Using Training Trajectories for Language Model Continual LearningYujie Feng, Jian Li, Xiaoyu Dong, Pengfei Xu 等EMNLP 2025
- From Schema to State: Zero-Shot Scheme-Only Dialogue State Tracking via Diverse Synthetic Dialogue and Step-by-Step DistillationHuan Xu, Zequn Li, Wen Tang, Jian Jun ZhangEMNLP 2025
- GeoEdit: Geometric Knowledge Editing for Large Language ModelsYujie Feng, Li-Ming Zhan, Zexin Lu, Yongxin Xu 等EMNLP 2025
它引用的顶会 Paper14
- A Simple Language Model for Task-Oriented DialogueEhsan Hosseini-Asl, Bryan McCann, Chien-Sheng Wu, Semih Yavuz 等NeurIPS 2020 · 被引用 590 次
- Large Language Models are Human-Level Prompt EngineersYongchao Zhou, Andrei Ioan Muresanu, Ziwen Han, Keiran Paster 等ICLR 2023 · 被引用 297 次
- Automatic Chain of Thought Prompting in Large Language ModelsZhuosheng Zhang, Aston Zhang, Mu Li, Alex SmolaICLR 2023 · 被引用 234 次
- Multi-Task Pre-Training for Plug-and-Play Task-Oriented Dialogue SystemYixuan Su, Lei Shu, Elman Mansimov, Arshit Gupta 等ACL 2022 · 被引用 218 次
- Read-only Prompt Optimization for Vision-Language Few-shot LearningDongjun Lee, Seokwon Song, Jihee Suh, Joonmyeong Choi 等ICCV 2023 · 被引用 87 次
相关 Paper
- Zero-Shot Dialogue State Tracking via Cross-Task TransferZhaojiang Lin, Bing Liu, Andrea Madotto, Seungwhan Moon 等EMNLP 2021
- Prompter: Zero-shot Adaptive Prefixes for Dialogue State Tracking Domain AdaptationIbrahim Taha Aksu, Min-Yen Kan, Nancy F. ChenACL 2023 · 被引用 3 次
- MA-DST: Multi-Attention-Based Scalable Dialog State TrackingAdarsh Kumar, Peter Ku, Anuj Kumar Goyal, Angeliki Metallinou 等AAAI 2020 · 被引用 61 次
- Dialogue State Tracking with a Language Model using Schema-Driven PromptingChia-Hsuan Lee, Hao Cheng, Mari OstendorfEMNLP 2021 · 被引用 87 次
- DiSTRICT: Dialogue State Tracking with Retriever Driven In-Context TuningPraveen Venkateswaran, Evelyn Duesterwald, Vatche IsahagianEMNLP 2023 · 被引用 7 次
