Enhancing Dialogue State Tracking Models through LLM-backed User-Agents Simulation
Cheng Niu, Xingguang Wang, Xuxin Cheng, Juntong Song, Tong Zhang
摘要
Dialogue State Tracking (DST) is designed to monitor the evolving dialogue state in the conversations and plays a pivotal role in developing task-oriented dialogue systems. However, obtaining the annotated data for the DST task is usually a costly endeavor. In this paper, we focus on employing LLMs to generate dialogue data to reduce dialogue collection and annotation costs. Specifically, GPT-4 is used to simulate the user and agent interaction, generating thousands of dialogues annotated with DST labels. Then a two-stage fine-tuning on LLaMA 2 is performed on the generated data and the real data for the DST prediction. Experimental results on two public DST benchmarks show that with the generated dialogue data, our model performs better than the baseline trained solely on real data. In addition, our approach is also capable of adapting to the dynamic demands in real-world scenarios, generating dialogues in new domains swiftly. After replacing dialogue segments in any domain with the corresponding generated ones, the model achieves comparable performance to the model trained on real data 1 .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper7
- Synergistic Interplay between Search and Large Language Models for Information RetrievalJiazhan Feng, Chongyang Tao, Xiubo Geng, Tao Shen 等ACL 2024 · 被引用 7 次
- OnGoal: Tracking and Visualizing Conversational Goals in Multi-Turn Dialogue with Large Language ModelsAdam Coscia, Shunan Guo, Eunyee Koh, Alex EndertUIST 2025 · 被引用 6 次
- From Daily Song to Daily Self: Supporting Emotional Growth of Deaf and Hard-of-Hearing Individuals through Generative AI SongwritingYoujin Choi, Jinyoung Yoo, JaeYoung Moon, Yoonjae Kim 等CHI 2026 · 被引用 2 次
- Can Large Language Models Keep Up? Benchmarking Online Adaptation to Continual Knowledge StreamsJiyeon Kim, Hyunji Lee, Dylan Zhou, Sue Hyun Park 等ACL 2026 · 被引用 1 次
- SQLWOZ: A Realistic Task-Oriented Dialogue Dataset with SQL-Based Dialogue State Representation for Complex User RequirementsHeng-Da Xu, Xian-Ling Mao, Fanshu Sun, Tian-Yi Che 等EMNLP 2025
它引用的顶会 Paper12
- LoRA: Low-Rank Adaptation of Large Language ModelsEdward J. Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu 等ICLR 2022 · 被引用 18,833 次
- HuggingGPT: Solving AI Tasks with ChatGPT and its Friends in Hugging FaceYongliang Shen, Kaitao Song, Xu Tan, Dongsheng Li 等NeurIPS 2023 · 被引用 1,778 次
- Efficient Memory Management for Large Language Model Serving with PagedAttentionWoosuk Kwon, Zhuohan Li, Siyuan Zhuang, Ying Sheng 等SOSP 2023 · 被引用 1,016 次
- Towards Scalable Multi-Domain Conversational Agents: The Schema-Guided Dialogue DatasetAbhinav Rastogi, Xiaoxue Zang, Srinivas Sunkara, Raghav Gupta 等AAAI 2020 · 被引用 707 次
- Schema-Guided Multi-Domain Dialogue State Tracking with Graph Attention Neural NetworksLu Chen, Boer Lv, Chi Wang, Su Zhu 等AAAI 2020 · 被引用 143 次
相关 Paper
- Towards LLM-driven Dialogue State TrackingYujie Feng, Zexin Lu, Bo Liu, Liming Zhan 等EMNLP 2023 · 被引用 25 次
- Large Language Models as Zero-shot Dialogue State Tracker through Function CallingZekun Li, Zhiyu Chen, Mike Ross, Patrick Huber 等ACL 2024 · 被引用 9 次
- A Dual Prompt Learning Framework for Few-Shot Dialogue State TrackingYuting Yang, Wenqiang Lei, Pei Huang, Juan Cao 等WWW 2023 · 被引用 20 次
- Turn-Level Active Learning for Dialogue State TrackingZihan Zhang, Meng Fang, Fanghua Ye, Ling Chen 等EMNLP 2023 · 被引用 3 次
- DiSTRICT: Dialogue State Tracking with Retriever Driven In-Context TuningPraveen Venkateswaran, Evelyn Duesterwald, Vatche IsahagianEMNLP 2023 · 被引用 7 次
