Delving into Global Dialogue Structures: Structure Planning Augmented Response Selection for Multi-turn Conversations
Tingchen Fu, Xueliang Zhao, Rui Yan
Abstract
Retrieval-based dialogue systems are a crucial component of natural language processing, employing information retrieval techniques to select responses from a predefined pool of candidates. The advent of pre-trained language models (PLMs) has significantly advanced the field, with a prevailing paradigm that involves post-training PLMs on specific dialogue corpora, followed by fine-tuning for the response selection (RS) task. This post-training process aims to capture dialogue-specific features, as most PLMs are originally trained on plain text. However, prior approaches predominantly rely on self-supervised tasks or session-level graph neural networks during post-training, focusing on capturing underlying patterns of coherent dialogues without explicitly refining the global pattern across the entire dialogue corpus. Consequently, the learned knowledge for organizing coherent dialogues remains isolated, heavily reliant on specific contexts. Additionally, interpreting or visualizing the implicit knowledge acquired through self-supervised tasks proves challenging. In this study, we address these limitations by explicitly refining the knowledge required for response selection and structuring it into a coherent global flow, known as "dialogue structure." This structure captures the inter-dependency of utterances and topic shifts, thereby enhancing the response selection task. To achieve this, we propose a novel structure model comprising a state recognizer and a structure planner. This model effectively captures the flow within the utterance history and plans the trajectory of future utterances. Importantly, the structure model operates orthogonally to the retrieval model, enabling seamless integration with existing retrieval models and facilitating collaborative training. Extensive experiments conducted on three benchmark datasets demonstrate the superior performance of our method over a wide range of competitive baselines, establishing a new state-of-the-art in the field.
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Your agent calls
Lunesearch_papers
Free to start. No credit card required.
Terminal
Install the CLIlune papers get 0a6aee2e-b501-4613-81e1-d61ee4d29522Cited by top-tier papers1
Ask how each one uses itRelated papers
- Learning an Effective Context-Response Matching Model with Self-Supervised Tasks for Retrieval-based DialoguesRuijian Xu, Chongyang Tao, Daxin Jiang, Xueliang Zhao et al.AAAI 2021 · 76 citations
- A Pre-training Strategy for Zero-Resource Response Selection in Knowledge-Grounded ConversationsChongyang Tao, Changyu Chen, Jiazhan Feng, Ji-Rong Wen et al.ACL 2021
- Knowledge-Grounded Dialogue Generation with Pre-trained Language ModelsXueliang Zhao, Wei Wu, Can Xu, Chongyang Tao et al.EMNLP 2020 · 153 citations
- Structural Pre-training for Dialogue ComprehensionZhuosheng Zhang, Hai ZhaoACL 2021
- Generative Subgraph Retrieval for Knowledge Graph-Grounded Dialog GenerationJinyoung Park, Minseok Joo, Joo-Kyung Kim, Hyunwoo J. KimEMNLP 2024 · 3 citations
