Improving Multi-party Dialogue Generation via Topic and Rhetorical Coherence
Yaxin Fan, Peifeng Li, Qiaoming Zhu
Abstract
Previous studies on multi-party dialogue generation predominantly concentrated on modeling the reply-to structure of dialogue histories, always overlooking the coherence between generated responses and target utterances. To address this issue, we propose a Reinforcement Learning approach emphasizing both Topic and Rhetorical Coherence (RL-TRC). In particular, the topic-and rhetorical-coherence tasks are designed to enhance the model's perception of coherence with the target utterance. Subsequently, an agent is employed to learn a coherence policy, which guides the generation of responses that are topically and rhetorically aligned with the target utterance. Furthermore, three discourse-aware rewards are developed to assess the coherence between the generated response and the target utterance, with the objective of optimizing the policy. The experimental results and in-depth analyses on two popular datasets demonstrate that our RL-TRC significantly outperforms the state-of-the-art baselines, particularly in generating responses that are more coherent with the target utterances.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 5baf5ae2-cab8-4093-955e-d3b1484f1d9fCited by top-tier papers2
- Enabling Chatbots with Eyes and Ears: An Immersive Multimodal Conversation System for Dynamic InteractionsJihyoung Jang, Minwook Bae, Minji Kim, Dilek Hakkani-Tür et al.ACL 2025
- Discourse Coherence and Response-Guided Context Rewriting for Multi-Party Dialogue GenerationZhiyu Cao, Peifeng Li, Qiaoming ZhuACL 2026
Builds on9
- BART: Denoising Sequence-to-Sequence Pre-training for Natural Language Generation, Translation, and ComprehensionMike Lewis, Yinhan Liu, Naman Goyal, Marjan Ghazvininejad et al.ACL 2020 · 1,224 citations
- AlpacaFarm: A Simulation Framework for Methods that Learn from Human FeedbackYann Dubois, Chen Xuechen Li, Rohan Taori, Tianyi Zhang et al.NeurIPS 2023 · 948 citations
- Zero-shot Text Classification via Reinforced Self-trainingZhiquan Ye, Yuxia Geng, Jiaoyan Chen, Jingmin Chen et al.ACL 2020 · 75 citations
- Learning to Memorize Entailment and Discourse Relations for Persona-Consistent DialoguesRuijun Chen, Jin Wang, Liang-Chih Yu, Xuejie ZhangAAAI 2023 · 32 citations
- EM Pre-training for Multi-party Dialogue Response GenerationYiyang Li, Hai ZhaoACL 2023 · 9 citations
Related papers
- Knowledge Graph Grounded Goal Planning for Open-Domain Conversation GenerationJun Xu, Haifeng Wang, Zhengyu Niu, Hua Wu et al.AAAI 2020 · 69 citations
- Improving Dialogue Discourse Parsing via Reply-to Structures of Addressee RecognitionYaxin Fan, Feng Jiang, Peifeng Li, Fang Kong et al.EMNLP 2023 · 4 citations
- R4: Nested Reasoning-Retrieval for Reward Modeling in Role-Playing AgentsRenzhi Wang, Chongqiang Wei, Zhisheng Wang, Piji LiICLR 2026
- DialogBERT: Discourse-Aware Response Generation via Learning to Recover and Rank UtterancesXiaodong Gu, Kang Min Yoo, Jung-Woo HaAAAI 2021 · 83 citations
- Hierarchical Reinforcement Learning for Open-Domain DialogAbdelrhman Saleh, Natasha Jaques, Asma Ghandeharioun, Judy Hanwen Shen et al.AAAI 2020 · 60 citations
