Improving Multi-party Dialogue Generation via Topic and Rhetorical Coherence
Yaxin Fan, Peifeng Li, Qiaoming Zhu
摘要
Previous studies on multi-party dialogue generation predominantly concentrated on modeling the reply-to structure of dialogue histories, always overlooking the coherence between generated responses and target utterances. To address this issue, we propose a Reinforcement Learning approach emphasizing both Topic and Rhetorical Coherence (RL-TRC). In particular, the topic-and rhetorical-coherence tasks are designed to enhance the model's perception of coherence with the target utterance. Subsequently, an agent is employed to learn a coherence policy, which guides the generation of responses that are topically and rhetorically aligned with the target utterance. Furthermore, three discourse-aware rewards are developed to assess the coherence between the generated response and the target utterance, with the objective of optimizing the policy. The experimental results and in-depth analyses on two popular datasets demonstrate that our RL-TRC significantly outperforms the state-of-the-art baselines, particularly in generating responses that are more coherent with the target utterances.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Enabling Chatbots with Eyes and Ears: An Immersive Multimodal Conversation System for Dynamic InteractionsJihyoung Jang, Minwook Bae, Minji Kim, Dilek Hakkani-Tür 等ACL 2025
- Discourse Coherence and Response-Guided Context Rewriting for Multi-Party Dialogue GenerationZhiyu Cao, Peifeng Li, Qiaoming ZhuACL 2026
它引用的顶会 Paper9
- BART: Denoising Sequence-to-Sequence Pre-training for Natural Language Generation, Translation, and ComprehensionMike Lewis, Yinhan Liu, Naman Goyal, Marjan Ghazvininejad 等ACL 2020 · 被引用 1,224 次
- AlpacaFarm: A Simulation Framework for Methods that Learn from Human FeedbackYann Dubois, Chen Xuechen Li, Rohan Taori, Tianyi Zhang 等NeurIPS 2023 · 被引用 948 次
- Zero-shot Text Classification via Reinforced Self-trainingZhiquan Ye, Yuxia Geng, Jiaoyan Chen, Jingmin Chen 等ACL 2020 · 被引用 75 次
- Learning to Memorize Entailment and Discourse Relations for Persona-Consistent DialoguesRuijun Chen, Jin Wang, Liang-Chih Yu, Xuejie ZhangAAAI 2023 · 被引用 32 次
- EM Pre-training for Multi-party Dialogue Response GenerationYiyang Li, Hai ZhaoACL 2023 · 被引用 9 次
相关 Paper
- Knowledge Graph Grounded Goal Planning for Open-Domain Conversation GenerationJun Xu, Haifeng Wang, Zhengyu Niu, Hua Wu 等AAAI 2020 · 被引用 69 次
- Improving Dialogue Discourse Parsing via Reply-to Structures of Addressee RecognitionYaxin Fan, Feng Jiang, Peifeng Li, Fang Kong 等EMNLP 2023 · 被引用 4 次
- R4: Nested Reasoning-Retrieval for Reward Modeling in Role-Playing AgentsRenzhi Wang, Chongqiang Wei, Zhisheng Wang, Piji LiICLR 2026
- DialogBERT: Discourse-Aware Response Generation via Learning to Recover and Rank UtterancesXiaodong Gu, Kang Min Yoo, Jung-Woo HaAAAI 2021 · 被引用 83 次
- Hierarchical Reinforcement Learning for Open-Domain DialogAbdelrhman Saleh, Natasha Jaques, Asma Ghandeharioun, Judy Hanwen Shen 等AAAI 2020 · 被引用 60 次
