EM Pre-training for Multi-party Dialogue Response Generation
Yiyang Li, Hai Zhao
摘要
Dialogue response generation requires an agent to generate a response according to the current dialogue history, in terms of which two-party dialogues have been well studied, but leaving a great gap for multi-party dialogues at the same time. Different from two-party dialogues where each response is a direct reply to its previous utterance, the addressee of a response utterance should be specified before it is generated in the multi-party scenario. Thanks to the huge amount of two-party conversational data, various pre-trained language models for two-party dialogue response generation have been proposed. However, due to the lack of annotated addressee labels in multi-party dialogue datasets, it is hard to use them to pre-train a response generation model for multi-party dialogues. To tackle this obstacle, we propose an Expectation-Maximization (EM) approach that iteratively performs the expectation steps to generate addressee labels, and the maximization steps to optimize a response generation model. Theoretical analyses and extensive experiments have justified the feasibility and effectiveness of our proposed method. The official implementation of this paper is available at https://github.com/EricLee8/MPDRG.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- Pre-training Multi-party Dialogue Models with Latent Discourse InferenceYiyang Li, Xinting Huang, Wei Bi, Hai ZhaoACL 2023 · 被引用 3 次
- Improving Multi-party Dialogue Generation via Topic and Rhetorical CoherenceYaxin Fan, Peifeng Li, Qiaoming ZhuEMNLP 2024 · 被引用 1 次
- Discourse Coherence and Response-Guided Context Rewriting for Multi-Party Dialogue GenerationZhiyu Cao, Peifeng Li, Qiaoming ZhuACL 2026
- MADNet: Maximizing Addressee Deduction Expectation for Multi-Party Conversation GenerationJia-Chen Gu, Chao-Hong Tan, Caiyuan Chu, Zhen-Hua Ling 等EMNLP 2023
它引用的顶会 Paper8
- BART: Denoising Sequence-to-Sequence Pre-training for Natural Language Generation, Translation, and ComprehensionMike Lewis, Yinhan Liu, Naman Goyal, Marjan Ghazvininejad 等ACL 2020 · 被引用 1,224 次
- ELECTRA: Pre-training Text Encoders as Discriminators Rather Than GeneratorsKevin Clark, Minh-Thang Luong, Quoc V. Le, Christopher D. ManningICLR 2020 · 被引用 541 次
- PLATO: Pre-trained Dialogue Generation Model with Discrete Latent VariableSiqi Bao, Huang He, Fan Wang, Hua Wu 等ACL 2020 · 被引用 229 次
- Response Selection for Multi-Party Conversations with Dynamic Topic TrackingWeishi Wang, Steven C. H. Hoi, Shafiq R. JotyEMNLP 2020 · 被引用 41 次
- Multi-turn Response Selection using Dialogue Dependency RelationsQi Jia, Yizhu Liu, Siyu Ren, Kenny Q. Zhu 等EMNLP 2020 · 被引用 31 次
相关 Paper
- HeterMPC: A Heterogeneous Graph Neural Network for Response Generation in Multi-Party ConversationsJia-Chen Gu, Chao-Hong Tan, Chongyang Tao, Zhen-Hua Ling 等ACL 2022
- Multi-Party Empathetic Dialogue Generation: A New Task for Dialog SystemsLingyu Zhu, Zhengkun Zhang, Jun Wang, Hongbin Wang 等ACL 2022
- MPC-BERT: A Pre-Trained Language Model for Multi-Party Conversation UnderstandingJia-Chen Gu, Chongyang Tao, Zhen-Hua Ling, Can Xu 等ACL 2021
- A Pre-Training Based Personalized Dialogue Generation Model with Persona-Sparse DataYinhe Zheng, Rongsheng Zhang, Minlie Huang, Xiaoxi MaoAAAI 2020 · 被引用 173 次
- DialogBERT: Discourse-Aware Response Generation via Learning to Recover and Rank UtterancesXiaodong Gu, Kang Min Yoo, Jung-Woo HaAAAI 2021 · 被引用 83 次
