MADNet: Maximizing Addressee Deduction Expectation for Multi-Party Conversation Generation
Jia-Chen Gu, Chao-Hong Tan, Caiyuan Chu, Zhen-Hua Ling, Chongyang Tao, Quan Liu, Cong Liu
摘要
Modeling multi-party conversations (MPCs) with graph neural networks has been proven effective at capturing complicated and graphical information flows. However, existing methods rely heavily on the necessary addressee labels and can only be applied to an ideal setting where each utterance must be tagged with an “@” or other equivalent addressee label. To study the scarcity of addressee labels which is a common issue in MPCs, we propose MADNet that maximizes addressee deduction expectation in heterogeneous graph neural networks for MPC generation. Given an MPC with a few addressee labels missing, existing methods fail to build a consecutively connected conversation graph, but only a few separate conversation fragments instead. To ensure message passing between these conversation fragments, four additional types of latent edges are designed to complete a fully-connected graph. Besides, to optimize the edge-type-dependent message passing for those utterances without addressee labels, an Expectation-Maximization-based method that iteratively generates silver addressee labels (E step), and optimizes the quality of generated responses (M step), is designed. Experimental results on two Ubuntu IRC channel benchmarks show that MADNet outperforms various baseline models on the task of MPC generation, especially under the more common and challenging setting where part of addressee labels are missing.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- Improving Multi-party Dialogue Generation via Topic and Rhetorical CoherenceYaxin Fan, Peifeng Li, Qiaoming ZhuEMNLP 2024 · 被引用 1 次
- Enabling Chatbots with Eyes and Ears: An Immersive Multimodal Conversation System for Dynamic InteractionsJihyoung Jang, Minwook Bae, Minji Kim, Dilek Hakkani-Tür 等ACL 2025
- Discourse Coherence and Response-Guided Context Rewriting for Multi-Party Dialogue GenerationZhiyu Cao, Peifeng Li, Qiaoming ZhuACL 2026
它引用的顶会 Paper6
- BART: Denoising Sequence-to-Sequence Pre-training for Natural Language Generation, Translation, and ComprehensionMike Lewis, Yinhan Liu, Naman Goyal, Marjan Ghazvininejad 等ACL 2020 · 被引用 1,224 次
- Response Selection for Multi-Party Conversations with Dynamic Topic TrackingWeishi Wang, Steven C. H. Hoi, Shafiq R. JotyEMNLP 2020 · 被引用 41 次
- EM Pre-training for Multi-party Dialogue Response GenerationYiyang Li, Hai ZhaoACL 2023 · 被引用 9 次
- An Equal-Size Hard EM Algorithm for Diverse Dialogue GenerationYuqiao Wen, Yongchang Hao, Yanshuai Cao, Lili MouICLR 2023 · 被引用 2 次
- MPC-BERT: A Pre-Trained Language Model for Multi-Party Conversation UnderstandingJia-Chen Gu, Chongyang Tao, Zhen-Hua Ling, Can Xu 等ACL 2021
相关 Paper
- HeterMPC: A Heterogeneous Graph Neural Network for Response Generation in Multi-Party ConversationsJia-Chen Gu, Chao-Hong Tan, Chongyang Tao, Zhen-Hua Ling 等ACL 2022
- Do LLMs suffer from Multi-Party Hangover? A Diagnostic Approach to Addressee Recognition and Response Selection in ConversationsNicolò Penzo, Maryam Sajedinia, Bruno Lepri, Sara Tonelli 等EMNLP 2024 · 被引用 2 次
- GIFT: Graph-Induced Fine-Tuning for Multi-Party Conversation UnderstandingJia-Chen Gu, Zhenhua Ling, Quan Liu, Cong Liu 等ACL 2023 · 被引用 3 次
- Pre-training Multi-party Dialogue Models with Latent Discourse InferenceYiyang Li, Xinting Huang, Wei Bi, Hai ZhaoACL 2023 · 被引用 3 次
- Infusing Multi-Source Knowledge with Heterogeneous Graph Neural Network for Emotional Conversation GenerationYunlong Liang, Fandong Meng, Ying Zhang, Yufeng Chen 等AAAI 2021 · 被引用 62 次
