GIFT: Graph-Induced Fine-Tuning for Multi-Party Conversation Understanding
Jia-Chen Gu, Zhenhua Ling, Quan Liu, Cong Liu, Guoping Hu
摘要
Addressing the issues of who saying what to whom in multi-party conversations (MPCs) has recently attracted a lot of research attention. However, existing methods on MPC understanding typically embed interlocutors and utterances into sequential information flows, or utilize only the superficial of inherent graph structures in MPCs. To this end, we present a plug-and-play and lightweight method named graph-induced fine-tuning (GIFT) which can adapt various Transformer-based pre-trained language models (PLMs) for universal MPC understanding. In detail, the full and equivalent connections among utterances in regular Transformer ignore the sparse but distinctive dependency of an utterance on another in MPCs. To distinguish different relationships between utterances, four types of edges are designed to integrate graph-induced signals into attention mechanisms to refine PLMs originally designed for processing sequential texts. We evaluate GIFT by implementing it into three PLMs, and test the performance on three downstream tasks including addressee recognition, speaker identification and response selection. Experimental results show that GIFT can significantly improve the performance of three PLMs on three downstream tasks and two benchmarks with only 4 additional parameters per encoding layer, achieving new state-of-the-art performance on MPC understanding.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- Friends-MMC: A Dataset for Multi-modal Multi-party Conversation UnderstandingYueqian Wang, Xiaojun Meng, Yuxuan Wang, Jianxin Liang 等AAAI 2025 · 被引用 6 次
- Do LLMs suffer from Multi-Party Hangover? A Diagnostic Approach to Addressee Recognition and Response Selection in ConversationsNicolò Penzo, Maryam Sajedinia, Bruno Lepri, Sara Tonelli 等EMNLP 2024 · 被引用 2 次
- MISP-Meeting: A Real-World Dataset with Multimodal Cues for Long-form Meeting Transcription and SummarizationHang Chen, Chao-Han Huck Yang, Jia-Chen Gu, Sabato Marco Siniscalchi 等ACL 2025
它引用的顶会 Paper4
- Response Selection for Multi-Party Conversations with Dynamic Topic TrackingWeishi Wang, Steven C. H. Hoi, Shafiq R. JotyEMNLP 2020 · 被引用 41 次
- Unsupervised Conversation Disentanglement through Co-TrainingHui Liu, Zhan Shi, Xiaodan ZhuEMNLP 2021 · 被引用 20 次
- MPC-BERT: A Pre-Trained Language Model for Multi-Party Conversation UnderstandingJia-Chen Gu, Chongyang Tao, Zhen-Hua Ling, Can Xu 等ACL 2021
- HeterMPC: A Heterogeneous Graph Neural Network for Response Generation in Multi-Party ConversationsJia-Chen Gu, Chao-Hong Tan, Chongyang Tao, Zhen-Hua Ling 等ACL 2022
相关 Paper
- GIFT: Guided Fine-Tuning and Transfer for Enhancing Instruction-Tuned Language ModelsZhiwen Ruan, Yichao Du, Jianjie Zheng, Longyue Wang 等ACL 2026
- MADNet: Maximizing Addressee Deduction Expectation for Multi-Party Conversation GenerationJia-Chen Gu, Chao-Hong Tan, Caiyuan Chu, Zhen-Hua Ling 等EMNLP 2023
- SLASH the Sink: Sharpening Structural Attention Inside LLMsYiming Liu, Bin Lu, Xinbing Wang, Chenghu Zhou 等ICML 2026
- EM Pre-training for Multi-party Dialogue Response GenerationYiyang Li, Hai ZhaoACL 2023 · 被引用 9 次
- GiLT: Augmenting Transformer Language Models with Dependency GraphsTianyu Huang, Yida Zhao, Chuyan Zhou, Kewei TuACL 2026
