Multi-document Summarization with Maximal Marginal Relevance-guided Reinforcement Learning
Yuning Mao, Yanru Qu, Yiqing Xie, Xiang Ren, Jiawei Han
摘要
While neural sequence learning methods have made significant progress in single-document summarization (SDS), they produce unsatisfactory results on multi-document summarization (MDS). We observe two major challenges when adapting SDS advances to MDS: (1) MDS involves larger search space and yet more limited training data, setting obstacles for neural methods to learn adequate representations; (2) MDS needs to resolve higher information redundancy among the source documents, which SDS methods are less effective to handle. To close the gap, we present RL-MMR, Maximal Margin Relevance-guided Reinforcement Learning for MDS, which unifies advanced neural SDS methods and statistical measures used in classical MDS. RL-MMR casts MMR guidance on fewer promising candidates, which restrains the search space and thus leads to better representation learning. Additionally, the explicit redundancy measure in MMR helps the neural representation of the summary to better capture redundancy. Extensive experiments demonstrate that RL-MMR achieves state-of-the-art performance on benchmark MDS datasets. In particular, we show the benefits of incorporating MMR into end-to-end learning when adapting SDS to MDS in terms of both learning effectiveness and efficiency. 1
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper10
- PRIMERA: Pyramid-based Masked Sentence Pre-training for Multi-document SummarizationWen Xiao, Iz Beltagy, Giuseppe Carenini, Arman CohanACL 2022 · 被引用 147 次
- A Multi-Document Coverage Reward for RELAXed Multi-Document SummarizationJacob Parnell, Inigo Jauregi Unanue, Massimo PiccardiACL 2022 · 被引用 16 次
- MDCure: A Scalable Pipeline for Multi-Document Instruction-FollowingGabrielle Kaili-May Liu, Bowen Shi, Avi Caciularu, Idan Szpektor 等ACL 2025 · 被引用 13 次
- PDSum: Prototype-driven Continuous Summarization of Evolving Multi-document Sets StreamSusik Yoon, Hou Pong Chan, Jiawei HanWWW 2023 · 被引用 13 次
- CiteSum: Citation Text-guided Scientific Extreme Summarization and Domain Adaptation with Limited SupervisionYuning Mao, Ming Zhong, Jiawei HanEMNLP 2022 · 被引用 11 次
它引用的顶会 Paper4
- PEGASUS: Pre-training with Extracted Gap-sentences for Abstractive SummarizationJingqing Zhang, Yao Zhao, Mohammad Saleh, Peter J. LiuICML 2020 · 被引用 2,453 次
- BART: Denoising Sequence-to-Sequence Pre-training for Natural Language Generation, Translation, and ComprehensionMike Lewis, Yinhan Liu, Naman Goyal, Marjan Ghazvininejad 等ACL 2020 · 被引用 1,224 次
- Generating Representative Headlines for News StoriesXiaotao Gu, Yuning Mao, Jiawei Han, Jialu Liu 等WWW 2020 · 被引用 77 次
- Facet-Aware Evaluation for Extractive SummarizationYuning Mao, Liyuan Liu, Qi Zhu, Xiang Ren 等ACL 2020 · 被引用 19 次
相关 Paper
- SgSum: Transforming Multi-document Summarization into Sub-graph SelectionMoye Chen, Wei Li, Jiachen Liu, Xinyan Xiao 等EMNLP 2021 · 被引用 21 次
- Content- and Topology-Aware Representation Learning for Scientific Multi-LiteratureKai Zhang, Kaisong Song, Yangyang Kang, Xiaozhong LiuEMNLP 2023
- A Unified Retrieval Framework with Document Ranking and EDU Filtering for Multi-document SummarizationShiyin Tan, Jaeeon Park, Dongyuan Li, Renhe Jiang 等SIGIR 2025
- Leveraging Graph to Improve Abstractive Multi-Document SummarizationWei Li, Xinyan Xiao, Jiachen Liu, Hua Wu 等ACL 2020 · 被引用 118 次
- Be Relevant, Non-Redundant, and Timely: Deep Reinforcement Learning for Real-Time Event SummarizationMin Yang, Chengming Li, Fei Sun, Zhou Zhao 等AAAI 2020 · 被引用 8 次
