MemSum: Extractive Summarization of Long Documents Using Multi-Step Episodic Markov Decision Processes
Nianlong Gu, Elliott Ash, Richard H. R. Hahnloser
摘要
We introduce MemSum (Multi-step Episodic Markov decision process extractive SUMmarizer), a reinforcement-learning-based extractive summarizer enriched at each step with information on the current extraction history. When MemSum iteratively selects sentences into the summary, it considers a broad information set that would intuitively also be used by humans in this task: 1) the text content of the sentence, 2) the global text context of the rest of the document, and 3) the extraction history consisting of the set of sentences that have already been extracted. With a lightweight architecture, MemSum obtains state-of-the-art test-set performance (ROUGE) in summarizing long documents taken from PubMed, arXiv, and GovReport. Ablation studies demonstrate the importance of local, global, and history information. A human evaluation confirms the high quality and low redundancy of the generated summaries, stemming from MemSum's awareness of extraction history.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- Text summarization via global structure awarenessJiaquan Zhang, Chaoning Zhang, Shuxu Chen, Yibei Liu 等ICLR 2026 · 被引用 14 次
- Record Once, Post Everywhere: Automatic Shortening of Audio Stories for Social MediaBryan Wang, Zeyu Jin, Gautham J. MysoreUIST 2022 · 被引用 13 次
- Unsupervised Extractive Summarization with Learnable Length Control StrategiesRenlong Jie, Xiaojun Meng, Xin Jiang, Qun LiuAAAI 2024 · 被引用 8 次
- TalkLess: Blending Extractive and Abstractive Summarization for Editing Speech to Preserve Content and StyleKarim Benharrak, Puyuan Peng, Amy PavelUIST 2025 · 被引用 1 次
- Deep Submodular Optimization and LLM for Multimodal Content Extraction and Automatic Poster Generation from Long DocumentVijay Jaisankar, Sambaran Bandyopadhyay, Kalp Vyas, Varre Suman Chaitanya 等AAAI 2025 · 被引用 1 次
它引用的顶会 Paper4
- Big Bird: Transformers for Longer SequencesManzil Zaheer, Guru Guruganesh, Kumar Avinava Dubey, Joshua Ainslie 等NeurIPS 2020 · 被引用 3,159 次
- Reformer: The Efficient TransformerNikita Kitaev, Lukasz Kaiser, Anselm LevskayaICLR 2020 · 被引用 2,878 次
- BART: Denoising Sequence-to-Sequence Pre-training for Natural Language Generation, Translation, and ComprehensionMike Lewis, Yinhan Liu, Naman Goyal, Marjan Ghazvininejad 等ACL 2020 · 被引用 1,224 次
- On Extractive and Abstractive Neural Document Summarization with Transformer Language ModelsJonathan Pilault, Raymond Li, Sandeep Subramanian, Chris PalEMNLP 2020 · 被引用 186 次
相关 Paper
- Preserve Context Information for Extract-Generate Long-Input Summarization FrameworkRuifeng Yuan, Zili Wang, Ziqiang Cao, Wenjie LiAAAI 2023 · 被引用 3 次
- Contextualized Rewriting for Text SummarizationGuangsheng Bao, Yue ZhangAAAI 2021 · 被引用 17 次
- Copy or Rewrite: Hybrid Summarization with Hierarchical Reinforcement LearningLiqiang Xiao, Lu Wang, Hao He, Yaohui JinAAAI 2020 · 被引用 29 次
- Keyword-aware Abstractive Summarization by Extracting Set-level Intermediate SummariesYizhu Liu, Qi Jia, Kenny Q. ZhuWWW 2021 · 被引用 14 次
- DYLE: Dynamic Latent Extraction for Abstractive Long-Input SummarizationZiming Mao, Chen Henry Wu, Ansong Ni, Yusen Zhang 等ACL 2022 · 被引用 62 次
