MemSum: Extractive Summarization of Long Documents Using Multi-Step Episodic Markov Decision Processes
Nianlong Gu, Elliott Ash, Richard H. R. Hahnloser
Abstract
We introduce MemSum (Multi-step Episodic Markov decision process extractive SUMmarizer), a reinforcement-learning-based extractive summarizer enriched at each step with information on the current extraction history. When MemSum iteratively selects sentences into the summary, it considers a broad information set that would intuitively also be used by humans in this task: 1) the text content of the sentence, 2) the global text context of the rest of the document, and 3) the extraction history consisting of the set of sentences that have already been extracted. With a lightweight architecture, MemSum obtains state-of-the-art test-set performance (ROUGE) in summarizing long documents taken from PubMed, arXiv, and GovReport. Ablation studies demonstrate the importance of local, global, and history information. A human evaluation confirms the high quality and low redundancy of the generated summaries, stemming from MemSum's awareness of extraction history.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 50e98e74-a984-4721-b1ee-6c8bc6b2ccf2Cited by top-tier papers6
- Text summarization via global structure awarenessJiaquan Zhang, Chaoning Zhang, Shuxu Chen, Yibei Liu et al.ICLR 2026 · 14 citations
- Record Once, Post Everywhere: Automatic Shortening of Audio Stories for Social MediaBryan Wang, Zeyu Jin, Gautham J. MysoreUIST 2022 · 13 citations
- Unsupervised Extractive Summarization with Learnable Length Control StrategiesRenlong Jie, Xiaojun Meng, Xin Jiang, Qun LiuAAAI 2024 · 8 citations
- TalkLess: Blending Extractive and Abstractive Summarization for Editing Speech to Preserve Content and StyleKarim Benharrak, Puyuan Peng, Amy PavelUIST 2025 · 1 citation
- Deep Submodular Optimization and LLM for Multimodal Content Extraction and Automatic Poster Generation from Long DocumentVijay Jaisankar, Sambaran Bandyopadhyay, Kalp Vyas, Varre Suman Chaitanya et al.AAAI 2025 · 1 citation
Builds on4
- Big Bird: Transformers for Longer SequencesManzil Zaheer, Guru Guruganesh, Kumar Avinava Dubey, Joshua Ainslie et al.NeurIPS 2020 · 3,159 citations
- Reformer: The Efficient TransformerNikita Kitaev, Lukasz Kaiser, Anselm LevskayaICLR 2020 · 2,878 citations
- BART: Denoising Sequence-to-Sequence Pre-training for Natural Language Generation, Translation, and ComprehensionMike Lewis, Yinhan Liu, Naman Goyal, Marjan Ghazvininejad et al.ACL 2020 · 1,224 citations
- On Extractive and Abstractive Neural Document Summarization with Transformer Language ModelsJonathan Pilault, Raymond Li, Sandeep Subramanian, Chris PalEMNLP 2020 · 186 citations
Related papers
- Preserve Context Information for Extract-Generate Long-Input Summarization FrameworkRuifeng Yuan, Zili Wang, Ziqiang Cao, Wenjie LiAAAI 2023 · 3 citations
- Contextualized Rewriting for Text SummarizationGuangsheng Bao, Yue ZhangAAAI 2021 · 17 citations
- Copy or Rewrite: Hybrid Summarization with Hierarchical Reinforcement LearningLiqiang Xiao, Lu Wang, Hao He, Yaohui JinAAAI 2020 · 29 citations
- Keyword-aware Abstractive Summarization by Extracting Set-level Intermediate SummariesYizhu Liu, Qi Jia, Kenny Q. ZhuWWW 2021 · 14 citations
- DYLE: Dynamic Latent Extraction for Abstractive Long-Input SummarizationZiming Mao, Chen Henry Wu, Ansong Ni, Yusen Zhang et al.ACL 2022 · 62 citations
