Compressive Summarization with Plausibility and Salience Modeling
Shrey Desai, Jiacheng Xu, Greg Durrett
Abstract
Compressive summarization systems typically rely on a crafted set of syntactic rules to determine what spans of possible summary sentences can be deleted, then learn a model of what to actually delete by optimizing for content selection (ROUGE). In this work, we propose to relax the rigid syntactic constraints on candidate spans and instead leave compression decisions to two data-driven criteria: plausibility and salience. Deleting a span is plausible if removing it maintains the grammaticality and factuality of a sentence, and spans are salient if they contain important information from the summary. Each of these is judged by a pre-trained Transformer model, and only deletions that are both plausible and not salient can be applied. When integrated into a simple extraction-compression pipeline, our method achieves strong in-domain results on benchmark summarization datasets, and human evaluation shows that the plausibility model generally selects for grammatical and factual deletions. Furthermore, the flexibility of our approach allows it to generalize cross-domain: our system fine-tuned on only 500 samples from a new domain can match or exceed an in-domain extractive model trained on much more data. 1
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext c7a5e258-10ad-4667-852f-41e508177414Cited by top-tier papers4
- An AI-Resilient Text Rendering Technique for Reading and Skimming DocumentsZiwei Gu, Ian Arawjo, Kenneth Li, Jonathan K. Kummerfeld et al.CHI 2024 · 31 citations
- BASS: Boosting Abstractive Summarization with Unified Semantic GraphWenhao Wu, Wei Li, Xinyan Xiao, Jiachen Liu et al.ACL 2021
- Dissecting Generation Modes for Abstractive Summarization Models via Ablation and AttributionJiacheng Xu, Greg DurrettACL 2021
- ConvoSumm: Conversation Summarization Benchmark and Improved Abstractive Summarization with Argument MiningAlexander R. Fabbri, Faiaz Rahman, Imad Rizvi, Borui Wang et al.ACL 2021
Builds on7
- PEGASUS: Pre-training with Extracted Gap-sentences for Abstractive SummarizationJingqing Zhang, Yao Zhao, Mohammad Saleh, Peter J. LiuICML 2020 · 2,453 citations
- ELECTRA: Pre-training Text Encoders as Discriminators Rather Than GeneratorsKevin Clark, Minh-Thang Luong, Quoc V. Le, Christopher D. ManningICLR 2020 · 541 citations
- Extractive Summarization as Text MatchingMing Zhong, Pengfei Liu, Yiran Chen, Danqing Wang et al.ACL 2020 · 410 citations
- Asking and Answering Questions to Evaluate the Factual Consistency of SummariesAlex Wang, Kyunghyun Cho, Mike LewisACL 2020 · 317 citations
- Discourse-Aware Neural Extractive Text SummarizationJiacheng Xu, Zhe Gan, Yu Cheng, Jingjing LiuACL 2020 · 264 citations
Related papers
- Syntactically Look-Ahead Attention Network for Sentence CompressionHidetaka Kamigaito, Manabu OkumuraAAAI 2020 · 22 citations
- Learning with Rejection for Abstractive Text SummarizationMeng Cao, Yue Dong, Jingyi He, Jackie Chi Kit CheungEMNLP 2022 · 10 citations
- Multi-Fact Correction in Abstractive Text SummarizationYue Dong, Shuohang Wang, Zhe Gan, Yu Cheng et al.EMNLP 2020 · 99 citations
- RECOMP: Improving Retrieval-Augmented LMs with Context Compression and Selective AugmentationFangyuan Xu, Weijia Shi, Eunsol ChoiICLR 2024 · 260 citations
- Summarization Programs: Interpretable Abstractive Summarization with Neural Modular TreesSwarnadeep Saha, Shiyue Zhang, Peter Hase, Mohit BansalICLR 2023 · 7 citations
