Stepwise Extractive Summarization and Planning with Structured Transformers
Shashi Narayan, Joshua Maynez, Jakub Adámek, Daniele Pighin, Blaz Bratanic, Ryan T. McDonald
Abstract
We propose encoder-centric stepwise models for extractive summarization using structured transformers -HiBERT (Zhang et al., 2019) and Extended Transformers (Ainslie et al., 2020). We enable stepwise summarization by injecting the previously generated summary into the structured transformer as an auxiliary sub-structure. Our models are not only efficient in modeling the structure of long inputs, but they also do not rely on task-specific redundancy-aware modeling, making them a general purpose extractive content planner for different tasks. When evaluated on CNN/DailyMail extractive summarization, stepwise models achieve state-of-the-art performance in terms of Rouge without any redundancy aware modeling or sentence filtering. This also holds true for Rotowire tableto-text generation, where our models surpass previously reported metrics for content selection, planning and ordering, highlighting the strength of stepwise modeling. Amongst the two structured transformers we test, stepwise Extended Transformers provides the best performance across both datasets and sets a new standard for these challenges. 1 * Equal contribution.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext bfa53c64-bd81-4f56-9871-03b6cd2ce39dCited by top-tier papers5
- Enriching and Controlling Global Semantics for Text SummarizationThong Nguyen, Anh Tuan Luu, Truc Lu, Tho QuanEMNLP 2021 · 26 citations
- Toward Unifying Text Segmentation and Long Document SummarizationSangwoo Cho, Kaiqiang Song, Xiaoyang Wang, Fei Liu et al.EMNLP 2022 · 19 citations
- LiSum: Open Source Software License Summarization with Multi-Task LearningLinyu Li, Sihan Xu, Yang Liu, Ya Gao et al.ASE 2023 · 4 citations
- Deep Differential Amplifier for Extractive SummarizationRuipeng Jia, Yanan Cao, Fang Fang, Yuchen Zhou et al.ACL 2021
- Long-Span Summarization via Local Attention and Content SelectionPotsawee Manakul, Mark J. F. GalesACL 2021
Builds on4
- Reformer: The Efficient TransformerNikita Kitaev, Lukasz Kaiser, Anselm LevskayaICLR 2020 · 2,878 citations
- Compressive Transformers for Long-Range Sequence ModellingJack W. Rae, Anna Potapenko, Siddhant M. Jayakumar, Chloe Hillier et al.ICLR 2020 · 833 citations
- Heterogeneous Graph Neural Networks for Extractive Document SummarizationDanqing Wang, Pengfei Liu, Yining Zheng, Xipeng Qiu et al.ACL 2020 · 275 citations
- Copy or Rewrite: Hybrid Summarization with Hierarchical Reinforcement LearningLiqiang Xiao, Lu Wang, Hao He, Yaohui JinAAAI 2020 · 29 citations
Related papers
- HIBRIDS: Attention with Hierarchical Biases for Structure-aware Long Document SummarizationShuyang Cao, Lu WangACL 2022
- Neural Extractive Summarization with Hierarchical Attentive Heterogeneous Graph NetworkRuipeng Jia, Yanan Cao, Hengzhu Tang, Fang Fang et al.EMNLP 2020 · 87 citations
- Pre-training for Abstractive Document Summarization by Reinstating Source TextYanyan Zou, Xingxing Zhang, Wei Lu, Furu Wei et al.EMNLP 2020 · 42 citations
- STRUDEL: Structured Dialogue Summarization for Dialogue ComprehensionBorui Wang, Chengcheng Feng, Arjun Nair, Madelyn Mao et al.EMNLP 2022 · 2 citations
- DYLE: Dynamic Latent Extraction for Abstractive Long-Input SummarizationZiming Mao, Chen Henry Wu, Ansong Ni, Yusen Zhang et al.ACL 2022 · 62 citations
