Unsupervised Extractive Summarization with Learnable Length Control Strategies
Renlong Jie, Xiaojun Meng, Xin Jiang, Qun Liu
摘要
Unsupervised extractive summarization is an important technique in information extraction and retrieval. Compared with supervised method, it does not require high-quality human-labelled summaries for training and thus can be easily applied for documents with different types, domains or languages. Most of existing unsupervised methods including TextRank and PACSUM rely on graph-based ranking on sentence centrality. However, this scorer can not be directly applied in end-to-end training, and the positional-related prior assumption is often needed for achieving good summaries. In addition, less attention is paid to length-controllable extractor, where users can decide to summarize texts under particular length constraint. This paper introduces an unsupervised extractive summarization model based on a siamese network, for which we develop a trainable bidirectional prediction objective between the selected summary and the original document. Different from the centrality-based ranking methods, our extractive scorer can be trained in an end-to-end manner, with no other requirement of positional assumption. In addition, we introduce a differentiable length control module by approximating 0-1 knapsack solver for end-to-end length-controllable extracting. Experiments show that our unsupervised method largely outperforms the centrality-based baseline using a same sentence encoder. In terms of length control ability, via our trainable knapsack module, the performance consistently outperforms the strong baseline without utilizing end-to-end training. Human evaluation further evidences that our method performs the best among baselines in terms of relevance and consistency.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper9
- SimCSE: Simple Contrastive Learning of Sentence EmbeddingsTianyu Gao, Xingcheng Yao, Danqi ChenEMNLP 2021 · 被引用 2,496 次
- Extractive Summarization as Text MatchingMing Zhong, Pengfei Liu, Yiran Chen, Danqing Wang 等ACL 2020 · 被引用 410 次
- Heterogeneous Graph Neural Networks for Extractive Document SummarizationDanqing Wang, Pengfei Liu, Yining Zheng, Xipeng Qiu 等ACL 2020 · 被引用 275 次
- Length Control in Abstractive Summarization by Pretraining Information SelectionYizhu Liu, Qi Jia, Kenny Q. ZhuACL 2022 · 被引用 39 次
- Deep Neural Network Approximated Dynamic Programming for Combinatorial OptimizationShenghe Xu, Shivendra S. Panwar, Murali S. Kodialam, T. V. LakshmanAAAI 2020 · 被引用 29 次
相关 Paper
- Learning Non-Autoregressive Models from Search for Unsupervised Sentence SummarizationPuyuan Liu, Chenyang Huang, Lili MouACL 2022 · 被引用 20 次
- RepSum: Unsupervised Dialogue Summarization based on Replacement StrategyXiyan Fu, Yating Zhang, Tianyi Wang, Xiaozhong Liu 等ACL 2021
- The Summary Loop: Learning to Write Abstractive Summaries Without ExamplesPhilippe Laban, Andrew Hsi, John F. Canny, Marti A. HearstACL 2020 · 被引用 26 次
- Flexible Non-Autoregressive Extractive Summarization with Threshold: How to Extract a Non-Fixed Number of Summary SentencesRuipeng Jia, Yanan Cao, Haichao Shi, Fang Fang 等AAAI 2021 · 被引用 15 次
- Improving Unsupervised Question Answering via Summarization-Informed Question GenerationChenyang Lyu, Lifeng Shang, Yvette Graham, Jennifer Foster 等EMNLP 2021 · 被引用 33 次
