Discourse-Aware Neural Extractive Text Summarization
Jiacheng Xu, Zhe Gan, Yu Cheng, Jingjing Liu
摘要
Recently BERT has been adopted for document encoding in state-of-the-art text summarization models. However, sentence-based extractive models often result in redundant or uninformative phrases in the extracted summaries. Also, long-range dependencies throughout a document are not well captured by BERT, which is pre-trained on sentence pairs instead of documents. To address these issues, we present a discourse-aware neural summarization model -DISCOBERT 1 . DISCOBERT extracts sub-sentential discourse units (instead of sentences) as candidates for extractive selection on a finer granularity. To capture the long-range dependencies among discourse units, structural discourse graphs are constructed based on RST trees and coreference mentions, encoded with Graph Convolutional Networks. Experiments show that the proposed model outperforms state-of-the-art methods by a significant margin on popular summarization benchmarks compared to other BERT-base models.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper37
- Language Model Tokenizers Introduce Unfairness Between LanguagesAleksandar Petrov, Emanuele La Malfa, Philip H. S. Torr, Adel BibiNeurIPS 2023 · 被引用 301 次
- Neural Extractive Summarization with Hierarchical Attentive Heterogeneous Graph NetworkRuipeng Jia, Yanan Cao, Hengzhu Tang, Fang Fang 等EMNLP 2020 · 被引用 87 次
- SKIER: A Symbolic Knowledge Integrated Model for Conversational Emotion RecognitionWei Li, Luyao Zhu, Rui Mao, Erik CambriaAAAI 2023 · 被引用 85 次
- Semantic Self-Segmentation for Abstractive Summarization of Long Documents in Low-Resource RegimesGianluca Moro, Luca RagazziAAAI 2022 · 被引用 67 次
- Educational Question Generation of Children Storybooks via Question Type Distribution Learning and Event-centric SummarizationZhenjie Zhao, Yufang Hou, Dakuo Wang, Mo Yu 等ACL 2022 · 被引用 50 次
相关 Paper
- Leveraging Graph to Improve Abstractive Multi-Document SummarizationWei Li, Xinyan Xiao, Jiachen Liu, Hua Wu 等ACL 2020 · 被引用 118 次
- Span Graph Transformer for Document-Level Named Entity RecognitionHongli Mao, Xian-Ling Mao, Hanlin Tang, Yuming Shang 等AAAI 2024 · 被引用 3 次
- HEGEL: Hypergraph Transformer for Long Document SummarizationHaopeng Zhang, Xiao Liu, Jiawei ZhangEMNLP 2022 · 被引用 33 次
- SemSUM: Semantic Dependency Guided Neural Abstractive SummarizationHanqi Jin, Tianming Wang, Xiaojun WanAAAI 2020 · 被引用 60 次
- DialogBERT: Discourse-Aware Response Generation via Learning to Recover and Rank UtterancesXiaodong Gu, Kang Min Yoo, Jung-Woo HaAAAI 2021 · 被引用 83 次
