Incorporating Distributions of Discourse Structure for Long Document Abstractive Summarization
Dongqi Liu, Yifan Wang, Vera Demberg
Abstract
For text summarization, the role of discourse structure is pivotal in discerning the core content of a text. Regrettably, prior studies on incorporating Rhetorical Structure Theory (RST) into transformer-based summarization models only consider the nuclearity annotation, thereby overlooking the variety of discourse relation types. This paper introduces the 'RSTformer', a novel summarization model that comprehensively incorporates both the types and uncertainty of rhetorical relations. Our RST-attention mechanism, rooted in document-level rhetorical structure, is an extension of the recently devised Longformer framework. Through rigorous evaluation, the model proposed herein exhibits significant superiority over state-of-theart models, as evidenced by its notable performance on several automatic metrics and human evaluation. 1
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 0d96f755-84fe-47dd-a446-67af4797141fCited by top-tier papers4
- BooookScore: A systematic exploration of book-length summarization in the era of LLMsYapei Chang, Kyle Lo, Tanya Goyal, Mohit IyyerICLR 2024 · 173 citations
- What Is That Talk About? A Video-to-Text Summarization Dataset for Scientific PresentationsDongqi Liu, Chenxi Whitehouse, Xi Yu, Louis Mahon et al.ACL 2025
- RST-Guarder: Enhancing Long-Context Robustness for Safeguards via RST Parsing and Probabilistic InferenceXu Zhang, Xiaojun WanACL 2026
- Disco-RAG: Discourse-Aware Retrieval-Augmented GenerationDongqi Liu, Hang Ding, Qiming Feng, Xurong Xie et al.ACL 2026
Builds on12
- BERTScore: Evaluating Text Generation with BERTTianyi Zhang, Varsha Kishore, Felix Wu, Kilian Q. Weinberger et al.ICLR 2020 · 8,443 citations
- Big Bird: Transformers for Longer SequencesManzil Zaheer, Guru Guruganesh, Kumar Avinava Dubey, Joshua Ainslie et al.NeurIPS 2020 · 3,159 citations
- Discourse-Aware Neural Extractive Text SummarizationJiacheng Xu, Zhe Gan, Yu Cheng, Jingjing LiuACL 2020 · 264 citations
- SG-Net: Syntax-Guided Machine Reading ComprehensionZhuosheng Zhang, Yuwei Wu, Junru Zhou, Sufeng Duan et al.AAAI 2020 · 192 citations
- Discourse Level Factors for Sentence Deletion in Text SimplificationYang Zhong, Chao Jiang, Wei Xu, Junyi Jessy LiAAAI 2020 · 57 citations
Related papers
- Beyond Chunking: Discourse-Aware Hierarchical Retrieval for Long Document Question AnsweringHuiyao Chen, Yi Yang, Yinghui Li, Meishan Zhang et al.ACL 2026 · 6 citations
- Exploring Discourse Structure in Document-level Machine TranslationXinyu Hu, Xiaojun WanEMNLP 2023 · 3 citations
- Top-Down RST Parsing Utilizing Granularity Levels in DocumentsNaoki Kobayashi, Tsutomu Hirao, Hidetaka Kamigaito, Manabu Okumura et al.AAAI 2020 · 48 citations
- A Top-down Neural Architecture towards Text-level Parsing of Discourse Rhetorical StructureLongyin Zhang, Yuqing Xing, Fang Kong, Peifeng Li et al.ACL 2020 · 39 citations
- HIBRIDS: Attention with Hierarchical Biases for Structure-aware Long Document SummarizationShuyang Cao, Lu WangACL 2022
