Language Models of Code Are Few-Shot Planners and Reasoners for Multi-Document Summarization with Attribution
Abhilash Nandy, Sambaran Bandyopadhyay
摘要
Document summarization has greatly benefited from advances in large language models (LLMs). In real-world situations, summaries often need to be generated from multiple documents with diverse sources and authors, lacking a clear information flow. Naively concatenating these documents and generating a summary can lead to poorly structured narratives and redundancy. Additionally, attributing each part of the generated summary to a specific source is crucial for reliability. In this study, we address multi-document summarization with attribution using our proposed solution MiDAS-PRo, consisting of three stages: (i) Planning the hierarchical organization of source documents, (ii) Reasoning by generating relevant entities/topics, and (iii) Summary Generation. We treat the first two sub-problems as a code completion task for LLMs. By incorporating well-selected in-context learning examples through a graph attention network, LLMs effectively generate plans and reason topics for a document collection. Experiments on summarizing scientific articles from public datasets show that our approach outperforms state-ofthe-art baselines in both automated and human evaluations.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- CoNewsReader: Supporting Comprehensive Understanding and Raising Critical Thoughts on Social Media News Through CommentsKangyu Yuan, Guanzheng Chen, Sizhe Liang, Hehai Lin 等CSCW 2026
- RSF-GLLM: Bridging the Semantic Gap in Multi-Hop Knowledge Graph QA via Recurrent Soft-Flow and Decoupled LLM GenerationSambaran Bandyopadhyay, Ananth MuppidiICML 2026
- WikiREVIEW: A Multi-Perspective Review Framework for Automatic Wiki-Style Article GenerationGuo-Biao Zhang, Zhijing Wu, Tian Lan, Ding-Yuan Liu 等AAAI 2026
它引用的顶会 Paper9
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah 等NeurIPS 2020 · 被引用 64,255 次
- BERTScore: Evaluating Text Generation with BERTTianyi Zhang, Varsha Kishore, Felix Wu, Kilian Q. Weinberger 等ICLR 2020 · 被引用 8,443 次
- PEGASUS: Pre-training with Extracted Gap-sentences for Abstractive SummarizationJingqing Zhang, Yao Zhao, Mohammad Saleh, Peter J. LiuICML 2020 · 被引用 2,453 次
- BART: Denoising Sequence-to-Sequence Pre-training for Natural Language Generation, Translation, and ComprehensionMike Lewis, Yinhan Liu, Naman Goyal, Marjan Ghazvininejad 等ACL 2020 · 被引用 1,224 次
- PAL: Program-aided Language ModelsLuyu Gao, Aman Madaan, Shuyan Zhou, Uri Alon 等ICML 2023 · 被引用 700 次
相关 Paper
- Principled Content Selection to Generate Diverse and Personalized Multi-Document SummariesVishakh Padmakumar, Zichao Wang, David Arbour, Jennifer HealeyACL 2025 · 被引用 1 次
- Attribute First, then Generate: Locally-attributable Grounded Text GenerationAviv Slobodkin, Eran Hirsch, Arie Cattan, Tal Schuster 等ACL 2024
- Attribute or Abstain: Large Language Models as Long Document AssistantsJan Buchmann, Xiao Liu, Iryna GurevychEMNLP 2024
- Mixture of Knowledge Minigraph Agents for Literature Review GenerationZhi Zhang, Yan Liu, Sheng-hua Zhong, Gong Chen 等AAAI 2025 · 被引用 1 次
- Multi-Level Explanations for Generative Language ModelsLucas Monteiro Paes, Dennis Wei, Hyo Jin Do, Hendrik Strobelt 等ACL 2025 · 被引用 16 次
