SARA: Salience-Aware Reinforced Adaptive Decoding for Large Language Models in Abstractive Summarization
Nayu Liu, Junnan Zhu, Yiming Ma, Zhicong Lu, Wenlei Xu, Yong Yang, Jiang Zhong, Kaiwen Wei
摘要
LLMs have improved the fluency and informativeness of abstractive summarization but remain prone to hallucinations, where generated content deviates from the source document. Recent PMI decoding strategies mitigate over-reliance on prior knowledge by comparing output probabilities with and without source documents, effectively enhancing contextual utilization and improving faithfulness. However, existing strategies often neglect the explicit use of salient contextual information and rely on static hyperparameters to fix the balance between contextual and prior knowledge, limiting their flexibility. In this work, we propose S alience-A ware R einforced A daptive decoding (SARA), which incorporates salient information and allows the model to adaptively determine reliance on the source document’s context, salient context, and the model’s prior knowledge based on pointwise mutual information. Moreover, a tokenwise adaptive decoding mechanism via reinforcement learning is proposed in SARA to dynamically adjust the contributions of context and prior knowledge at each decoding timestep. Experiments on CNN/DM, WikiHow, and NYT50 datasets show that SARA consistently improves the quality and faithfulness of summaries across various LLM backbones without modifying their weights.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- DeAR: Fine-Grained VLM Adaptation by Decomposing Attention Head RolesYiming Ma, Hongkun Yang, Lionel Z. Wang, Bin Chen 等CVPR 2026 · 被引用 2 次
- HyCoRA: Hyper-Contrastive Role-Adaptive Learning for Role-PlayingShihao Yang, Zhicong Lu, Yong Yang, Bo Lv 等AAAI 2026 · 被引用 2 次
- From Past To Path: Masked History Learning for Next-Item Prediction in Generative RecommendationKaiwen Wei, Kejun He, Xiaomian Kang, Jie Zhang 等ACL 2026
- Rectify Evaluation Preference: Improving LLMs' Critique on Math Reasoning via Perplexity-aware Reinforcement LearningChangyuan Tian, Zhicong Lu, Shuang Qian, Nayu Liu 等AAAI 2026
它引用的顶会 Paper17
- BERTScore: Evaluating Text Generation with BERTTianyi Zhang, Varsha Kishore, Felix Wu, Kilian Q. Weinberger 等ICLR 2020 · 被引用 8,443 次
- PEGASUS: Pre-training with Extracted Gap-sentences for Abstractive SummarizationJingqing Zhang, Yao Zhao, Mohammad Saleh, Peter J. LiuICML 2020 · 被引用 2,453 次
- CRITIC: Large Language Models Can Self-Correct with Tool-Interactive CritiquingZhibin Gou, Zhihong Shao, Yeyun Gong, Yelong Shen 等ICLR 2024 · 被引用 699 次
- SelfCheckGPT: Zero-Resource Black-Box Hallucination Detection for Generative Large Language ModelsPotsawee Manakul, Adian Liusie, Mark J. F. GalesEMNLP 2023 · 被引用 331 次
- Factuality Enhanced Language Models for Open-Ended Text GenerationNayeon Lee, Wei Ping, Peng Xu, Mostofa Patwary 等NeurIPS 2022 · 被引用 318 次
相关 Paper
- Towards Improving Faithfulness in Abstractive SummarizationXiuying Chen, Mingzhe Li, Xin Gao, Xiangliang ZhangNeurIPS 2022 · 被引用 39 次
- Breaking the Trade-Off Between Faithfulness and Expressiveness for Large Language ModelsChenxu Yang, Qingyi Si, Lanrui Wang, Zheng LinAAAI 2026
- Factually Consistent Summarization via Reinforcement Learning with Textual Entailment FeedbackPaul Roit, Johan Ferret, Lior Shani, Roee Aharoni 等ACL 2023 · 被引用 21 次
- Keyword-aware Abstractive Summarization by Extracting Set-level Intermediate SummariesYizhu Liu, Qi Jia, Kenny Q. ZhuWWW 2021 · 被引用 14 次
- LLMInertia: Adaptive Counter-Inertial Reasoning to Improve Evidence Faithfulness in Large Language ModelsXinxin You, Xien Liu, Chenwei Yan, Siqi Song 等ICML 2026
