Learning with Rejection for Abstractive Text Summarization
Meng Cao, Yue Dong, Jingyi He, Jackie Chi Kit Cheung
Abstract
State-of-the-art abstractive summarization systems frequently hallucinate content that is not supported by the source document, mainly due to noise in the training dataset. Existing methods opt to drop the noisy samples or tokens from the training set entirely, reducing the effective training set size and creating an artificial propensity to copy words from the source. In this work, we propose a training objective for abstractive summarization based on rejection learning, in which the model learns whether or not to reject potentially noisy tokens. We further propose a regularized decoding objective that penalizes nonfactual candidate summaries during inference by using the rejection probability learned during training. We show that our method considerably improves the factuality of generated summaries in automatic and human evaluations when compared to five baseline models, and that it does so while increasing the abstractiveness of the generated summaries. 1
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext fc003f62-46de-4c9e-ad45-5964e2883e0bCited by top-tier papers7
- Teaching Language Models to Hallucinate Less with Synthetic TasksErik Jones, Hamid Palangi, Clarisse Simões, Varun Chandrasekaran et al.ICLR 2024 · 43 citations
- Detecting and Mitigating Hallucinations in Multilingual SummarisationYifu Qiu, Yftah Ziser, Anna Korhonen, Edoardo Maria Ponti et al.EMNLP 2023 · 12 citations
- Beyond Next Token Probabilities: Learnable, Fast Detection of Hallucinations and Data Contamination on LLM Output DistributionsGuy Bar-Shalom, Fabrizio Frasca, Derek Lim, Yoav Gelberg et al.AAAI 2026 · 7 citations
- Enhancing Reinforcement Learning with Dense Rewards from Language Model CriticMeng Cao, Lei Shu, Lei Yu, Yun Zhu et al.EMNLP 2024 · 7 citations
- Promoting Topic Coherence and Inter-Document Consorts in Multi-Document Summarization via Simplicial Complex and Sheaf GraphYash Kumar Atri, Arun Iyer, Tanmoy Chakraborty, Vikram GoyalEMNLP 2023 · 2 citations
Builds on11
- PEGASUS: Pre-training with Extracted Gap-sentences for Abstractive SummarizationJingqing Zhang, Yao Zhao, Mohammad Saleh, Peter J. LiuICML 2020 · 2,453 citations
- BART: Denoising Sequence-to-Sequence Pre-training for Natural Language Generation, Translation, and ComprehensionMike Lewis, Yinhan Liu, Naman Goyal, Marjan Ghazvininejad et al.ACL 2020 · 1,224 citations
- BRIO: Bringing Order to Abstractive SummarizationYixin Liu, Pengfei Liu, Dragomir R. Radev, Graham NeubigACL 2022 · 329 citations
- Multi-Fact Correction in Abstractive Text SummarizationYue Dong, Shuohang Wang, Zhe Gan, Yu Cheng et al.EMNLP 2020 · 99 citations
- FEQA: A Question Answering Evaluation Framework for Faithfulness Assessment in Abstractive SummarizationEsin Durmus, He He, Mona T. DiabACL 2020 · 90 citations
Related papers
- Hallucinated but Factual! Inspecting the Factuality of Hallucinations in Abstractive SummarizationMeng Cao, Yue Dong, Jackie Chi Kit CheungACL 2022
- On Faithfulness and Factuality in Abstractive SummarizationJoshua Maynez, Shashi Narayan, Bernd Bohnet, Ryan T. McDonaldACL 2020 · 54 citations
- Unsupervised Opinion Summarization with Noising and DenoisingReinald Kim Amplayo, Mirella LapataACL 2020 · 8 citations
- Improving Factual Consistency of Abstractive Summarization via Question AnsweringFeng Nan, Cícero Nogueira dos Santos, Henghui Zhu, Patrick Ng et al.ACL 2021
- Pre-training for Abstractive Document Summarization by Reinstating Source TextYanyan Zou, Xingxing Zhang, Wei Lu, Furu Wei et al.EMNLP 2020 · 42 citations
