Play the Shannon Game with Language Models: A Human-Free Approach to Summary Evaluation
Nicholas Egan, Oleg V. Vasilyev, John Bohannon
摘要
The goal of a summary is to concisely state the most important information in a document. With this principle in mind, we introduce new reference-free summary evaluation metrics that use a pretrained language model to estimate the information content shared between a document and its summary. These metrics are a modern take on the Shannon Game, a method for summary quality scoring proposed decades ago, where we replace human annotators with language models. We also view these metrics as an extension of BLANC, a recently proposed approach to summary quality measurement based on the performance of a language model with and without the help of a summary. Using transformer based language models, we empirically verify that our metrics achieve state-of-the-art correlation with human judgement of the summary quality dimensions of both coherence and relevance, as well as competitive correlation with human judgement of consistency and fluency.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- Compression, Transduction, and Creation: A Unified Framework for Evaluating Natural Language GenerationMingkai Deng, Bowen Tan, Zhengzhong Liu, Eric P. Xing 等EMNLP 2021 · 被引用 49 次
- Zero-shot Faithfulness Evaluation for Text Summarization with Foundation Language ModelQi Jia, Siyu Ren, Yizhu Liu, Kenny Q. ZhuEMNLP 2023 · 被引用 4 次
- COSMIC: Mutual Information for Task-Agnostic Summarization EvaluationMaxime Darrin, Philippe Formont, Jackie Chi Kit Cheung, Pablo PiantanidaACL 2024
它引用的顶会 Paper2
相关 Paper
- A Training-free and Reference-free Summarization Evaluation Metric via Centrality-weighted Relevance and Self-referenced RedundancyWang Chen, Piji Li, Irwin KingACL 2021
- QuestEval: Summarization Asks for Fact-based EvaluationThomas Scialom, Paul-Alexis Dray, Sylvain Lamprier, Benjamin Piwowarski 等EMNLP 2021
- Spurious Correlations in Reference-Free Evaluation of Text GenerationEsin Durmus, Faisal Ladhak, Tatsunori HashimotoACL 2022
- MTAS: A Reference-Free Approach for Evaluating Abstractive Summarization SystemsXiaoyan Zhu, Mingyue Jiang, Xiao-Yi Zhang, Liming Nie 等FSE 2024 · 被引用 2 次
- Unsupervised Reference-Free Summary Quality Evaluation via Contrastive LearningHanlu Wu, Tengfei Ma, Lingfei Wu, Tariro Manyumwa 等EMNLP 2020 · 被引用 47 次
