Unsupervised Reference-Free Summary Quality Evaluation via Contrastive Learning
Hanlu Wu, Tengfei Ma, Lingfei Wu, Tariro Manyumwa, Shouling Ji
Abstract
Evaluation of a document summarization system has been a critical factor to impact the success of the summarization task. Previous approaches, such as ROUGE, mainly consider the informativeness of the assessed summary and require human-generated references for each test summary. In this work, we propose to evaluate the summary qualities without reference summaries by unsupervised contrastive learning. Specifically, we design a new metric which covers both linguistic qualities and semantic informativeness based on BERT. To learn the metric, for each summary, we construct different types of negative samples with respect to different aspects of the summary qualities, and train our model with a ranking loss. Experiments on Newsroom and CNN/Daily Mail demonstrate that our new evaluation method outperforms other metrics even without reference summaries. Furthermore, we show that our method is general and transferable across datasets.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 237f6563-afab-4905-9856-3a9a1ee0eb76Cited by top-tier papers6
- Sequence Level Contrastive Learning for Text SummarizationShusheng Xu, Xingxing Zhang, Yi Wu, Furu WeiAAAI 2022 · 113 citations
- What are the Desired Characteristics of Calibration Sets? Identifying Correlates on Long Form Scientific SummarizationGriffin Adams, Bichlien Nguyen, Jake Smith, Yingce Xia et al.ACL 2023 · 8 citations
- Feeding What You Need by Understanding What You LearnedXiaoqiang Wang, Bang Liu, Fangli Xu, Bo Long et al.ACL 2022 · 6 citations
- QRelScore: Better Evaluating Generated Questions with Deeper Understanding of Context-aware RelevanceXiaoqiang Wang, Bang Liu, Siliang Tang, Lingfei WuEMNLP 2022 · 6 citations
- Improving Translation Quality Estimation with Bias MitigationHui Huang, Shuangzhi Wu, Kehai Chen, Hui Di et al.ACL 2023 · 2 citations
Builds on4
- BERTScore: Evaluating Text Generation with BERTTianyi Zhang, Varsha Kishore, Felix Wu, Kilian Q. Weinberger et al.ICLR 2020 · 8,443 citations
- Reinforcement Learning Based Graph-to-Sequence Model for Natural Question GenerationYu Chen, Lingfei Wu, Mohammed J. ZakiICLR 2020 · 167 citations
- Knowledge Graph-Augmented Abstractive Summarization with Semantic-Driven Cloze RewardLuyang Huang, Lingfei Wu, Lu WangACL 2020 · 152 citations
- Learning to Compare for Better Training and Evaluation of Open Domain Natural Language Generation ModelsWangchunshu Zhou, Ke XuAAAI 2020 · 49 citations
Related papers
- Evaluating Code Summarization Techniques: A New Metric and an Empirical CharacterizationAntonio Mastropaolo, Matteo Ciniselli, Massimiliano Di Penta, Gabriele BavotaICSE 2024 · 28 citations
- Discrete Optimization for Unsupervised Sentence Summarization with Word-Level ExtractionRaphael Schumann, Lili Mou, Yao Lu, Olga Vechtomova et al.ACL 2020 · 4 citations
- SemCSE: Semantic Contrastive Sentence Embeddings Using LLM-Generated Summaries For Scientific AbstractsMarc Felix Brinner, Sina ZarrießEMNLP 2025
- MTAS: A Reference-Free Approach for Evaluating Abstractive Summarization SystemsXiaoyan Zhu, Mingyue Jiang, Xiao-Yi Zhang, Liming Nie et al.FSE 2024 · 2 citations
- QuestEval: Summarization Asks for Fact-based EvaluationThomas Scialom, Paul-Alexis Dray, Sylvain Lamprier, Benjamin Piwowarski et al.EMNLP 2021
