On the Relation between Quality-Diversity Evaluation and Distribution-Fitting Goal in Text Generation
Jianing Li, Yanyan Lan, Jiafeng Guo, Xueqi Cheng
Abstract
The goal of text generation models is to fit the underlying real probability distribution of text. For performance evaluation, quality and diversity metrics are usually applied. However, it is still not clear to what extend can the quality-diversity evaluation reflect the distribution-fitting goal. In this paper, we try to reveal such relation in a theoretical approach. We prove that under certain conditions, a linear combination of quality and diversity constitutes a divergence metric between the generated distribution and the real distribution. We also show that the commonly used BLEU/Self-BLEU metric pair fails to match any divergence metric, thus propose CR/NRR as a substitute for quality/diversity metric pair.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 39364c55-cadb-4228-890d-c2fb332e12b4Cited by top-tier papers2
- Evade the Trap of Mediocrity: Promoting Diversity and Novelty in Text Generation via Concentrating AttentionWenhao Li, Xiaoyuan Yi, Jinyi Hu, Maosong Sun et al.EMNLP 2022 · 1 citation
- Distributional Open-Ended Evaluation of LLM Cultural Value Alignment Based on Value CodebookJaehyeok Lee, Xiaoyuan Yi, Jing Yao, Hyunjin Hwang et al.ICML 2026 · 1 citation
Builds on1
Related papers
- On the Usefulness of Embeddings, Clusters and Strings for Text Generation EvaluationTiago Pimentel, Clara Meister, Ryan CotterellICLR 2023
- MAUVE: Measuring the Gap Between Neural Text and Human Text using Divergence FrontiersKrishna Pillutla, Swabha Swayamdipta, Rowan Zellers, John Thickstun et al.NeurIPS 2021 · 606 citations
- SESCORE2: Learning Text Generation Evaluation via Synthesizing Realistic MistakesWenda Xu, Xian Qian, Mingxuan Wang, Lei Li et al.ACL 2023 · 3 citations
- Open-Domain Text Evaluation via Contrastive Distribution MethodsSidi Lu, Hongyi Liu, Asli Celikyilmaz, Tianlu Wang et al.ICML 2024 · 2 citations
- Tailoring Language Generation Models under Total Variation DistanceHaozhe Ji, Pei Ke, Zhipeng Hu, Rongsheng Zhang et al.ICLR 2023 · 2 citations
