On the Relation between Quality-Diversity Evaluation and Distribution-Fitting Goal in Text Generation
Jianing Li, Yanyan Lan, Jiafeng Guo, Xueqi Cheng
摘要
The goal of text generation models is to fit the underlying real probability distribution of text. For performance evaluation, quality and diversity metrics are usually applied. However, it is still not clear to what extend can the quality-diversity evaluation reflect the distribution-fitting goal. In this paper, we try to reveal such relation in a theoretical approach. We prove that under certain conditions, a linear combination of quality and diversity constitutes a divergence metric between the generated distribution and the real distribution. We also show that the commonly used BLEU/Self-BLEU metric pair fails to match any divergence metric, thus propose CR/NRR as a substitute for quality/diversity metric pair.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Evade the Trap of Mediocrity: Promoting Diversity and Novelty in Text Generation via Concentrating AttentionWenhao Li, Xiaoyuan Yi, Jinyi Hu, Maosong Sun 等EMNLP 2022 · 被引用 1 次
- Distributional Open-Ended Evaluation of LLM Cultural Value Alignment Based on Value CodebookJaehyeok Lee, Xiaoyuan Yi, Jing Yao, Hyunjin Hwang 等ICML 2026 · 被引用 1 次
它引用的顶会 Paper1
相关 Paper
- On the Usefulness of Embeddings, Clusters and Strings for Text Generation EvaluationTiago Pimentel, Clara Meister, Ryan CotterellICLR 2023
- MAUVE: Measuring the Gap Between Neural Text and Human Text using Divergence FrontiersKrishna Pillutla, Swabha Swayamdipta, Rowan Zellers, John Thickstun 等NeurIPS 2021 · 被引用 606 次
- SESCORE2: Learning Text Generation Evaluation via Synthesizing Realistic MistakesWenda Xu, Xian Qian, Mingxuan Wang, Lei Li 等ACL 2023 · 被引用 3 次
- Open-Domain Text Evaluation via Contrastive Distribution MethodsSidi Lu, Hongyi Liu, Asli Celikyilmaz, Tianlu Wang 等ICML 2024 · 被引用 2 次
- Tailoring Language Generation Models under Total Variation DistanceHaozhe Ji, Pei Ke, Zhipeng Hu, Rongsheng Zhang 等ICLR 2023 · 被引用 2 次
