A Unifying Information-theoretic Perspective on Evaluating Generative Models
Alexis Fox, Samarth Swarup, Abhijin Adiga
Abstract
Considering the difficulty of interpreting generative model output, there is significant current research focused on determining meaningful evaluation metrics. Several recent approaches utilize "precision" and "recall," borrowed from the classification domain, to individually quantify the output fidelity (realism) and output diversity (representation of the real data variation), respectively. With the increase in metric proposals, there is a need for a unifying perspective, allowing for easier comparison and clearer explanation of their benefits and drawbacks. To this end, we unify a class of kth-nearest neighbors (kNN)-based metrics under an information-theoretic lens using approaches from kNN density estimation. Additionally, we propose a tri-dimensional metric composed of Precision Cross-Entropy (PCE), Recall Cross-Entropy (RCE), and Recall Entropy (RE), which separately measure fidelity and two distinct aspects of diversity, inter- and intra-class. Our domain-agnostic metric, derived from the information-theoretic concepts of entropy and cross-entropy, can be dissected for both sample- and mode-level analysis. Our detailed experimental results demonstrate the sensitivity of our metric components to their respective qualities and reveal undesirable behaviors of other metrics.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 2729ee7b-5e88-47ff-8648-eb8f05e9f55aBuilds on8
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 13,211 citations
- Reliable Fidelity and Diversity Metrics for Generative ModelsMuhammad Ferjad Naeem, Seong Joon Oh, Youngjung Uh, Yunjey Choi et al.ICML 2020 · 553 citations
- How Faithful is your Synthetic Data? Sample-level Metrics for Evaluating and Auditing Generative ModelsAhmed M. Alaa, Boris van Breugel, Evgeny S. Saveliev, Mihaela van der SchaarICML 2022 · 287 citations
- Exposing flaws of generative model evaluation metrics and their unfair treatment of diffusion modelsGeorge Stein, Jesse C. Cresswell, Rasa Hosseinzadeh, Yi Sui et al.NeurIPS 2023 · 260 citations
- Divergence Frontiers for Generative Models: Sample Complexity, Quantization Effects, and Frontier IntegralsLang Liu, Krishna Pillutla, Sean Welleck, Sewoong Oh et al.NeurIPS 2021 · 22 citations
Related papers
- Probabilistic Precision and Recall Towards Reliable Evaluation of Generative ModelsDogyun Park, Suhyun KimICCV 2023 · 12 citations
- Emergent Asymmetry of Precision and Recall for Measuring Fidelity and Diversity of Generative Models in High DimensionsMahyar Khayatkhoei, Wael Abd-AlmageedICML 2023 · 11 citations
- Feature Likelihood Score: Evaluating the Generalization of Generative Models Using SamplesMarco Jiralerspong, Avishek Joey Bose, Ian Gemp, Chongli Qin et al.NeurIPS 2023 · 39 citations
- An Information-Theoretic Evaluation of Generative Models in Learning Multi-modal DistributionsMohammad Jalali, Cheuk Ting Li, Farzan FarniaNeurIPS 2023 · 46 citations
- A Theoretical Framework for Statistical Evaluability of Generative ModelsShashaank Aiyer, Yishay Mansour, Shay Moran, Han ShaoICML 2026 · 1 citation
