Effectively Unbiased FID and Inception Score and Where to Find Them
Min Jin Chong, David A. Forsyth
摘要
This paper shows that two commonly used evaluation metrics for generative models, the Fréchet Inception Distance (FID) and the Inception Score (IS), are biased -the expected value of the score computed for a finite sample set is not the true value of the score. Worse, the paper shows that the bias term depends on the particular model being evaluated, so model A may get a better score than model B simply because model A's bias term is smaller. This effect cannot be fixed by evaluating at a fixed number of samples. This means all comparisons using FID or IS as currently computed are unreliable.
We then show how to extrapolate the score to obtain an effectively bias-free estimate of scores computed with an infinite number of samples, which we term FID ∞ and IS ∞ . In turn, this effectively bias-free estimate requires good estimates of scores with a finite number of samples. We show that using Quasi-Monte Carlo integration notably improves estimates of FID and IS for finite sample sets. Our extrapolated scores are simple, drop-in replacements for the finite sample scores. Additionally, we show that using low discrepancy sequence in GAN training offers small improvements in the resulting generator. The code for calculating FID ∞ and IS ∞ is at https://github.com/ mchong6/FID_IS_infinity.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper40
- A Continuous Time Framework for Discrete Denoising ModelsAndrew Campbell, Joe Benton, Valentin De Bortoli, Thomas Rainforth 等NeurIPS 2022 · 被引用 496 次
- Exposing flaws of generative model evaluation metrics and their unfair treatment of diffusion modelsGeorge Stein, Jesse C. Cresswell, Rasa Hosseinzadeh, Yi Sui 等NeurIPS 2023 · 被引用 260 次
- On Aliased Resizing and Surprising Subtleties in GAN EvaluationGaurav Parmar, Richard Zhang, Jun-Yan ZhuCVPR 2022 · 被引用 250 次
- Visual Fourier Prompt TuningRunjia Zeng, Cheng Han, Qifan Wang, Chunshu Wu 等NeurIPS 2024 · 被引用 58 次
- ChatScratch: An AI-Augmented System Toward Autonomous Visual Programming Learning for Children Aged 6-12Liuqing Chen, Shuhong Xiao, Yunnong Chen, Yaxuan Song 等CHI 2024 · 被引用 56 次
它引用的顶会 Paper1
相关 Paper
- Rethinking FID: Towards a Better Evaluation Metric for Image GenerationSadeep Jayasumana, Srikumar Ramalingam, Andreas Veit, Daniel Glasner 等CVPR 2024
- The Role of ImageNet Classes in Fréchet Inception DistanceTuomas Kynkäänniemi, Tero Karras, Miika Aittala, Timo Aila 等ICLR 2023 · 被引用 44 次
- TopP&R: Robust Support Estimation Approach for Evaluating Fidelity and Diversity in Generative ModelsPum Jun Kim, Yoojin Jang, Jisu Kim, Jaejun YooNeurIPS 2023 · 被引用 16 次
- PSA-GAN: Progressive Self Attention GANs for Synthetic Time SeriesPaul Jeha, Michael Bohlke-Schneider, Pedro Mercado, Shubham Kapoor 等ICLR 2022 · 被引用 92 次
- Influence Estimation for Generative Adversarial NetworksNaoyuki Terashita, Hiroki Ohashi, Yuichi Nonaka, Takashi KanemaruICLR 2021 · 被引用 12 次
