Lune

NeurIPS2021Top-tier venue

Diffusion Models Beat GANs on Image Synthesis

Prafulla Dhariwal, Alexander Quinn Nichol

2021Year
13,211Citations
3,075Top-tier citations

Abstract

We show that diffusion models can achieve image sample quality superior to the current state-of-the-art generative models. We achieve this on unconditional image synthesis by finding a better architecture through a series of ablations. For conditional image synthesis, we further improve sample quality with classifier guidance: a simple, compute-efficient method for trading off diversity for fidelity using gradients from a classifier. We achieve an FID of 2.97 on ImageNet 128×\times128, 4.59 on ImageNet 256×\times256, and 7.72 on ImageNet 512×\times512, and we match BigGAN-deep even with as few as 25 forward passes per sample, all while maintaining better coverage of the distribution. Finally, we find that classifier guidance combines well with upsampling diffusion models, further improving FID to 3.94 on ImageNet 256×\times256 and 3.85 on ImageNet 512×\times512. We release our code at https://github.com/openai/guided-diffusion

Ask about this paper

Your agent reads all of it.

Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.

Questions to start from

Your agent calls

Luneget_paper_fulltext

Ask in Lune

Free to start. No credit card required.

lune papers fulltext 831a8d93-e50f-4ba0-b952-ca340af982cc

Cited by top-tier papers3,075

Ask how each one uses it

Builds on17

Related papers

Dusk over the sea between two cliffs drawn in fine vertical lines