When, Why, and Which Pretrained GANs Are Useful?
Timofey Grigoryev, Andrey Voynov, Artem Babenko
Abstract
The literature has proposed several methods to finetune pretrained GANs on new datasets, which typically results in higher performance compared to training from scratch, especially in the limited-data regime. However, despite the apparent empirical benefits of GAN pretraining, its inner mechanisms were not analyzed in-depth, and understanding of its role is not entirely clear. Moreover, the essential practical details, e.g., selecting a proper pretrained GAN checkpoint, currently do not have rigorous grounding and are typically determined by trial and error. This work aims to dissect the process of GAN finetuning. First, we show that initializing the GAN training process by a pretrained checkpoint primarily affects the model's coverage rather than the fidelity of individual samples. Second, we explicitly describe how pretrained generators and discriminators contribute to the finetuning process and explain the previous evidence on the importance of pretraining both of them. Finally, as an immediate practical benefit of our analysis, we describe a simple recipe to choose an appropriate GAN checkpoint that is the most suitable for finetuning to a particular target task. Importantly, for most of the target tasks, Imagenet-pretrained GAN, despite having poor visual quality, appears to be an excellent starting point for finetuning, resembling the typical pretraining scenario of discriminative computer vision models.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 2ea95b5d-b46e-4597-a221-7825573550daCited by top-tier papers2
- StyleGAN-XL: Scaling StyleGAN to Large Diverse DatasetsAxel Sauer, Katja Schwarz, Andreas GeigerSIGGRAPH 2022 · 326 citations
- Restyling Unsupervised Concept Based Interpretable Networks with Generative ModelsJayneel Parekh, Quentin Bouniot, Pavlo Mozharovskyi, Alasdair Newson et al.ICLR 2025
Builds on11
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 24,064 citations
- Training Generative Adversarial Networks with Limited DataTero Karras, Miika Aittala, Janne Hellsten, Samuli Laine et al.NeurIPS 2020 · 2,345 citations
- Unsupervised Discovery of Interpretable Directions in the GAN Latent SpaceAndrey Voynov, Artem BabenkoICML 2020 · 459 citations
- Image Generation From Small Datasets via Batch Statistics AdaptationAtsuhiro Noguchi, Tatsuya HaradaICCV 2019 · 211 citations
- Few-shot Image Generation with Elastic Weight ConsolidationYijun Li, Richard Zhang, Jingwan Lu, Eli ShechtmanNeurIPS 2020 · 193 citations
Related papers
- Ensembling Off-the-shelf Models for GAN TrainingNupur Kumari, Richard Zhang, Eli Shechtman, Jun-Yan ZhuCVPR 2022 · 75 citations
- On Leveraging Pretrained GANs for Generation with Limited DataMiaoyun Zhao, Yulai Cong, Lawrence CarinICML 2020 · 102 citations
- MineGAN: Effective Knowledge Transfer From GANs to Target Domains With Few ImagesYaxing Wang, Abel Gonzalez-Garcia, David Berga, Luis Herranz et al.CVPR 2020
- FEditNet: Few-Shot Editing of Latent Semantics in GAN SpacesMengfei Xia, Yezhi Shu, Yuji Wang, Yu-Kun Lai et al.AAAI 2023 · 4 citations
- Finding an Unsupervised Image Segmenter in each of your Deep Generative ModelsLuke Melas-Kyriazi, Christian Rupprecht, Iro Laina, Andrea VedaldiICLR 2022 · 61 citations
