A U-Net Based Discriminator for Generative Adversarial Networks
Edgar Schönfeld, Bernt Schiele, Anna Khoreva
Abstract
Among the major remaining challenges for generative adversarial networks (GANs ) is the capacity to synthesize globally and locally coherent images with object shapes and textures indistinguishable from real images. To target this issue we propose an alternative U-Net based discriminator architecture, borrowing the insights from the segmentation literature. The proposed U-Net based architecture allows to provide detailed per-pixel feedback to the generator while maintaining the global coherence of synthesized images, by providing the global image feedback as well. Empowered by the per-pixel response of the discriminator, we further propose a per-pixel consistency regularization technique based on the CutMix data augmentation, encouraging the U-Net discriminator to focus more on semantic and structural changes between real and fake images. This improves the U-Net discriminator training, further enhancing the quality of generated samples. The novel discriminator improves over the state of the art in terms of the standard distribution and image quality metrics, enabling the generator to synthesize images with varying structure, appearance and levels of detail, maintaining global and local realism. Compared to the BigGAN baseline, we achieve an average improvement of 2.7 FID points across FFHQ, CelebA, and the proposed COCO-Animals dataset.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 08e329dd-cec2-4079-866c-0a78b87b09adCited by top-tier papers48
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Training Generative Adversarial Networks with Limited DataTero Karras, Miika Aittala, Janne Hellsten, Samuli Laine et al.NeurIPS 2020 · 2,345 citations
- Vector-quantized Image Modeling with Improved VQGANJiahui Yu, Xin Li, Jing Yu Koh, Han Zhang et al.ICLR 2022 · 753 citations
- TransGAN: Two Pure Transformers Can Make One Strong GAN, and That Can Scale UpYifan Jiang, Shiyu Chang, Zhangyang WangNeurIPS 2021 · 515 citations
- Projected GANs Converge FasterAxel Sauer, Kashyap Chitta, Jens Müller, Andreas GeigerNeurIPS 2021 · 325 citations
Builds on4
- CutMix: Regularization Strategy to Train Strong Classifiers With Localizable FeaturesSangdoo Yun, Dongyoon Han, Sanghyuk Chun, Seong Joon Oh et al.ICCV 2019 · 5,843 citations
- Consistency Regularization for Generative Adversarial NetworksHan Zhang, Zizhao Zhang, Augustus Odena, Honglak LeeICLR 2020 · 305 citations
- Improved Consistency Regularization for GANsZhengli Zhao, Sameer Singh, Honglak Lee, Zizhao Zhang et al.AAAI 2021 · 166 citations
- COCO-GAN: Generation by Parts via Conditional CoordinatingChieh Hubert Lin, Chia-Che Chang, Yu-Sheng Chen, Da-Cheng Juan et al.ICCV 2019 · 147 citations
Related papers
- You Only Need Adversarial Supervision for Semantic Image SynthesisEdgar Schönfeld, Vadim Sushko, Dan Zhang, Juergen Gall et al.ICLR 2021 · 219 citations
- Self-Supervised Dense Consistency Regularization for Image-to-Image TranslationMinsu Ko, Eunju Cha, Sungjoo Suh, Huijin Lee et al.CVPR 2022 · 25 citations
- Feature Quantization Improves GAN TrainingYang Zhao, Chunyuan Li, Ping Yu, Jianfeng Gao et al.ICML 2020 · 49 citations
- On Positive-Unlabeled Classification in GANTianyu Guo, Chang Xu, Jiajun Huang, Yunhe Wang et al.CVPR 2020
- Detail Me More: Improving GAN's photo-realism of complex scenesRaghudeep Gadde, Qianli Feng, Aleix M. MartinezICCV 2021 · 21 citations
