You Only Need Adversarial Supervision for Semantic Image Synthesis
Edgar Schönfeld, Vadim Sushko, Dan Zhang, Juergen Gall, Bernt Schiele, Anna Khoreva
Abstract
Despite their recent successes, GAN models for semantic image synthesis still suffer from poor image quality when trained with only adversarial supervision. Historically, additionally employing the VGG-based perceptual loss has helped to overcome this issue, significantly improving the synthesis quality, but at the same time limiting the progress of GAN models for semantic image synthesis. In this work, we propose a novel, simplified GAN model, which needs only adversarial supervision to achieve high quality results. We re-design the discriminator as a semantic segmentation network, directly using the given semantic label maps as the ground truth for training. By providing stronger supervision to the discriminator as well as to the generator through spatially- and semantically-aware discriminator feedback, we are able to synthesize images of higher fidelity with better alignment to their input label maps, making the use of the perceptual loss superfluous. Moreover, we enable high-quality multi-modal image synthesis through global and local sampling of a 3D noise tensor injected into the generator, which allows complete or partial image change. We show that images synthesized by our model are more diverse and follow the color and texture distributions of real images more closely. We achieve an average improvement of FID and mIoU points over the state of the art across different datasets using only adversarial supervision.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 2229b0c3-370e-4d1d-a27a-8a6d6a60161bCited by top-tier papers49
- T2I-Adapter: Learning Adapters to Dig Out More Controllable Ability for Text-to-Image Diffusion ModelsChong Mou, Xintao Wang, Liangbin Xie, Yanze Wu et al.AAAI 2024 · 1,641 citations
- StyleGAN-NADA: CLIP-guided domain adaptation of image generatorsRinon Gal, Or Patashnik, Haggai Maron, Amit H. Bermano et al.SIGGRAPH 2022 · 501 citations
- Dense Text-to-Image Generation with Attention ModulationYunji Kim, Jiyoung Lee, Jin-Hwa Kim, Jung-Woo Ha et al.ICCV 2023 · 204 citations
- GANcraft: Unsupervised 3D Neural Rendering of Minecraft WorldsZekun Hao, Arun Mallya, Serge J. Belongie, Ming-Yu LiuICCV 2021 · 131 citations
- FreeMask: Synthetic Images with Dense Annotations Make Stronger Segmentation ModelsLihe Yang, Xiaogang Xu, Bingyi Kang, Yinghuan Shi et al.NeurIPS 2023 · 94 citations
Builds on6
- CutMix: Regularization Strategy to Train Strong Classifiers With Localizable FeaturesSangdoo Yun, Dongyoon Han, Sanghyuk Chun, Seong Joon Oh et al.ICCV 2019 · 5,843 citations
- Diverse Image Synthesis From Semantic Layouts via Conditional IMLEKe Li, Tianhao Zhang, Jitendra MalikICCV 2019 · 102 citations
- Dual Attention GANs for Semantic Image SynthesisHao Tang, Song Bai, Nicu SebeACM MM 2020 · 81 citations
- A U-Net Based Discriminator for Generative Adversarial NetworksEdgar Schönfeld, Bernt Schiele, Anna KhorevaCVPR 2020
- Local Class-Specific and Global Image-Level Generative Adversarial Networks for Semantic-Guided Scene GenerationHao Tang, Dan Xu, Yan Yan, Philip H. S. Torr et al.CVPR 2020
Related papers
- Unlocking Pre-Trained Image Backbones for Semantic Image SynthesisTariq Berrada, Jakob Verbeek, Camille Couprie, Karteek AlahariCVPR 2024
- SeD: Semantic-Aware Discriminator for Image Super-ResolutionBingchen Li, Xin Li, Hanxin Zhu, Yeying Jin et al.CVPR 2024
- Network-Free, Unsupervised Semantic Segmentation with Synthetic ImagesQianli Feng, Raghudeep Gadde, Wentong Liao, Eduard Ramon et al.CVPR 2023
- Collaging Class-specific GANs for Semantic Image SynthesisYuheng Li, Yijun Li, Jingwan Lu, Eli Shechtman et al.ICCV 2021 · 36 citations
- Semantic Image Analogy with a Conditional Single-Image GANJiacheng Li, Zhiwei Xiong, Dong Liu, Xuejin Chen et al.ACM MM 2020 · 4 citations
