Improving GAN Equilibrium by Raising Spatial Awareness
Jianyuan Wang, Ceyuan Yang, Yinghao Xu, Yujun Shen, Hongdong Li, Bolei Zhou
Abstract
The success of Generative Adversarial Networks (GANs) is largely built upon the adversarial training between a generator (G) and a discriminator (D). They are expected to reach a certain equilibrium where D cannot distinguish the generated images from the real ones. However, such an equilibrium is rarely achieved in practical GAN training, instead, D almost always surpasses G. We attribute one of its sources to the information asymmetry between D and G. We observe that D learns its own visual attention when determining whether an image is real or fake, but G has no explicit clue on which regions to focus on for a particular synthesis. To alleviate the issue of D dominating the competition in GANs, we aim to raise the spatial awareness of G. Randomly sampled multi-level heatmaps are encoded into the intermediate layers of G as an inductive bias. Thus G can purposefully improve the synthesis of certain image regions. We further propose to align the spatial awareness of G with the attention map induced from D. Through this way we effectively lessen the information gap between D and G. Extensive results show that our method pushes the two-player game in GANs closer to the equilibrium, leading to a better synthesis performance. As a byproduct, the introduced spatial awareness facilitates interactive editing over the output synthesis. Demo video and code are available at https://genforce.github.io/eqgan-sa/ .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 35a46063-38ed-4b76-bf9f-2ac309fdf630Cited by top-tier papers14
- Drag Your GAN: Interactive Point-based Manipulation on the Generative Image ManifoldXingang Pan, Ayush Tewari, Thomas Leimkühler, Lingjie Liu et al.SIGGRAPH 2023 · 206 citations
- MoVQ: Modulating Quantized Vectors for High-Fidelity Image GenerationChuanxia Zheng, Tung-Long Vuong, Jianfei Cai, Dinh PhungNeurIPS 2022 · 156 citations
- DirectGPT: A Direct Manipulation Interface to Interact with Large Language ModelsDamien Masson, Sylvain Malacria, Géry Casiez, Daniel VogelCHI 2024 · 104 citations
- Improving GANs with A Dynamic DiscriminatorCeyuan Yang, Yujun Shen, Yinghao Xu, Deli Zhao et al.NeurIPS 2022 · 42 citations
- Move Anything with Layered Scene DiffusionJiawei Ren, Mengmeng Xu, Jui-Chieh Wu, Ziwei Liu et al.CVPR 2024 · 7 citations
Builds on5
- Training Generative Adversarial Networks with Limited DataTero Karras, Miika Aittala, Janne Hellsten, Samuli Laine et al.NeurIPS 2020 · 2,345 citations
- Instance-Conditioned GANArantxa Casanova, Marlène Careil, Jakob Verbeek, Michal Drozdzal et al.NeurIPS 2021 · 167 citations
- GIRAFFE: Representing Scenes As Compositional Generative Neural Feature FieldsMichael Niemeyer, Andreas GeigerCVPR 2021
- A U-Net Based Discriminator for Generative Adversarial NetworksEdgar Schönfeld, Bernt Schiele, Anna KhorevaCVPR 2020
- Analyzing and Improving the Image Quality of StyleGANTero Karras, Samuli Laine, Miika Aittala, Janne Hellsten et al.CVPR 2020
Related papers
- GLeaD: Improving GANs with A Generator-Leading TaskQingyan Bai, Ceyuan Yang, Yinghao Xu, Xihui Liu et al.CVPR 2023
- Unpaired Image Enhancement with Quality-Attention Generative Adversarial NetworkZhangkai Ni, Wenhan Yang, Shiqi Wang, Lin Ma et al.ACM MM 2020 · 23 citations
- Toward Spatially Unbiased Generative ModelsJooyoung Choi, Jungbeom Lee, Yonghyun Jeong, Sungroh YoonICCV 2021 · 17 citations
- Diagonal Attention and Style-based GAN for Content-Style Disentanglement in Image Generation and TranslationGihyun Kwon, Jong Chul YeICCV 2021 · 59 citations
- Dual Attention GANs for Semantic Image SynthesisHao Tang, Song Bai, Nicu SebeACM MM 2020 · 81 citations
