Detail Me More: Improving GAN's photo-realism of complex scenes
Raghudeep Gadde, Qianli Feng, Aleix M. Martinez
Abstract
Generative models can synthesize photo-realistic images of a single object. For example, for human faces, algorithms learn to model the local shape and shading of the face components, i.e., changes in the brows, eyes, nose, mouth, jaw line, etc. This is possible because all faces have two brows, two eyes, a nose and a mouth, approximately in the same location. The modeling of complex scenes is however much more challenging because the scene components and their location vary from image to image. For example, living rooms contain a varying number of products belonging to many possible categories and locations, e.g., a lamp may or may not be present in an endless number of possible locations. In the present work, we propose to add a "broker" module in Generative Adversarial Networks (GAN) to solve this problem. The broker is tasked to mediate the use of multiple discriminators in the appropriate image locales. For example, if a lamp is detected or wanted in a specific area of the scene, the broker assigns a fine-grained lamp discriminator to that image patch. This allows the generator to learn the shape and shading models of the lamp. The resulting multi-fine-grained optimization problem is able to synthesize complex scenes with almost the same level of photo-realism as single object images. We demonstrate the generability of the proposed approach on several GAN algorithms (BigGAN, ProGAN, StyleGAN, StyleGAN2), image resolutions (2562 to 10242), and datasets. Our approach yields significant improvements over state-of-the-art GAN algorithms.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 6e4068c4-007f-49a4-a5e9-446068bb7a91Cited by top-tier papers9
- Ensembling Off-the-shelf Models for GAN TrainingNupur Kumari, Richard Zhang, Eli Shechtman, Jun-Yan ZhuCVPR 2022 · 75 citations
- UrbanGIRAFFE: Representing Urban Scenes as Compositional Generative Neural Feature FieldsYuanbo Yang, Yifei Yang, Hanlei Guo, Rong Xiong et al.ICCV 2023 · 26 citations
- Rewriting geometric rules of a GANSheng-Yu Wang, David Bau, Jun-Yan ZhuSIGGRAPH 2022 · 20 citations
- DPGEN: Differentially Private Generative Energy-Guided Network for Natural Image SynthesisJia-Wei Chen, Chia-Mu Yu, Ching-Chia Kao, Tzai-Wei Pang et al.CVPR 2022 · 14 citations
- ICAR: Image-Based Complementary Auto ReasoningXijun Wang, Anqi Liang, Junbang Liang, Ming C. Lin et al.AAAI 2024 · 1 citation
Builds on10
- AutoGAN: Neural Architecture Search for Generative Adversarial NetworksXinyu Gong, Shiyu Chang, Yifan Jiang, Zhangyang WangICCV 2019 · 286 citations
- Object-Centric Image Generation from LayoutsTristan Sylvain, Pengchuan Zhang, Yoshua Bengio, R. Devon Hjelm et al.AAAI 2021 · 107 citations
- Instance Selection for GANsTerrance DeVries, Michal Drozdzal, Graham W. TaylorNeurIPS 2020 · 41 citations
- A U-Net Based Discriminator for Generative Adversarial NetworksEdgar Schönfeld, Bernt Schiele, Anna KhorevaCVPR 2020
- MSG-GAN: Multi-Scale Gradients for Generative Adversarial NetworksAnimesh Karnewar, Oliver WangCVPR 2020
Related papers
- BodyGAN: General-purpose Controllable Neural Human Body GenerationChaojie Yang, Hanhui Li, Shengjie Wu, Shengkai Zhang et al.CVPR 2022 · 8 citations
- BlockGAN: Learning 3D Object-aware Scene Representations from Unlabelled ImagesThu Nguyen-Phuoc, Christian Richardt, Long Mai, Yong-Liang Yang et al.NeurIPS 2020 · 256 citations
- R-GAN: Exploring Human-like Way for Reasonable Text-to-Image Synthesis via Generative Adversarial NetworksYanyuan Qiao, Qi Chen, Chaorui Deng, Ning Ding et al.ACM MM 2021 · 18 citations
- InsetGAN for Full-Body Image GenerationAnna Frühstück, Krishna Kumar Singh, Eli Shechtman, Niloy J. Mitra et al.CVPR 2022 · 52 citations
- Multi-Class Multi-Instance Count Conditioned Adversarial Image GenerationAmrutha Saseendran, Kathrin Skubch, Margret KeuperICCV 2021 · 2 citations
