Semantic Palette: Guiding Scene Generation With Class Proportions
Guillaume Le Moing, Tuan-Hung Vu, Himalaya Jain, Patrick Pérez, Matthieu Cord
Abstract
Despite the recent progress of generative adversarial networks (GANs) at synthesizing photo-realistic images, producing complex urban scenes remains a challenging problem. Previous works break down scene generation into two consecutive phases: unconditional semantic layout synthesis and image synthesis conditioned on layouts. In this work, we propose to condition layout generation as well for higher semantic control: given a vector of class proportions, we generate layouts with matching composition. To this end, we introduce a conditional framework with novel architecture designs and learning objectives, which effectively accommodates class proportions to guide the scene generation process. The proposed architecture also allows partial layout editing with interesting applications. Thanks to the semantic control, we can produce layouts close to the real distribution, helping enhance the whole scene generation process. On different metrics and urban scene benchmarks, our models outperform existing baselines. Moreover, we demonstrate the merit of our approach for data augmentation: semantic segmenters trained on real layoutimage pairs along with additional ones generated by our approach outperform models only trained on real pairs.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 1be0ab83-e79a-4c4b-aa15-f6b24a0a80e5Cited by top-tier papers2
- Learning to Generate Semantic Layouts for Higher Text-Image Correspondence in Text-to-Image SynthesisMinho Park, Jooyeol Yun, Seunghwan Choi, Jaegul ChooICCV 2023 · 12 citations
- Adapting Diffusion Models for Improved Prompt Compliance and Controllable Image SynthesisDeepak Sridhar, Abhishek Peri, Rohith Rachala, Nuno VasconcelosNeurIPS 2024 · 5 citations
Builds on3
- Seeing What a GAN Cannot GenerateDavid Bau, Jun-Yan Zhu, Jonas Wulff, William S. Peebles et al.ICCV 2019 · 342 citations
- SEAN: Image Synthesis With Semantic Region-Adaptive NormalizationPeihao Zhu, Rameen Abdal, Yipeng Qin, Peter WonkaCVPR 2020
- Analyzing and Improving the Image Quality of StyleGANTero Karras, Samuli Laine, Miika Aittala, Janne Hellsten et al.CVPR 2020
Related papers
- CC3D: Layout-Conditioned Generation of Compositional 3D ScenesSherwin Bahmani, Jeong Joon Park, Despoina Paschalidou, Xingguang Yan et al.ICCV 2023 · 66 citations
- Image Synthesis via Semantic CompositionYi Wang, Lu Qi, Ying-Cong Chen, Xiangyu Zhang et al.ICCV 2021 · 72 citations
- Diverse Semantic Image Synthesis via Probability Distribution ModelingZhentao Tan, Menglei Chai, Dongdong Chen, Jing Liao et al.CVPR 2021
- End-to-End Optimization of Scene LayoutAndrew Luo, Zhoutong Zhang, Jiajun Wu, Joshua B. TenenbaumCVPR 2020
- Edge Guided GANs with Contrastive Learning for Semantic Image SynthesisHao Tang, Xiaojuan Qi, Guolei Sun, Dan Xu et al.ICLR 2023 · 2 citations
