Interactive Image Synthesis with Panoptic Layout Generation
Bo Wang, Tao Wu, Minfeng Zhu, Peng Du
Abstract
Interactive image synthesis from user-guided input is a challenging task when users wish to control the scene structure of a generated image with ease. Although remarkable progress has been made on layout-based image synthesis approaches, existing methods require high-precision inputs such as accurately placed bounding boxes, which might be constantly violated in an interactive setting. When placement of bounding boxes is subject to perturbation, layout-based models suffer from “missing regions” in the constructed semantic layouts and hence undesirable artifacts in the generated images. In this work, we propose Panoptic Layout Generative Adversarial Network (PLGAN) to address this challenge. The PLGAN employs panoptic theory which distinguishes object categories between “stuff” with amorphous boundaries and “things” with well-defined shapes, such that stuff and instance layouts are constructed through separate branches and later fused into panoptic layouts. In particular, the stuff layouts can take amorphous shapes and fill up the missing regions left out by the instance layouts. We experimentally compare our PLGAN with state-of-the-art layout-based models on the COCO-Stuff, Visual Genome, and Landscape datasets. The advantages of PLGAN are not only visually demonstrated but quantitatively verified in terms of inception score, Fréchet inception distance, classification accuracy score, and coverage. The code is available at https://github.com/wb-finalking/PLGAN.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 3c9caa09-af55-4035-bbd5-a00b9cde35d8Cited by top-tier papers8
- LayoutGPT: Compositional Visual Planning and Generation with Large Language ModelsWeixi Feng, Wanrong Zhu, Tsu-Jui Fu, Varun Jampani et al.NeurIPS 2023 · 462 citations
- HiCo: Hierarchical Controllable Diffusion Model for Layout-to-image GenerationBo Cheng, Yuhang Ma, Liebucha Wu, Shanyuan Liu et al.NeurIPS 2024 · 53 citations
- PlantoGraphy: Incorporating Iterative Design Process into Generative Artificial Intelligence for Landscape RenderingRong Huang, Haichuan Lin, Chuanzhang Chen, Kang Zhang et al.CHI 2024 · 45 citations
- DC-ControlNet: Decoupling Inter- and Intra-Element Conditions in Image Generation with Diffusion ModelsHongji Yang, Wencheng Han, Yucheng Zhou, Jianbing ShenICCV 2025 · 4 citations
- Zero-Painter: Training-Free Layout Control for Text-to-Image SynthesisMarianna Ohanyan, Hayk Manukyan, Zhangyang Wang, Shant Navasardyan et al.CVPR 2024 · 4 citations
Builds on9
- Specifying Object Attributes and Relations in Interactive Scene GenerationOron Ashual, Lior WolfICCV 2019 · 190 citations
- Image Synthesis From Reconfigurable Layout and StyleWei Sun, Tianfu WuICCV 2019 · 160 citations
- Context-Aware Layout to Image Generation With Enhanced Object AppearanceSen He, Wentong Liao, Michael Ying Yang, Yongxin Yang et al.CVPR 2021
- LayoutTransformer: Scene Layout Generation With Conceptual and Spatial DiversityCheng-Fu Yang, Wan-Cyuan Fan, Fu-En Yang, Yu-Chiang Frank WangCVPR 2021
- Panoptic-DeepLab: A Simple, Strong, and Fast Baseline for Bottom-Up Panoptic SegmentationBowen Cheng, Maxwell D. Collins, Yukun Zhu, Ting Liu et al.CVPR 2020
Related papers
- Contextual Outpainting with Object-Level Contrastive LearningJiacheng Li, Chang Chen, Zhiwei XiongCVPR 2022 · 10 citations
- Object-Centric Image Generation from LayoutsTristan Sylvain, Pengchuan Zhang, Yoshua Bengio, R. Devon Hjelm et al.AAAI 2021 · 107 citations
- Image Synthesis from Layout with Locality-Aware Mask AdaptionZejian Li, Jingyu Wu, Immanuel Koh, Yongchuan Tang et al.ICCV 2021 · 88 citations
- LayoutVAE: Stochastic Scene Layout Generation From a Label SetAkash Abdu Jyothi, Thibaut Durand, Jiawei He, Leonid Sigal et al.ICCV 2019 · 194 citations
- LAW-Diffusion: Complex Scene Generation by Diffusion with LayoutsBinbin Yang, Yi Luo, Ziliang Chen, Guangrun Wang et al.ICCV 2023 · 21 citations
