ObjBlur: A Curriculum Learning Approach With Progressive Object-Level Blurring for Improved Layout-to-Image Generation
Stanislav Frolov, Brian B. Moser, Sebastian Palacio, Andreas Dengel
Abstract
We present ObjBlur, a novel curriculum learning approach to improve layout-to-image generation models, where the task is to produce realistic images from layouts composed of boxes and labels. Our method is based on progressive object-level blurring, which effectively stabilizes training and enhances the quality of generated images. This curriculum learning strategy systematically applies varying degrees of blurring to individual objects or the background during training, starting from strong blurring to progressively cleaner images. Our findings reveal that this approach yields significant performance improvements, stabilized training, smoother convergence, and reduced variance between multiple runs. Moreover, our technique demonstrates its versatility by being compatible with generative adversarial networks and diffusion models, underlining its applicability across various generative modeling paradigms. With ObjBlur, we reach new state-of-the-art results on the complex COCO and Visual Genome datasets.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Builds on15
- Training Generative Adversarial Networks with Limited DataTero Karras, Miika Aittala, Janne Hellsten, Samuli Laine et al.NeurIPS 2020 · 2,345 citations
- Tackling the Generative Learning Trilemma with Denoising Diffusion GANsZhisheng Xiao, Karsten Kreis, Arash VahdatICLR 2022 · 726 citations
- Differentiable Augmentation for Data-Efficient GAN TrainingShengyu Zhao, Zhijian Liu, Ji Lin, Jun-Yan Zhu et al.NeurIPS 2020 · 707 citations
- Consistency Regularization for Generative Adversarial NetworksHan Zhang, Zizhao Zhang, Augustus Odena, Honglak LeeICLR 2020 · 305 citations
- Improved Consistency Regularization for GANsZhengli Zhao, Sameer Singh, Honglak Lee, Zizhao Zhang et al.AAAI 2021 · 166 citations
Related papers
- Denoising Task Difficulty-based Curriculum for Training Diffusion ModelsJin-Young Kim, Hyojun Go, Soonwoo Kwon, Hyun-Gyoon KimICLR 2025
- Context-Aware Layout to Image Generation With Enhanced Object AppearanceSen He, Wentong Liao, Michael Ying Yang, Yongxin Yang et al.CVPR 2021
- Interactive Image Synthesis with Panoptic Layout GenerationBo Wang, Tao Wu, Minfeng Zhu, Peng DuCVPR 2022 · 19 citations
- Image Synthesis From Reconfigurable Layout and StyleWei Sun, Tianfu WuICCV 2019 · 160 citations
- LAW-Diffusion: Complex Scene Generation by Diffusion with LayoutsBinbin Yang, Yi Luo, Ziliang Chen, Guangrun Wang et al.ICCV 2023 · 21 citations
