Stochastic Conditional Diffusion Models for Robust Semantic Image Synthesis
Juyeon Ko, Inho Kong, Dogyun Park, Hyunwoo J. Kim
摘要
Semantic image synthesis (SIS) is a task to generate realistic images corresponding to semantic maps (labels). However, in real-world applications, SIS often encounters noisy user inputs. To address this, we propose Stochastic Conditional Diffusion Model (SCDM), which is a robust conditional diffusion model that features novel forward and generation processes tailored for SIS with noisy labels. It enhances robustness by stochastically perturbing the semantic label maps through Label Diffusion, which diffuses the labels with discrete diffusion. Through the diffusion of labels, the noisy and clean semantic maps become similar as the timestep increases, eventually becoming identical at t = T . This facilitates the generation of an image close to a clean image, enabling robust generation. Furthermore, we propose a class-wise noise schedule to differentially diffuse the labels depending on the class. We demonstrate that the proposed method generates high-quality samples through extensive experiments and analyses on benchmark datasets, including a novel experimental setup simulating human errors during real-world applications. Code is available at https://github.com/mlvlab/SCDM .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper7
- Constant Acceleration FlowDogyun Park, Sojin Lee, Sihyeon Kim, Taehoon Lee 等NeurIPS 2024 · 被引用 14 次
- SPRINT: Sparse-Dense Residual Fusion for Efficient Diffusion TransformersDogyun Park, Moayed Haji-Ali, Yanyu Li, Willi Menapace 等ICLR 2026 · 被引用 6 次
- Symmetrical Flow Matching: Unified Image Generation, Segmentation, and Classification with Score-Based Generative ModelsFrancisco Caetano, Christiaan G. A. Viviers, Peter H. N. de With, Fons van der SommenAAAI 2026 · 被引用 4 次
- Blockwise Flow Matching: Improving Flow Matching Models For Efficient High-Quality GenerationDogyun Park, Taehoon Lee, Minseok Joo, Hyunwoo J. KimNeurIPS 2025 · 被引用 4 次
- DA-Font: Few-Shot Font Generation via Dual-Attention Hybrid IntegrationWeiran Chen, Guiqian Zhu, Ying Li, Yi Ji 等ACM MM 2025 · 被引用 2 次
它引用的顶会 Paper24
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 被引用 13,211 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Photorealistic Text-to-Image Diffusion Models with Deep Language UnderstandingChitwan Saharia, William Chan, Saurabh Saxena, Lala Li 等NeurIPS 2022 · 被引用 8,965 次
- Adding Conditional Control to Text-to-Image Diffusion ModelsLvmin Zhang, Anyi Rao, Maneesh AgrawalaICCV 2023 · 被引用 6,759 次
相关 Paper
- Stochastic Segmentation with Conditional Categorical Diffusion ModelsLukas Zbinden, Lars Doorenbos, Theodoros Pissas, Adrian Thomas Huber 等ICCV 2023 · 被引用 57 次
- Label-Noise Robust Diffusion ModelsByeonghu Na, Yeongmin Kim, HeeSun Bae, Jung Hyun Lee 等ICLR 2024 · 被引用 20 次
- Unlocking Pre-Trained Image Backbones for Semantic Image SynthesisTariq Berrada, Jakob Verbeek, Camille Couprie, Karteek AlahariCVPR 2024
- Directional Label Diffusion Model for Learning from Noisy LabelsSenyu Hou, Gaoxia Jiang, Jia Zhang, Shangrong Yang 等CVPR 2025
- An End-to-End Robust Point Cloud Semantic Segmentation Network with Single-Step Conditional Diffusion ModelsWentao Qu, Jing Wang, Yongshun Gong, Xiaoshui Huang 等CVPR 2025
