Diverse Semantic Image Synthesis via Probability Distribution Modeling
Zhentao Tan, Menglei Chai, Dongdong Chen, Jing Liao, Qi Chu, Bin Liu, Gang Hua, Nenghai Yu
摘要
Semantic image synthesis, translating semantic layouts to photo-realistic images, is a one-to-many mapping problem. Though impressive progress has been recently made, diverse semantic synthesis that can efficiently produce semantic-level multimodal results, still remains a challenge. In this paper, we propose a novel diverse semantic image synthesis framework from the perspective of semantic class distributions, which naturally supports diverse generation at semantic or even instance level. We achieve this by modeling class-level conditional modulation parameters as continuous probability distributions instead of discrete values, and sampling per-instance modulation parameters through instance-adaptive stochastic sampling that is consistent across the network. Moreover, we propose prior noise remapping, through linear perturbation parameters encoded from paired references, to facilitate supervised training and exemplar-based instance style control at test time. Extensive experiments on multiple datasets show that our method can achieve superior diversity and comparable quality compared to state-of-the-art methods. Code will be available at https://github.com/tzt101/ INADE.git
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper21
- Reduce Information Loss in Transformers for Pluralistic Image InpaintingQiankun Liu, Zhentao Tan, Dongdong Chen, Qi Chu 等CVPR 2022 · 被引用 99 次
- HairCLIP: Design Your Hair by Text and Reference ImageTianyi Wei, Dongdong Chen, Wenbo Zhou, Jing Liao 等CVPR 2022 · 被引用 94 次
- FreeMask: Synthetic Images with Dense Annotations Make Stronger Segmentation ModelsLihe Yang, Xiaogang Xu, Bingyi Kang, Yinghuan Shi 等NeurIPS 2023 · 被引用 94 次
- Scalable Multi-Temporal Remote Sensing Change Data Generation via Simulating Stochastic Change ProcessZhuo Zheng, Shiqi Tian, Ailong Ma, Liangpei Zhang 等ICCV 2023 · 被引用 31 次
- SemFlow: Binding Semantic Segmentation and Image Synthesis via Rectified FlowChaoyang Wang, Xiangtai Li, Lu Qi, Henghui Ding 等NeurIPS 2024 · 被引用 25 次
它引用的顶会 Paper7
- Dual Attention GANs for Semantic Image SynthesisHao Tang, Song Bai, Nicu SebeACM MM 2020 · 被引用 81 次
- MichiGAN: multi-input-conditioned hair image generation for portrait editingZhentao Tan, Menglei Chai, Dongdong Chen, Jing Liao 等SIGGRAPH 2020 · 被引用 75 次
- Cross-Domain Correspondence Learning for Exemplar-Based Image TranslationPan Zhang, Bo Zhang, Dong Chen, Lu Yuan 等CVPR 2020
- Semantically Multi-Modal Image SynthesisZhen Zhu, Zhiliang Xu, Ansheng You, Xiang BaiCVPR 2020
- MaskGAN: Towards Diverse and Interactive Facial Image ManipulationCheng-Han Lee, Ziwei Liu, Lingyun Wu, Ping LuoCVPR 2020
相关 Paper
- Semantic Palette: Guiding Scene Generation With Class ProportionsGuillaume Le Moing, Tuan-Hung Vu, Himalaya Jain, Patrick Pérez 等CVPR 2021
- PLACE: Adaptive Layout-Semantic Fusion for Semantic Image SynthesisZhengyao Lv, Yuxiang Wei, Wangmeng Zuo, Kwan-Yee K. WongCVPR 2024 · 被引用 14 次
- Diverse Image Synthesis From Semantic Layouts via Conditional IMLEKe Li, Tianhao Zhang, Jitendra MalikICCV 2019 · 被引用 102 次
- Semantic Image Analogy with a Conditional Single-Image GANJiacheng Li, Zhiwei Xiong, Dong Liu, Xuejin Chen 等ACM MM 2020 · 被引用 4 次
- Stochastic Conditional Diffusion Models for Robust Semantic Image SynthesisJuyeon Ko, Inho Kong, Dogyun Park, Hyunwoo J. KimICML 2024 · 被引用 14 次
