CHIMLE: Conditional Hierarchical IMLE for Multimodal Conditional Image Synthesis
Shichong Peng, Seyed Alireza Moazenipourasil, Ke Li
摘要
A persistent challenge in conditional image synthesis has been to generate diverse output images from the same input image despite only one output image being observed per input image. GAN-based methods are prone to mode collapse, which leads to low diversity. To get around this, we leverage Implicit Maximum Likelihood Estimation (IMLE) which can overcome mode collapse fundamentally. IMLE uses the same generator as GANs but trains it with a different, non-adversarial objective which ensures each observed image has a generated sample nearby. Unfortunately, to generate high-fidelity images, prior IMLE-based methods require a large number of samples, which is expensive. In this paper, we propose a new method to get around this limitation, which we dub Conditional Hierarchical IMLE (CHIMLE), which can generate high-fidelity images without requiring many samples. We show CHIMLE significantly outperforms the prior best IMLE, GAN and diffusion-based methods in terms of image fidelity and mode coverage across four tasks, namely night-to-day, 16× single image super-resolution, image colourization and image decompression. Quantitatively, our method improves Fréchet Inception Distance (FID) by 36.9% on average compared to the prior best IMLE-based method, and by 27.5% on average compared to the best non-IMLE-based generalpurpose methods. More results and code are available on the project website at https://niopeng.github.io/CHIMLE/ .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper14
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- Denoising Diffusion Restoration ModelsBahjat Kawar, Michael Elad, Stefano Ermon, Jiaming SongNeurIPS 2022 · 被引用 1,439 次
- NVAE: A Deep Hierarchical Variational AutoencoderArash Vahdat, Jan KautzNeurIPS 2020 · 被引用 1,141 次
- Generative Adversarial Networks for Extreme Learned Image CompressionEirikur Agustsson, Michael Tschannen, Fabian Mentzer, Radu Timofte 等ICCV 2019 · 被引用 648 次
- QS-Attn: Query-Selected Attention for Contrastive Learning in I2I TranslationXueqi Hu, Xinyue Zhou, Qiusheng Huang, Zhengyi Shi 等CVPR 2022 · 被引用 103 次
相关 Paper
- Adaptive IMLE for Few-shot Pretraining-free Generative ModellingMehran Aghabozorgi, Shichong Peng, Ke LiICML 2023 · 被引用 8 次
- Diverse Image Synthesis From Semantic Layouts via Conditional IMLEKe Li, Tianhao Zhang, Jitendra MalikICCV 2019 · 被引用 102 次
- Hierarchical Modes Exploring in Generative Adversarial NetworksMengxiao Hu, Jinlong Li, Maolin Hu, Tao HuAAAI 2020 · 被引用 2 次
- Omni-GAN: On the Secrets of cGANs and BeyondPeng Zhou, Lingxi Xie, Bingbing Ni, Cong Geng 等ICCV 2021 · 被引用 22 次
- Diverse Image Generation via Self-Conditioned GANsSteven Liu, Tongzhou Wang, David Bau, Jun-Yan Zhu 等CVPR 2020
