CHIMLE: Conditional Hierarchical IMLE for Multimodal Conditional Image Synthesis
Shichong Peng, Seyed Alireza Moazenipourasil, Ke Li
Abstract
A persistent challenge in conditional image synthesis has been to generate diverse output images from the same input image despite only one output image being observed per input image. GAN-based methods are prone to mode collapse, which leads to low diversity. To get around this, we leverage Implicit Maximum Likelihood Estimation (IMLE) which can overcome mode collapse fundamentally. IMLE uses the same generator as GANs but trains it with a different, non-adversarial objective which ensures each observed image has a generated sample nearby. Unfortunately, to generate high-fidelity images, prior IMLE-based methods require a large number of samples, which is expensive. In this paper, we propose a new method to get around this limitation, which we dub Conditional Hierarchical IMLE (CHIMLE), which can generate high-fidelity images without requiring many samples. We show CHIMLE significantly outperforms the prior best IMLE, GAN and diffusion-based methods in terms of image fidelity and mode coverage across four tasks, namely night-to-day, 16× single image super-resolution, image colourization and image decompression. Quantitatively, our method improves Fréchet Inception Distance (FID) by 36.9% on average compared to the prior best IMLE-based method, and by 27.5% on average compared to the best non-IMLE-based generalpurpose methods. More results and code are available on the project website at https://niopeng.github.io/CHIMLE/ .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 71ee428a-01a8-448d-b795-8b27d4b017f9Builds on14
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- Denoising Diffusion Restoration ModelsBahjat Kawar, Michael Elad, Stefano Ermon, Jiaming SongNeurIPS 2022 · 1,439 citations
- NVAE: A Deep Hierarchical Variational AutoencoderArash Vahdat, Jan KautzNeurIPS 2020 · 1,141 citations
- Generative Adversarial Networks for Extreme Learned Image CompressionEirikur Agustsson, Michael Tschannen, Fabian Mentzer, Radu Timofte et al.ICCV 2019 · 648 citations
- QS-Attn: Query-Selected Attention for Contrastive Learning in I2I TranslationXueqi Hu, Xinyue Zhou, Qiusheng Huang, Zhengyi Shi et al.CVPR 2022 · 103 citations
Related papers
- Adaptive IMLE for Few-shot Pretraining-free Generative ModellingMehran Aghabozorgi, Shichong Peng, Ke LiICML 2023 · 8 citations
- Diverse Image Synthesis From Semantic Layouts via Conditional IMLEKe Li, Tianhao Zhang, Jitendra MalikICCV 2019 · 102 citations
- Hierarchical Modes Exploring in Generative Adversarial NetworksMengxiao Hu, Jinlong Li, Maolin Hu, Tao HuAAAI 2020 · 2 citations
- Omni-GAN: On the Secrets of cGANs and BeyondPeng Zhou, Lingxi Xie, Bingbing Ni, Cong Geng et al.ICCV 2021 · 22 citations
- Diverse Image Generation via Self-Conditioned GANsSteven Liu, Tongzhou Wang, David Bau, Jun-Yan Zhu et al.CVPR 2020
