Rethinking conditional GAN training: An approach using geometrically structured latent manifolds
Sameera Ramasinghe, Moshiur R. Farazi, Salman H. Khan, Nick Barnes, Stephen Gould
Abstract
Conditional GANs (cGAN), in their rudimentary form, suffer from critical drawbacks such as the lack of diversity in generated outputs and distortion between the latent and output manifolds. Although efforts have been made to improve results, they can suffer from unpleasant side-effects such as the topology mismatch between latent and output spaces. In contrast, we tackle this problem from a geometrical perspective and propose a novel training mechanism that increases both the diversity and the visual quality of a vanilla cGAN, by systematically encouraging a bi-lipschitz mapping between the latent and the output manifolds. We validate the efficacy of our solution on a baseline cGAN (i.e., Pix2Pix) which lacks diversity, and show that by only modifying its training mechanism (i.e., with our proposed Pix2Pix-Geo), one can achieve more diverse and realistic outputs on a broad set of image-to-image translation tasks. Codes are available at https://github.com/samgregoost/Rethinking-CGANs .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 8ee667dd-ec69-4565-978d-37dbe119ccc3Cited by top-tier papers2
- Collapse by Conditioning: Training Class-conditional GANs with Limited DataMohamad Shahbazi, Martin Danelljan, Danda Pani Paudel, Luc Van GoolICLR 2022 · 39 citations
- Precise and Generalized Robustness Certification for Neural NetworksYuanyuan Yuan, Shuai Wang, Zhendong SuUSENIX Security 2023
Builds on4
- Conditional Generative Modeling via Learning the Latent SpaceSameera Ramasinghe, Kanchana Nisal Ranasinghe, Salman H. Khan, Nick Barnes et al.ICLR 2021 · 10 citations
- Sym-Parameterized Dynamic Inference for Mixed-Domain Image TranslationSimyung Chang, Seonguk Park, John Yang, Nojun KwakICCV 2019 · 8 citations
- Unpaired Image Super-Resolution Using Pseudo-SupervisionShunta MaedaCVPR 2020
- MaskGAN: Towards Diverse and Interactive Facial Image ManipulationCheng-Han Lee, Ziwei Liu, Lingyun Wu, Ping LuoCVPR 2020
Related papers
- Unsupervised Image-to-Image Translation with Generative PriorShuai Yang, Liming Jiang, Ziwei Liu, Chen Change LoyCVPR 2022 · 51 citations
- UniGAN: Reducing Mode Collapse in GANs using a Uniform GeneratorZiqi Pan, Li Niu, Liqing ZhangNeurIPS 2022 · 17 citations
- Smoothing the Disentangled Latent Style Space for Unsupervised Image-to-Image TranslationYahui Liu, Enver Sangineto, Yajing Chen, Linchao Bao et al.CVPR 2021
- DivCo: Diverse Conditional Image Synthesis via Contrastive Generative Adversarial NetworkRui Liu, Yixiao Ge, Ching Lam Choi, Xiaogang Wang et al.CVPR 2021
- Posterior Promoted GAN With Distribution Discriminator for Unsupervised Image SynthesisXianchao Zhang, Ziyang Cheng, Xiaotong Zhang, Han LiuCVPR 2021
