Rethinking conditional GAN training: An approach using geometrically structured latent manifolds
Sameera Ramasinghe, Moshiur R. Farazi, Salman H. Khan, Nick Barnes, Stephen Gould
摘要
Conditional GANs (cGAN), in their rudimentary form, suffer from critical drawbacks such as the lack of diversity in generated outputs and distortion between the latent and output manifolds. Although efforts have been made to improve results, they can suffer from unpleasant side-effects such as the topology mismatch between latent and output spaces. In contrast, we tackle this problem from a geometrical perspective and propose a novel training mechanism that increases both the diversity and the visual quality of a vanilla cGAN, by systematically encouraging a bi-lipschitz mapping between the latent and the output manifolds. We validate the efficacy of our solution on a baseline cGAN (i.e., Pix2Pix) which lacks diversity, and show that by only modifying its training mechanism (i.e., with our proposed Pix2Pix-Geo), one can achieve more diverse and realistic outputs on a broad set of image-to-image translation tasks. Codes are available at https://github.com/samgregoost/Rethinking-CGANs .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Collapse by Conditioning: Training Class-conditional GANs with Limited DataMohamad Shahbazi, Martin Danelljan, Danda Pani Paudel, Luc Van GoolICLR 2022 · 被引用 39 次
- Precise and Generalized Robustness Certification for Neural NetworksYuanyuan Yuan, Shuai Wang, Zhendong SuUSENIX Security 2023
它引用的顶会 Paper4
- Conditional Generative Modeling via Learning the Latent SpaceSameera Ramasinghe, Kanchana Nisal Ranasinghe, Salman H. Khan, Nick Barnes 等ICLR 2021 · 被引用 10 次
- Sym-Parameterized Dynamic Inference for Mixed-Domain Image TranslationSimyung Chang, Seonguk Park, John Yang, Nojun KwakICCV 2019 · 被引用 8 次
- Unpaired Image Super-Resolution Using Pseudo-SupervisionShunta MaedaCVPR 2020
- MaskGAN: Towards Diverse and Interactive Facial Image ManipulationCheng-Han Lee, Ziwei Liu, Lingyun Wu, Ping LuoCVPR 2020
相关 Paper
- Unsupervised Image-to-Image Translation with Generative PriorShuai Yang, Liming Jiang, Ziwei Liu, Chen Change LoyCVPR 2022 · 被引用 51 次
- UniGAN: Reducing Mode Collapse in GANs using a Uniform GeneratorZiqi Pan, Li Niu, Liqing ZhangNeurIPS 2022 · 被引用 17 次
- Smoothing the Disentangled Latent Style Space for Unsupervised Image-to-Image TranslationYahui Liu, Enver Sangineto, Yajing Chen, Linchao Bao 等CVPR 2021
- DivCo: Diverse Conditional Image Synthesis via Contrastive Generative Adversarial NetworkRui Liu, Yixiao Ge, Ching Lam Choi, Xiaogang Wang 等CVPR 2021
- Posterior Promoted GAN With Distribution Discriminator for Unsupervised Image SynthesisXianchao Zhang, Ziyang Cheng, Xiaotong Zhang, Han LiuCVPR 2021
