Discovering Density-Preserving Latent Space Walks in GANs for Semantic Image Transformations
Guanyue Li, Yi Liu, Xiwen Wei, Yang Zhang, Si Wu, Yong Xu, Hau-San Wong
Abstract
Generative adversarial network (GAN)-based models possess superior capability of high-fidelity image synthesis. There are a wide range of semantically meaningful directions in the latent representation space of well-trained GANs, and the corresponding latent space walks are meaningful for semantic controllability in the synthesized images. To explore the underlying organization of a latent space, we propose an unsupervised Density-Preserving Latent Semantics Exploration model (DP-LaSE). The important latent directions are determined by maximizing the variations in intermediate features, while the correlation between the directions is minimized. Considering that latent codes are sampled from a prior distribution, we adopt a density-preserving regularization approach to ensure latent space walks are maintained in iso-density regions, since moving to a higher/lower density region tends to cause unexpected transformations. To further refine semantics-specific transformations, we perform subspace learning over intermediate feature channels, such that the transformations are limited to the most relevant subspaces. Extensive experiments on a variety of benchmark datasets demonstrate that DP-LaSE is able to discover interpretable latent space walks, and specific properties of synthesized images can thus be precisely controlled.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 801bde9c-42fb-4eb4-8c9a-2825737f6d2cCited by top-tier papers2
- Leveraging GAN Priors for Few-Shot Part SegmentationMengya Han, Heliang Zheng, Chaoyue Wang, Yong Luo et al.ACM MM 2022 · 5 citations
- Cycle Encoding of a StyleGAN Encoder for Improved Reconstruction and EditabilityXudong Mao, Liujuan Cao, Aurele Tohokantche Gnanha, Zhenguo Yang et al.ACM MM 2022 · 5 citations
Builds on10
- GANSpace: Discovering Interpretable GAN ControlsErik Härkönen, Aaron Hertzmann, Jaakko Lehtinen, Sylvain ParisNeurIPS 2020 · 1,049 citations
- Unsupervised Discovery of Interpretable Directions in the GAN Latent SpaceAndrey Voynov, Artem BabenkoICML 2020 · 459 citations
- On the "steerability" of generative adversarial networksAli Jahanian, Lucy Chai, Phillip IsolaICLR 2020 · 421 citations
- GANalyze: Toward Visual Definitions of Cognitive Image PropertiesLore Goetschalckx, Alex Andonian, Aude Oliva, Phillip IsolaICCV 2019 · 345 citations
- Controlling generative models with continuous factors of variationsAntoine Plumerault, Hervé Le Borgne, Céline HudelotICLR 2020 · 132 citations
Related papers
- High Fidelity GAN Inversion via Prior Multi-Subspace Feature CompositionGuanyue Li, Qianfen Jiao, Sheng Qian, Si Wu et al.AAAI 2021
- LatentCLR: A Contrastive Learning Approach for Unsupervised Discovery of Interpretable DirectionsOguz Kaan Yüksel, Enis Simsar, Ezgi Gülperi Er, Pinar YanardagICCV 2021 · 71 citations
- Do Not Escape From the Manifold: Discovering the Local Coordinates on the Latent Space of GANsJaewoong Choi, Junho Lee, Changyeon Yoon, Jung Ho Park et al.ICLR 2022 · 34 citations
- Finding the Global Semantic Representation in GAN through Fréchet MeanJaewoong Choi, Geonho Hwang, Hyunsoo Cho, Myungjoo KangICLR 2023
- Interpreting the Latent Space of GANs for Semantic Face EditingYujun Shen, Jinjin Gu, Xiaoou Tang, Bolei ZhouCVPR 2020
