Unsupervised Discovery of Interpretable Directions in the GAN Latent Space
Andrey Voynov, Artem Babenko
摘要
The latent spaces of GAN models often have semantically meaningful directions. Moving in these directions corresponds to humaninterpretable image transformations, such as zooming or recoloring, enabling a more controllable generation process. However, the discovery of such directions is currently performed in a supervised manner, requiring human labels, pretrained models, or some form of self-supervision. These requirements severely restrict a range of directions existing approaches can discover. In this paper, we introduce an unsupervised method to identify interpretable directions in the latent space of a pretrained GAN model. By a simple model-agnostic procedure, we find directions corresponding to sensible semantic manipulations without any form of (self-)supervision. Furthermore, we reveal several non-trivial findings, which would be difficult to obtain by existing methods, e.g., a direction corresponding to background removal. As an immediate practical benefit of our work, we show how to exploit this finding to achieve competitive performance for weakly-supervised saliency detection. The implementation of our method is available online 1 .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper128
- StyleCLIP: Text-Driven Manipulation of StyleGAN ImageryOr Patashnik, Zongze Wu, Eli Shechtman, Daniel Cohen-Or 等ICCV 2021 · 被引用 1,437 次
- GANSpace: Discovering Interpretable GAN ControlsErik Härkönen, Aaron Hertzmann, Jaakko Lehtinen, Sylvain ParisNeurIPS 2020 · 被引用 1,049 次
- Label-Efficient Semantic Segmentation with Diffusion ModelsDmitry Baranchuk, Andrey Voynov, Ivan Rubachev, Valentin Khrulkov 等ICLR 2022 · 被引用 700 次
- Designing an encoder for StyleGAN image manipulationOmer Tov, Yuval Alaluf, Yotam Nitzan, Or Patashnik 等SIGGRAPH 2021 · 被引用 692 次
- ReStyle: A Residual-Based StyleGAN Encoder via Iterative RefinementYuval Alaluf, Or Patashnik, Daniel Cohen-OrICCV 2021 · 被引用 377 次
它引用的顶会 Paper2
相关 Paper
- LatentCLR: A Contrastive Learning Approach for Unsupervised Discovery of Interpretable DirectionsOguz Kaan Yüksel, Enis Simsar, Ezgi Gülperi Er, Pinar YanardagICCV 2021 · 被引用 71 次
- GAN "Steerability" without optimizationNurit Spingarn, Ron Banner, Tomer MichaeliICLR 2021 · 被引用 59 次
- Controlling generative models with continuous factors of variationsAntoine Plumerault, Hervé Le Borgne, Céline HudelotICLR 2020 · 被引用 132 次
- Toward a Visual Concept Vocabulary for GAN Latent SpaceSarah Schwettmann, Evan Hernandez, David Bau, Samuel Klein 等ICCV 2021 · 被引用 16 次
- Finding an Unsupervised Image Segmenter in each of your Deep Generative ModelsLuke Melas-Kyriazi, Christian Rupprecht, Iro Laina, Andrea VedaldiICLR 2022 · 被引用 61 次
