GAN "Steerability" without optimization
Nurit Spingarn, Ron Banner, Tomer Michaeli
摘要
Recent research has shown remarkable success in revealing "steering" directions in the latent spaces of pre-trained GANs. These directions correspond to semantically meaningful image transformations (e.g., shift, zoom, color manipulations), and have similar interpretable effects across all categories that the GAN can generate. Some methods focus on user-specified transformations, while others discover transformations in an unsupervised manner. However, all existing techniques rely on an optimization procedure to expose those directions, and offer no control over the degree of allowed interaction between different transformations. In this paper, we show that "steering" trajectories can be computed in closed form directly from the generator's weights without any form of training or optimization. This applies to user-prescribed geometric transformations, as well as to unsupervised discovery of more complex effects. Our approach allows determining both linear and nonlinear trajectories, and has many advantages over previous methods. In particular, we can control whether one transformation is allowed to come on the expense of another (e.g., zoom-in with or without allowing translation to keep the object centered). Moreover, we can determine the natural end-point of the trajectory, which corresponds to the largest extent to which a transformation can be applied without incurring degradation. Finally, we show how transferring attributes between images can be achieved without optimization, even across different categories.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper23
- Designing an encoder for StyleGAN image manipulationOmer Tov, Yuval Alaluf, Yotam Nitzan, Or Patashnik 等SIGGRAPH 2021 · 被引用 692 次
- StyleGAN-XL: Scaling StyleGAN to Large Diverse DatasetsAxel Sauer, Katja Schwarz, Andreas GeigerSIGGRAPH 2022 · 被引用 326 次
- GAN-Control: Explicitly Controllable GANsAlon Shoshan, Nadav Bhonker, Igor Kviatkovsky, Gérard G. MedioniICCV 2021 · 被引用 151 次
- Low-Rank Subspaces in GANsJiapeng Zhu, Ruili Feng, Yujun Shen, Deli Zhao 等NeurIPS 2021 · 被引用 80 次
- LatentCLR: A Contrastive Learning Approach for Unsupervised Discovery of Interpretable DirectionsOguz Kaan Yüksel, Enis Simsar, Ezgi Gülperi Er, Pinar YanardagICCV 2021 · 被引用 71 次
它引用的顶会 Paper5
- GANSpace: Discovering Interpretable GAN ControlsErik Härkönen, Aaron Hertzmann, Jaakko Lehtinen, Sylvain ParisNeurIPS 2020 · 被引用 1,049 次
- On the "steerability" of generative adversarial networksAli Jahanian, Lucy Chai, Phillip IsolaICLR 2020 · 被引用 421 次
- GANalyze: Toward Visual Definitions of Cognitive Image PropertiesLore Goetschalckx, Alex Andonian, Aude Oliva, Phillip IsolaICCV 2019 · 被引用 345 次
- Specifying Object Attributes and Relations in Interactive Scene GenerationOron Ashual, Lior WolfICCV 2019 · 被引用 190 次
- Analyzing and Improving the Image Quality of StyleGANTero Karras, Samuli Laine, Miika Aittala, Janne Hellsten 等CVPR 2020
相关 Paper
- Unsupervised Discovery of Interpretable Directions in the GAN Latent SpaceAndrey Voynov, Artem BabenkoICML 2020 · 被引用 459 次
- Controlling generative models with continuous factors of variationsAntoine Plumerault, Hervé Le Borgne, Céline HudelotICLR 2020 · 被引用 132 次
- WarpedGANSpace: Finding non-linear RBF paths in GAN latent spaceChristos Tzelepis, Georgios Tzimiropoulos, Ioannis PatrasICCV 2021 · 被引用 66 次
- Toward a Visual Concept Vocabulary for GAN Latent SpaceSarah Schwettmann, Evan Hernandez, David Bau, Samuel Klein 等ICCV 2021 · 被引用 16 次
- A Geometric Analysis of Deep Generative Image Models and Its ApplicationsBinxu Wang, Carlos R. PonceICLR 2021 · 被引用 42 次
