PandA: Unsupervised Learning of Parts and Appearances in the Feature Maps of GANs
James Oldfield, Christos Tzelepis, Yannis Panagakis, Mihalis Nicolaou, Ioannis Patras
Abstract
Recent advances in the understanding of Generative Adversarial Networks (GANs) have led to remarkable progress in visual editing and synthesis tasks, capitalizing on the rich semantics that are embedded in the latent spaces of pre-trained GANs. However, existing methods are often tailored to specific GAN architectures and are limited to either discovering global semantic directions that do not facilitate localized control, or require some form of supervision through manually provided regions or segmentation masks. In this light, we present an architecture-agnostic approach that jointly discovers factors representing spatial parts and their appearances in an entirely unsupervised fashion. These factors are obtained by applying a semi-nonnegative tensor factorization on the feature maps, which in turn enables context-aware local image editing with pixel-level control. In addition, we show that the discovered appearance factors correspond to saliency maps that localize concepts of interest, without using any labels. Experiments on a wide range of GAN architectures and datasets show that, in comparison to the state of the art, our method is far more efficient in terms of training time and, most importantly, provides much more accurate localized control. Our code is available at: https://github.com/james-oldfield/PandA.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers9
- HyperReenact: One-Shot Reenactment via Jointly Learning to Refine and Retarget FacesStella Bounareli, Christos Tzelepis, Vasileios Argyriou, Ioannis Patras et al.ICCV 2023 · 63 citations
- EmerDiff: Emerging Pixel-level Semantic Knowledge in Diffusion ModelsKoichi Namekata, Amirmojtaba Sabour, Sanja Fidler, Seung Wook KimICLR 2024 · 41 citations
- Latent Traversals in Generative Models as Potential FlowsYue Song, T. Anderson Keller, Nicu Sebe, Max WellingICML 2023 · 15 citations
- Flow Factorized Representation LearningYue Song, Andy Keller, Nicu Sebe, Max WellingNeurIPS 2023 · 12 citations
- Householder Projector for Unsupervised Latent Semantics DiscoveryYue Song, Jichao Zhang, Nicu Sebe, Wei WangICCV 2023 · 9 citations
Builds on23
- Training Generative Adversarial Networks with Limited DataTero Karras, Miika Aittala, Janne Hellsten, Samuli Laine et al.NeurIPS 2020 · 2,345 citations
- Alias-Free Generative Adversarial NetworksTero Karras, Miika Aittala, Samuli Laine, Erik Härkönen et al.NeurIPS 2021 · 2,126 citations
- GANSpace: Discovering Interpretable GAN ControlsErik Härkönen, Aaron Hertzmann, Jaakko Lehtinen, Sylvain ParisNeurIPS 2020 · 1,049 citations
- Unsupervised Discovery of Interpretable Directions in the GAN Latent SpaceAndrey Voynov, Artem BabenkoICML 2020 · 459 citations
- GANalyze: Toward Visual Definitions of Cognitive Image PropertiesLore Goetschalckx, Alex Andonian, Aude Oliva, Phillip IsolaICCV 2019 · 345 citations
Related papers
- Region-Based Semantic Factorization in GANsJiapeng Zhu, Yujun Shen, Yinghao Xu, Deli Zhao et al.ICML 2022 · 42 citations
- LatentCLR: A Contrastive Learning Approach for Unsupervised Discovery of Interpretable DirectionsOguz Kaan Yüksel, Enis Simsar, Ezgi Gülperi Er, Pinar YanardagICCV 2021 · 71 citations
- Editing in Style: Uncovering the Local Semantics of GANsEdo Collins, Raja Bala, Bob Price, Sabine SüsstrunkCVPR 2020
- Low-Rank Subspaces in GANsJiapeng Zhu, Ruili Feng, Yujun Shen, Deli Zhao et al.NeurIPS 2021 · 80 citations
- Linear Semantics in Generative Adversarial NetworksJianjin Xu, Changxi ZhengCVPR 2021
