Unsupervised Disentanglement of Linear-Encoded Facial Semantics
Yutong Zheng, Yu-Kai Huang, Ran Tao, Zhiqiang Shen, Marios Savvides
Abstract
We propose a method to disentangle linear-encoded facial semantics from StyleGAN without external supervision. The method derives from linear regression and sparse representation learning concepts to make the disentangled latent representations easily interpreted as well. We start by coupling StyleGAN with a stabilized 3D deformable facial reconstruction method to decompose single-view GAN generations into multiple semantics. Latent representations are then extracted to capture interpretable facial semantics. In this work, we make it possible to get rid of labels for disentangling meaningful facial semantics. Also, we demonstrate that the guided extrapolation along the disentangled representations can help with data augmentation, which sheds light on handling unbalanced data. Finally, we provide an analysis of our learned localized facial representations and illustrate that the semantic information is encoded, which surprisingly complies with human intuition. The overall unsupervised design brings more flexibility to representation learning in the wild.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 61b6abc4-8b0d-4cc5-a518-79fac2619dfcCited by top-tier papers4
- Interpretable Generative Adversarial NetworksChao Li, Kelu Yao, Jin Wang, Boyu Diao et al.AAAI 2022 · 19 citations
- Rethinking Controllable Variational AutoencodersHuajie Shao, Yifei Yang, Haohong Lin, Longzhong Lin et al.CVPR 2022 · 14 citations
- Text-Guided Unsupervised Latent Transformation for Multi-Attribute Image ManipulationXiwen Wei, Zhen Xu, Cheng Liu, Si Wu et al.CVPR 2023
- Progressive Disentangled Representation Learning for Fine-Grained Controllable Talking Head SynthesisDuomin Wang, Yu Deng, Zixin Yin, Heung-Yeung Shum et al.CVPR 2023
Builds on6
- NVAE: A Deep Hierarchical Variational AutoencoderArash Vahdat, Jan KautzNeurIPS 2020 · 1,141 citations
- On the "steerability" of generative adversarial networksAli Jahanian, Lucy Chai, Phillip IsolaICLR 2020 · 421 citations
- Unsupervised Learning of Probably Symmetric Deformable 3D Objects From Images in the WildShangzhe Wu, Christian Rupprecht, Andrea VedaldiCVPR 2020
- Disentangled and Controllable Face Image Generation via 3D Imitative-Contrastive LearningYu Deng, Jiaolong Yang, Dong Chen, Fang Wen et al.CVPR 2020
- Analyzing and Improving the Image Quality of StyleGANTero Karras, Samuli Laine, Miika Aittala, Janne Hellsten et al.CVPR 2020
Related papers
- Self-Supervised Emotion Representation Disentanglement for Speech-Preserving Facial Expression ManipulationZhihua Xu, Tianshui Chen, Zhijing Yang, Chunmei Qing et al.ACM MM 2024 · 5 citations
- Only a matter of style: age transformation using a style-based regression modelYuval Alaluf, Or Patashnik, Daniel Cohen-OrSIGGRAPH 2021 · 143 citations
- Self-Supervised Geometry-Aware Encoder for Style-Based 3D GAN InversionYushi Lan, Xuyi Meng, Shuai Yang, Chen Change Loy et al.CVPR 2023
- Interpreting the Latent Space of GANs for Semantic Face EditingYujun Shen, Jinjin Gu, Xiaoou Tang, Bolei ZhouCVPR 2020
- Retrieve in Style: Unsupervised Facial Feature Transfer and RetrievalMin Jin Chong, Wen-Sheng Chu, Abhishek Kumar, David A. ForsythICCV 2021 · 27 citations
