Disentangled Image Generation Through Structured Noise Injection
Yazeed Alharbi, Peter Wonka
摘要
We explore different design choices for injecting noise into generative adversarial networks (GANs) with the goal of disentangling the latent space. Instead of traditional approaches, we propose feeding multiple noise codes through separate fully-connected layers respectively. The aim is restricting the influence of each noise code to specific parts of the generated image. We show that disentanglement in the first layer of the generator network leads to disentanglement in the generated image. Through a grid-based structure, we achieve several aspects of disentanglement without complicating the network architecture and without requiring labels. We achieve spatial disentanglement, scale-space disentanglement, and disentanglement of the foreground object from the background style allowing fine-grained control over the generated images. Examples include changing facial expressions in face images, changing beak length in bird images, and changing car dimensions in car images. This empirically leads to better disentanglement scores than state-of-the-art methods on the FFHQ dataset.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper13
- EditGAN: High-Precision Semantic Image EditingHuan Ling, Karsten Kreis, Daiqing Li, Seung Wook Kim 等NeurIPS 2021 · 被引用 248 次
- You Only Need Adversarial Supervision for Semantic Image SynthesisEdgar Schönfeld, Vadim Sushko, Dan Zhang, Juergen Gall 等ICLR 2021 · 被引用 219 次
- TryOnGAN: body-aware try-on via layered interpolationKathleen M. Lewis, Srivatsan Varadharajan, Ira Kemelmacher-ShlizermanSIGGRAPH 2021 · 被引用 86 次
- Diagonal Attention and Style-based GAN for Content-Style Disentanglement in Image Generation and TranslationGihyun Kwon, Jong Chul YeICCV 2021 · 被引用 59 次
- TransEditor: Transformer-Based Dual-Space GAN for Highly Controllable Facial EditingYanbo Xu, Yueqin Yin, Liming Jiang, Qianyi Wu 等CVPR 2022 · 被引用 53 次
它引用的顶会 Paper3
- Unsupervised Robust Disentangling of Latent Characteristics for Image SynthesisPatrick Esser, Johannes Haux, Björn OmmerICCV 2019 · 被引用 40 次
- Identity From Here, Pose From There: Self-Supervised Disentanglement and Generation of Objects Using Unlabeled VideosFanyi Xiao, Haotian Liu, Yong Jae LeeICCV 2019 · 被引用 16 次
- MaskGAN: Towards Diverse and Interactive Facial Image ManipulationCheng-Han Lee, Ziwei Liu, Lingyun Wu, Ping LuoCVPR 2020
相关 Paper
- SemanticStyleGAN: Learning Compositional Generative Priors for Controllable Image Synthesis and EditingYichun Shi, Xiao Yang, Yangyue Wan, Xiaohui ShenCVPR 2022 · 被引用 88 次
- SDGAN: Disentangling Semantic Manipulation for Facial Attribute EditingWenmin Huang, Weiqi Luo, Jiwu Huang, Xiaochun CaoAAAI 2024 · 被引用 20 次
- A Latent Transformer for Disentangled Face Editing in Images and VideosXu Yao, Alasdair Newson, Yann Gousseau, Pierre HellierICCV 2021 · 被引用 97 次
- ContraFeat: Contrasting Deep Features for Semantic DiscoveryXinqi Zhu, Chang Xu, Dacheng TaoAAAI 2023 · 被引用 2 次
- StyleSpace Analysis: Disentangled Controls for StyleGAN Image GenerationZongze Wu, Dani Lischinski, Eli ShechtmanCVPR 2021
