Linear Semantics in Generative Adversarial Networks
Jianjin Xu, Changxi Zheng
摘要
Generative Adversarial Networks (GANs) are able to generate high-quality images, but it remains difficult to explicitly specify the semantics of synthesized images. In this work, we aim to better understand the semantic representation of GANs, and thereby enable semantic control in GAN's generation process. Interestingly, we find that a well-trained GAN encodes image semantics in its internal feature maps in a surprisingly simple way: a linear transformation of feature maps suffices to extract the generated image semantics. To verify this simplicity, we conduct extensive experiments on various GANs and datasets; and thanks to this simplicity, we are able to learn a semantic segmentation model for a trained GAN from a small number (e.g., 8) of labeled images. Last but not least, leveraging our finding, we propose two few-shot image editing approaches, namely Semantic-Conditional Sampling and Semantic Image Editing. Given a trained GAN and as few as eight semantic annotations, the user is able to generate diverse images subject to a userprovided semantic layout, and control the synthesized image semantics. We have made the code publicly available 1 .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper11
- Label-Efficient Semantic Segmentation with Diffusion ModelsDmitry Baranchuk, Andrey Voynov, Ivan Rubachev, Valentin Khrulkov 等ICLR 2022 · 被引用 700 次
- EditGAN: High-Precision Semantic Image EditingHuan Ling, Karsten Kreis, Daiqing Li, Seung Wook Kim 等NeurIPS 2021 · 被引用 248 次
- BigDatasetGAN: Synthesizing ImageNet with Pixel-wise AnnotationsDaiqing Li, Huan Ling, Seung Wook Kim, Karsten Kreis 等CVPR 2022 · 被引用 71 次
- EmerDiff: Emerging Pixel-level Semantic Knowledge in Diffusion ModelsKoichi Namekata, Amirmojtaba Sabour, Sanja Fidler, Seung Wook KimICLR 2024 · 被引用 41 次
- Householder Projector for Unsupervised Latent Semantics DiscoveryYue Song, Jichao Zhang, Nicu Sebe, Wei WangICCV 2023 · 被引用 9 次
它引用的顶会 Paper9
- Unsupervised Discovery of Interpretable Directions in the GAN Latent SpaceAndrey Voynov, Artem BabenkoICML 2020 · 被引用 459 次
- On the "steerability" of generative adversarial networksAli Jahanian, Lucy Chai, Phillip IsolaICLR 2020 · 被引用 421 次
- Image GANs meet Differentiable Rendering for Inverse Graphics and Interpretable 3D Neural RenderingYuxuan Zhang, Wenzheng Chen, Huan Ling, Jun Gao 等ICLR 2021 · 被引用 140 次
- Editing in Style: Uncovering the Local Semantics of GANsEdo Collins, Raja Bala, Bob Price, Sabine SüsstrunkCVPR 2020
- Closed-Form Factorization of Latent Semantics in GANsYujun Shen, Bolei ZhouCVPR 2021
相关 Paper
- FEditNet: Few-Shot Editing of Latent Semantics in GAN SpacesMengfei Xia, Yezhi Shu, Yuji Wang, Yu-Kun Lai 等AAAI 2023 · 被引用 4 次
- Repurposing GANs for One-Shot Semantic Part SegmentationNontawat Tritrong, Pitchaporn Rewatbowornwong, Supasorn SuwajanakornCVPR 2021
- Attribute Group Editing for Reliable Few-shot Image GenerationGuanqi Ding, Xinzhe Han, Shuhui Wang, Shuzhe Wu 等CVPR 2022 · 被引用 36 次
- Semantic Image Analogy with a Conditional Single-Image GANJiacheng Li, Zhiwei Xiong, Dong Liu, Xuejin Chen 等ACM MM 2020 · 被引用 4 次
- Interpreting the Latent Space of GANs for Semantic Face EditingYujun Shen, Jinjin Gu, Xiaoou Tang, Bolei ZhouCVPR 2020
