Orthogonal Jacobian Regularization for Unsupervised Disentanglement in Image Generation
Yuxiang Wei, Yupeng Shi, Xiao Liu, Zhilong Ji, Yuan Gao, Zhongqin Wu, Wangmeng Zuo
摘要
Unsupervised disentanglement learning is a crucial issue for understanding and exploiting deep generative models. Recently, SeFa tries to find latent disentangled directions by performing SVD on the first projection of a pretrained GAN. However, it is only applied to the first layer and works in a post-processing way. Hessian Penalty minimizes the off-diagonal entries of the output's Hessian matrix to facilitate disentanglement, and can be applied to multi-layers. However, it constrains each entry of output independently, making it not sufficient in disentangling the latent directions (e.g., shape, size, rotation, etc.) of spatially correlated variations. In this paper, we propose a simple Orthogonal Jacobian Regularization (OroJaR) to encourage deep generative model to learn disentangled representations. It simply encourages the variation of output caused by perturbations on different latent dimensions to be orthogonal, and the Jacobian with respect to the input is calculated to represent this variation. We show that our OroJaR also encourages the output's Hessian matrix to be diagonal in an indirect manner. In contrast to the Hessian Penalty, our OroJaR constrains the output in a holistic way, making it very effective in disentangling latent dimensions corresponding to spatially correlated variations. Quantitative and qualitative experimental results show that our method is effective in disentangled and controllable image generation, and performs favorably against the state-of-the-art methods. Our code is available at https://github.com/csyxwei/OroJaR .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper24
- Content-Aware Local GAN for Photo-Realistic Super-ResolutionJoonKyu Park, Sanghyun Son, Kyoung Mu LeeICCV 2023 · 被引用 73 次
- StyleT2I: Toward Compositional and High-Fidelity Text-to-Image SynthesisZhiheng Li, Martin Renqiang Min, Kai Li, Chenliang XuCVPR 2022 · 被引用 38 次
- Attribute Group Editing for Reliable Few-shot Image GenerationGuanqi Ding, Xinzhe Han, Shuhui Wang, Shuzhe Wu 等CVPR 2022 · 被引用 36 次
- LinkGAN: Linking GAN Latents to Pixels for Controllable Image SynthesisJiapeng Zhu, Ceyuan Yang, Yujun Shen, Zifan Shi 等ICCV 2023 · 被引用 29 次
- Principal Component FlowsEdmond Cunningham, Adam D. Cobb, Susmit JhaICML 2022 · 被引用 18 次
它引用的顶会 Paper7
- GANSpace: Discovering Interpretable GAN ControlsErik Härkönen, Aaron Hertzmann, Jaakko Lehtinen, Sylvain ParisNeurIPS 2020 · 被引用 1,049 次
- Unsupervised Discovery of Interpretable Directions in the GAN Latent SpaceAndrey Voynov, Artem BabenkoICML 2020 · 被引用 459 次
- HoloGAN: Unsupervised Learning of 3D Representations From Natural ImagesThu Nguyen-Phuoc, Chuan Li, Lucas Theis, Christian Richardt 等ICCV 2019 · 被引用 98 次
- Closed-Form Factorization of Latent Semantics in GANsYujun Shen, Bolei ZhouCVPR 2021
- Analyzing and Improving the Image Quality of StyleGANTero Karras, Samuli Laine, Miika Aittala, Janne Hellsten 等CVPR 2020
相关 Paper
- OOGAN: Disentangling GAN with One-Hot Sampling and Orthogonal RegularizationBingchen Liu, Yizhe Zhu, Zuohui Fu, Gerard de Melo 等AAAI 2020 · 被引用 42 次
- Householder Projector for Unsupervised Latent Semantics DiscoveryYue Song, Jichao Zhang, Nicu Sebe, Wei WangICCV 2023 · 被引用 9 次
- Learning Disentangled Representation by Exploiting Pretrained Generative Models: A Contrastive Learning ViewXuanchi Ren, Tao Yang, Yuwang Wang, Wenjun ZengICLR 2022 · 被引用 54 次
- Unsupervised Disentanglement Without Compromises : How Functional Orthogonality Enforces IdentifiabilityMathieu Simon, Pascal Frossard, Christophe De VleeschouwerICML 2026
- Unsupervised Disentanglement with Tensor Product Representations on the TorusMichael Rotman, Amit Dekel, Shir Gur, Yaron Oz 等ICLR 2022 · 被引用 5 次
