StyleGAN knows Normal, Depth, Albedo, and More
Anand Bhattad, Daniel McKee, Derek Hoiem, David A. Forsyth
Abstract
Intrinsic images, in the original sense, are image-like maps of scene properties like depth, normal, albedo or shading. This paper demonstrates that StyleGAN can easily be induced to produce intrinsic images. The procedure is straightforward. We show that, if StyleGAN produces from latents , then for each type of intrinsic image, there is a fixed offset so that is that type of intrinsic image for . Here is independent of . The StyleGAN we used was pretrained by others, so this property is not some accident of our training regime. We show that there are image transformations StyleGAN will not produce in this fashion, so StyleGAN is not a generic image regression engine. It is conceptually exciting that an image generator should ``know'' and represent intrinsic images. There may also be practical advantages to using a generative model to produce intrinsic images. The intrinsic images obtained from StyleGAN compare well both qualitatively and quantitatively with those obtained by using SOTA image regression techniques; but StyleGAN's intrinsic images are robust to relighting effects, unlike SOTA methods.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 50778274-805b-4bf3-a4b0-2d9b0d286d71Cited by top-tier papers25
- RGB↔X: Image decomposition and synthesis using material- and lighting-aware diffusion modelsZheng Zeng, Valentin Deschaintre, Iliyan Georgiev, Yannick Hold-Geoffroy et al.SIGGRAPH 2024 · 61 citations
- UniRelight: Learning Joint Decomposition and Synthesis for Video RelightingKai He, Ruofan Liang, Jacob Munkberg, Jon Hasselgren et al.NeurIPS 2025 · 42 citations
- A General Protocol to Probe Large Vision Models for 3D Physical UnderstandingGuanqi Zhan, Chuanxia Zheng, Weidi Xie, Andrew ZissermanNeurIPS 2024 · 37 citations
- Shadows Don't Lie and Lines Can't Bend! Generative Models Don't know Projective Geometry...for NowAyush Sarkar, Hanlin Mai, Amitabh Mahapatra, Svetlana Lazebnik et al.CVPR 2024 · 22 citations
- Latent Intrinsics Emerge from Training to RelightXiao Zhang, William Gao, Seemandhar Jain, Michael Maire et al.NeurIPS 2024 · 21 citations
Builds on37
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 13,211 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Adding Conditional Control to Text-to-Image Diffusion ModelsLvmin Zhang, Anyi Rao, Maneesh AgrawalaICCV 2023 · 6,759 citations
- Vision Transformers for Dense PredictionRené Ranftl, Alexey Bochkovskiy, Vladlen KoltunICCV 2021 · 2,647 citations
- Alias-Free Generative Adversarial NetworksTero Karras, Miika Aittala, Samuli Laine, Erik Härkönen et al.NeurIPS 2021 · 2,126 citations
Related papers
- LumiNet: Latent Intrinsics Meets Diffusion Models for Indoor Scene RelightingXiaoyan Xing, Konrad Groh, Sezer Karaoglu, Theo Gevers et al.CVPR 2025
- IntrinsicDiffusion: Joint Intrinsic Layers from Latent Diffusion ModelsJundan Luo, Duygu Ceylan, Jae Shin Yoon, Nanxuan Zhao et al.SIGGRAPH 2024 · 18 citations
- Self-Supervised Geometry-Aware Encoder for Style-Based 3D GAN InversionYushi Lan, Xuyi Meng, Shuai Yang, Chen Change Loy et al.CVPR 2023
- Only a matter of style: age transformation using a style-based regression modelYuval Alaluf, Or Patashnik, Daniel Cohen-OrSIGGRAPH 2021 · 143 citations
- StyleCineGAN: Landscape Cinemagraph Generation Using a Pre-trained StyleGANJongwoo Choi, Kwanggyoon Seo, Amirsaman Ashtari, Junyong NohCVPR 2024
