LumiNet: Latent Intrinsics Meets Diffusion Models for Indoor Scene Relighting
Xiaoyan Xing, Konrad Groh, Sezer Karaoglu, Theo Gevers, Anand Bhattad
摘要
Abstract We introduce LumiNet, a novel architecture that leverages generative models and latent intrinsic representations for transferring lighting from one image to another. Given a source image and a target lighting image, LumiNet generates a relit version of the source scene that captures the target's lighting. Our approach makes two key contributions: a data curation strategy from the StyleGAN-based relighting model for our training, and a modified diffusion-based Con-trolNet that processes both latent intrinsic properties from the source image and latent extrinsic properties from the target image. We further improve lighting transfer through a learned adaptor that injects the target's latent extrinsic properties via cross-attention and light-weight fine-tuning. Unlike traditional ControlNet, which generates images with conditional maps from a single scene, LumiNet processes latent representations from two different imagespreserving geometry and albedo from the source while transferring lighting characteristics from the target. Experiments demonstrate that our method successfully transfers complex lighting phenomena including specular highlights and indirect illumination across scenes with varying spatial layouts and materials, outperforming existing approaches on challenging indoor scenes using only images as input.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper16
- WorldGen: From Text to Traversable and Interactive 3D WorldsDilin Wang, Hyunyoung Jung, Tom Monnier, Kihyuk Sohn 等CVPR 2026 · 被引用 24 次
- LuxRemix: Lighting Decomposition and Remixing for Indoor ScenesRuofan Liang, Norman Müller, Ethan Weber, Duncan Zauss 等CVPR 2026 · 被引用 7 次
- LightLab: Controlling Light Sources in Images with Diffusion ModelsNadav Magar, Amir Hertz, Eric Tabellion, Yael Pritch 等SIGGRAPH 2025 · 被引用 6 次
- FlowPortal: Residual-Corrected Flow for Training-Free Video Relighting and Background ReplacementWenshuo Gao, Junyi Fan, Jiangyue Zeng, Shuai YangCVPR 2026 · 被引用 6 次
- Generative Blocks World: Moving Things Around in PicturesVaibhav Vavilala, Seemandhar Jain, Rahul Vasanth, David Forsyth 等ICLR 2026 · 被引用 4 次
它引用的顶会 Paper32
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- A ConvNet for the 2020sZhuang Liu, Hanzi Mao, Chao-Yuan Wu, Christoph Feichtenhofer 等CVPR 2022 · 被引用 6,782 次
- Adding Conditional Control to Text-to-Image Diffusion ModelsLvmin Zhang, Anyi Rao, Maneesh AgrawalaICCV 2023 · 被引用 6,759 次
相关 Paper
- IntrinsicControlNet: Cross-Distribution Image Generation with Real and UnrealJiayuan Lu, Rengan Xie, Zixuan Xie, Zhizhen Wu 等ICCV 2025 · 被引用 4 次
- DiLightNet: Fine-grained Lighting Control for Diffusion-based Image GenerationChong Zeng, Yue Dong, Pieter Peers, Youkang Kong 等SIGGRAPH 2024 · 被引用 41 次
- Latent Intrinsics Emerge from Training to RelightXiao Zhang, William Gao, Seemandhar Jain, Michael Maire 等NeurIPS 2024 · 被引用 21 次
- GeoRelight: Learning Joint Geometrical Relighting and Reconstruction with Flexible Multi-Modal Diffusion TransformersYuxuan Xue, Ruofan Liang, Egor Zakharov, Timur M. Bagautdinov 等CVPR 2026 · 被引用 4 次
- StyleGAN knows Normal, Depth, Albedo, and MoreAnand Bhattad, Daniel McKee, Derek Hoiem, David A. ForsythNeurIPS 2023 · 被引用 61 次
