DiFaReli: Diffusion Face Relighting
Puntawat Ponglertnapakorn, Nontawat Tritrong, Supasorn Suwajanakorn
Abstract
We introduce a novel approach to single-view face relighting in the wild, addressing challenges such as global illumination and cast shadows. A common scheme in recent methods involves intrinsicallydecomposing an input image into 3D shape, albedo, and lighting, then recomposing it with the target lighting. However, estimating these components is error-prone and requires many training examples with ground-truth lighting to generalize well. Our work bypasses the need for accurate intrinsic estimation and can be trained solely on 2D images without any light stage data, relit pairs, multi-view images, or lighting ground truth. Our key idea is to leverage a conditional diffusion implicit model (DDIM) for decoding a disentangled light encoding along with other encodings related to 3D shape and facial identity inferred from off-the-shelf estimators. We propose a novel conditioning technique that simplifies modeling the complex interaction between light and geometry. It uses a rendered shading reference along with a shadow map, inferred using a simple and effective technique, to spatially modulate the DDIM. Moreover, we propose a single-shot relighting framework that requires just one network pass, given pre-processed data, and even outperforms the teacher model across all metrics. Our method realistically relights in-the-wild images with temporally consistent cast shadows under varying lighting conditions. We achieve state-of-the-art performance on the standard benchmark Multi-PIE and rank highest in user studies. Please visit our page: https://diffusion-face-relighting-pp.github.io
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext a428a950-cd71-455f-9d21-acba4b9f5cccCited by top-tier papers26
- DiLightNet: Fine-grained Lighting Control for Diffusion-based Image GenerationChong Zeng, Yue Dong, Pieter Peers, Youkang Kong et al.SIGGRAPH 2024 · 41 citations
- Text2Relight: Creative Portrait Relighting with Text GuidanceJunuk Cha, Mengwei Ren, Krishna Kumar Singh, He Zhang et al.AAAI 2025 · 9 citations
- DiffTV: Identity-Preserved Thermal-to-Visible Face Translation via Feature Alignment and Dual-Stage ConditionsJingyu Lin, Guiqin Zhao, Jing Xu, Guoli Wang et al.ACM MM 2024 · 9 citations
- DreamLight: Towards Harmonious and Consistent Image RelightingYong Liu, Wenpeng Xiao, Qianqian Wang, Junlin Chen et al.NeurIPS 2025 · 9 citations
- CineVision: An Interactive Pre-visualization Storyboard System for Director-Cinematographer CollaborationZheng Wei, Hongtao Wu, Lvmin Zhang, Xian Xu et al.UIST 2025 · 8 citations
Builds on40
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 13,211 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 11,743 citations
Related papers
- Neural Gaffer: Relighting Any Object via DiffusionHaian Jin, Yuan Li, Fujun Luan, Yuanbo Xiangli et al.NeurIPS 2024 · 112 citations
- UniRelight: Learning Joint Decomposition and Synthesis for Video RelightingKai He, Ruofan Liang, Jacob Munkberg, Jon Hasselgren et al.NeurIPS 2025 · 42 citations
- GeoRelight: Learning Joint Geometrical Relighting and Reconstruction with Flexible Multi-Modal Diffusion TransformersYuxuan Xue, Ruofan Liang, Egor Zakharov, Timur M. Bagautdinov et al.CVPR 2026 · 4 citations
- IFS-Light: An Interactive Framework for Single-view Face Relighting with both Facial and Lighting ConsistencyShuyang Wang, Chunxiao Li, Anlong MingACM MM 2025
- End-to-End 3D Face Reconstruction with Expressions and Specular Albedos from Single In-the-wild ImagesQixin Deng, Binh Huy Le, Aobo Jin, Zhigang DengACM MM 2022
