DiFaReli: Diffusion Face Relighting
Puntawat Ponglertnapakorn, Nontawat Tritrong, Supasorn Suwajanakorn
摘要
We introduce a novel approach to single-view face relighting in the wild, addressing challenges such as global illumination and cast shadows. A common scheme in recent methods involves intrinsicallydecomposing an input image into 3D shape, albedo, and lighting, then recomposing it with the target lighting. However, estimating these components is error-prone and requires many training examples with ground-truth lighting to generalize well. Our work bypasses the need for accurate intrinsic estimation and can be trained solely on 2D images without any light stage data, relit pairs, multi-view images, or lighting ground truth. Our key idea is to leverage a conditional diffusion implicit model (DDIM) for decoding a disentangled light encoding along with other encodings related to 3D shape and facial identity inferred from off-the-shelf estimators. We propose a novel conditioning technique that simplifies modeling the complex interaction between light and geometry. It uses a rendered shading reference along with a shadow map, inferred using a simple and effective technique, to spatially modulate the DDIM. Moreover, we propose a single-shot relighting framework that requires just one network pass, given pre-processed data, and even outperforms the teacher model across all metrics. Our method realistically relights in-the-wild images with temporally consistent cast shadows under varying lighting conditions. We achieve state-of-the-art performance on the standard benchmark Multi-PIE and rank highest in user studies. Please visit our page: https://diffusion-face-relighting-pp.github.io
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper26
- DiLightNet: Fine-grained Lighting Control for Diffusion-based Image GenerationChong Zeng, Yue Dong, Pieter Peers, Youkang Kong 等SIGGRAPH 2024 · 被引用 41 次
- Text2Relight: Creative Portrait Relighting with Text GuidanceJunuk Cha, Mengwei Ren, Krishna Kumar Singh, He Zhang 等AAAI 2025 · 被引用 9 次
- DiffTV: Identity-Preserved Thermal-to-Visible Face Translation via Feature Alignment and Dual-Stage ConditionsJingyu Lin, Guiqin Zhao, Jing Xu, Guoli Wang 等ACM MM 2024 · 被引用 9 次
- DreamLight: Towards Harmonious and Consistent Image RelightingYong Liu, Wenpeng Xiao, Qianqian Wang, Junlin Chen 等NeurIPS 2025 · 被引用 9 次
- CineVision: An Interactive Pre-visualization Storyboard System for Director-Cinematographer CollaborationZheng Wei, Hongtao Wu, Lvmin Zhang, Xian Xu 等UIST 2025 · 被引用 8 次
它引用的顶会 Paper40
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 被引用 13,211 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 被引用 11,743 次
相关 Paper
- Neural Gaffer: Relighting Any Object via DiffusionHaian Jin, Yuan Li, Fujun Luan, Yuanbo Xiangli 等NeurIPS 2024 · 被引用 112 次
- UniRelight: Learning Joint Decomposition and Synthesis for Video RelightingKai He, Ruofan Liang, Jacob Munkberg, Jon Hasselgren 等NeurIPS 2025 · 被引用 42 次
- GeoRelight: Learning Joint Geometrical Relighting and Reconstruction with Flexible Multi-Modal Diffusion TransformersYuxuan Xue, Ruofan Liang, Egor Zakharov, Timur M. Bagautdinov 等CVPR 2026 · 被引用 4 次
- IFS-Light: An Interactive Framework for Single-view Face Relighting with both Facial and Lighting ConsistencyShuyang Wang, Chunxiao Li, Anlong MingACM MM 2025
- End-to-End 3D Face Reconstruction with Expressions and Specular Albedos from Single In-the-wild ImagesQixin Deng, Binh Huy Le, Aobo Jin, Zhigang DengACM MM 2022
