Pixel Cube: Diffusion-based Portrait Video Relighting Through Realistic Lighting Reproduction
Yufan Zhang, Yu Ji, Ayo Ajiboye, Rundi Wu, Yu Guo, Changxi Zheng, Jinwei Ye
摘要
We present a diffusion-based method for relighting dynamic portrait videos with photorealism and temporal consistency. Our method is fueled by a hybrid training dataset that consists of real-captured and rendered dynamic portrait videos with diverse subject appearances, facial motions, head poses, and known lighting conditions. Specifically, we construct an LED-based lighting system for realistic lighting emulation and high-speed video relighting data acquisition. By leveraging the image priors embedded in pre-trained video diffusion models, and using per-frame high dynamic range (HDR) environment map as lighting control, we train a high-performance generative model for realistic and identity-preserving dynamic portrait video relighting. In addition to the environment map control, our model uses a synthesized background image to enable control on the camera's exposure level and color tone. Our model can produce temporally consistent relit portrait video that looks realistic and harmonious under a provided new environment and faithfully preserve the subject's expression and fine facial features, including skin tone, wrinkles, and facial hair. Our model generalizes well to unseen data, in terms of the subject appearance, motion, and lighting condition. We perform extensive experiments on relighting in-the-wild videos with various environment maps and demonstrate practical applications on portrait photography. Results show that our method achieves state-of-the-art performance in photorealism, lighting harmony, and temporal consistency. Our project page: https://yufanzhang82.github.io/PixelCube/.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper28
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Elucidating the Design Space of Diffusion-Based Generative ModelsTero Karras, Miika Aittala, Timo Aila, Samuli LaineNeurIPS 2022 · 被引用 3,959 次
- Video Diffusion ModelsJonathan Ho, Tim Salimans, Alexey A. Gritsenko, William Chan 等NeurIPS 2022 · 被引用 2,948 次
- DreamFusion: Text-to-3D using 2D DiffusionBen Poole, Ajay Jain, Jonathan T. Barron, Ben MildenhallICLR 2023 · 被引用 463 次
- DiffIR: Efficient Diffusion Model for Image RestorationBin Xia, Yulun Zhang, Shiyin Wang, Yitong Wang 等ICCV 2023 · 被引用 410 次
相关 Paper
- Lux Post Facto: Learning Portrait Performance Relighting with Conditional Video Diffusion and a Hybrid DatasetYiqun Mei, Mingming He, Li Ma, Julien Philip 等CVPR 2025
- Total relighting: learning to relight portraits for background replacementRohit Pandey, Sergio Orts-Escolano, Chloe LeGendre, Christian Häne 等SIGGRAPH 2021 · 被引用 138 次
- Relightful Harmonization: Lighting-Aware Portrait Background ReplacementMengwei Ren, Wei Xiong, Jae Shin Yoon, Zhixin Shu 等CVPR 2024 · 被引用 17 次
- Comprehensive Relighting: Generalizable and Consistent Monocular Human Relighting and HarmonizationJunying Wang, Jingyuan Liu, Xin Sun, Krishna Kumar Singh 等CVPR 2025
- SynthLight: Portrait Relighting with Diffusion Model by Learning to Re-render Synthetic FacesSumit Chaturvedi, Mengwei Ren, Yannick Hold-Geoffroy, Jingyuan Liu 等CVPR 2025
