WildCap: Facial Albedo Capture in the Wild via Hybrid Inverse Rendering
Yuxuan Han, Xin Ming, Tianxiao Li, Zhuofan Shen, Qixuan Zhang, Lan Xu, Feng Xu
Abstract
Existing methods achieve high-quality facial appearance capture under controllable lighting, which increases capture cost and limits usability. We propose WildCap, a novel method for high-quality facial appearance capture from a smartphone video recorded in the wild. To disentangle high-quality reflectance from complex lighting effects in in-the-wild captures, we propose a novel hybrid inverse rendering framework. Specifically, we first apply a data-driven method, i.e., SwitchLight, to convert the captured images into more constrained conditions and then adopt model-based inverse rendering. However, unavoidable local artifacts in network predictions, such as shadow-baking, are non-physical and thus hinder accurate inverse rendering of lighting and material. To address this, we propose a novel texel grid lighting model to explain non-physical effects as clean albedo illuminated by local physical lighting. During optimization, we jointly sample a diffusion prior for reflectance maps and optimize the lighting, effectively resolving scale ambiguity between local lights and albedo. Our method achieves significantly better results than prior arts in the same capture setup, closing the quality gap between in-the-wild and controllable recordings by a large margin.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext d830a0d4-2236-49cc-8638-6641871d4e4cCited by top-tier papers1
Ask how each one uses itBuilds on38
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- 3D Gaussian Splatting for Real-Time Radiance Field RenderingBernhard Kerbl, Georgios Kopanas, Thomas Leimkühler, George DrettakisSIGGRAPH 2023 · 5,687 citations
- 2D Gaussian Splatting for Geometrically Accurate Radiance FieldsBinbin Huang, Zehao Yu, Anpei Chen, Andreas Geiger et al.SIGGRAPH 2024 · 660 citations
- NeRD: Neural Reflectance Decomposition from Image CollectionsMark Boss, Raphael Braun, Varun Jampani, Jonathan T. Barron et al.ICCV 2021 · 608 citations
- Extracting Triangular 3D Models, Materials, and Lighting From ImagesJacob Munkberg, Wenzheng Chen, Jon Hasselgren, Alex Evans et al.CVPR 2022 · 306 citations
Related papers
- WildLight: In-the-wild Inverse Rendering with a FlashlightZiang Cheng, Junxuan Li, Hongdong LiCVPR 2023
- Facial Appearance Capture at Home with Patch-Level Reflectance PriorYuxuan Han, Junfeng Lyu, Kuan Sheng, Minghao Que et al.SIGGRAPH 2025 · 2 citations
- End-to-End 3D Face Reconstruction with Expressions and Specular Albedos from Single In-the-wild ImagesQixin Deng, Binh Huy Le, Aobo Jin, Zhigang DengACM MM 2022
- Monocular Facial Appearance Capture in the WildYingyan Xu, Kate Gadola, Prashanth Chandran, Sebastian Weiss et al.ICCV 2025
- Real-Time 3D-Aware Portrait Video RelightingZiqi Cai, Kaiwen Jiang, Shu-Yu Chen, Yu-Kun Lai et al.CVPR 2024
