Neural Reconstruction of Relightable Human Model from Monocular Video
Wenzhang Sun, Yunlong Che, Yandong Guo, Han Huang
Abstract
Creating relightable and animatable human characters from monocular video at a low cost is a critical task for digital human modeling and virtual reality applications. This task is complex due to intricate articulation motion, a wide range of ambient lighting conditions, and pose-dependent clothing deformations. In this paper, we introduce a novel self-supervised framework that takes a monocular video of a moving human as input and generates a 3D neural representation capable of being rendered with novel poses under arbitrary lighting conditions. Our framework decomposes dynamic humans under varying illumination into neural fields in canonical space, taking into account geometry and spatially varying BRDF material properties. Additionally, we introduce pose-driven deformation fields, enabling bidirectional mapping between canonical space and observation. Leveraging the proposed appearance decomposition and deformation fields, our framework learns in a self-supervised manner. Ultimately, based on pose-driven deformation, recovered appearance, and physically-based rendering, the reconstructed human figure becomes relightable and can be explicitly driven by novel poses. We demonstrate significant performance improvements over previous works and provide compelling examples of relighting from monocular videos of moving humans in challenging, uncontrolled capture scenarios.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 30a1572d-0e70-4d50-81aa-9e047bf37aebCited by top-tier papers7
- OccFusion: Rendering Occluded Humans with Generative Diffusion PriorsAdam Sun, Tiange Xiang, Scott L. Delp, Li Fei-Fei et al.NeurIPS 2024 · 13 citations
- IntrinsicAvatar: Physically Based Inverse Rendering of Dynamic Humans from Monocular Videos via Explicit Ray TracingShaofei Wang, Bozidar Antic, Andreas Geiger, Siyu TangCVPR 2024 · 13 citations
- NECA: Neural Customizable Human AvatarJunjin Xiao, Qing Zhang, Zhan Xu, Wei-Shi ZhengCVPR 2024 · 5 citations
- Relightable and Dynamic Gaussian Avatar Reconstruction from Monocular VideoSeonghwa Choi, Moonkyeong Choi, Mingyu Jang, Jaekyung Kim et al.ACM MM 2025 · 1 citation
- Generalizable and Relightable Gaussian Splatting for Human Novel View SynthesisYipengjing Sun, Shengping Zhang, Chenyang Wang, Shunyuan Zheng et al.SIGGRAPH 2026 · 1 citation
Builds on16
- Fourier Features Let Networks Learn High Frequency Functions in Low Dimensional DomainsMatthew Tancik, Pratul P. Srinivasan, Ben Mildenhall, Sara Fridovich-Keil et al.NeurIPS 2020 · 4,036 citations
- PIFu: Pixel-Aligned Implicit Function for High-Resolution Clothed Human DigitizationShunsuke Saito, Zeng Huang, Ryota Natsume, Shigeo Morishima et al.ICCV 2019 · 1,411 citations
- Multiview Neural Surface Reconstruction by Disentangling Geometry and AppearanceLior Yariv, Yoni Kasten, Dror Moran, Meirav Galun et al.NeurIPS 2020 · 1,010 citations
- StyleNeRF: A Style-based 3D Aware Generator for High-resolution Image SynthesisJiatao Gu, Lingjie Liu, Peng Wang, Christian TheobaltICLR 2022 · 622 citations
- HumanNeRF: Free-viewpoint Rendering of Moving People from Monocular VideoChung-Yi Weng, Brian Curless, Pratul P. Srinivasan, Jonathan T. Barron et al.CVPR 2022 · 411 citations
Related papers
- Relightable and Animatable Neural Avatars from VideosWenbin Lin, Chengwei Zheng, Jun-Hai Yong, Feng XuAAAI 2024 · 22 citations
- Physically Controllable Relighting of PhotographsChris Careaga, Yagiz AksoySIGGRAPH 2025 · 3 citations
- MonoHuman: Animatable Human Neural Field from Monocular VideoZhengming Yu, Wei Cheng, Xian Liu, Wayne Wu et al.CVPR 2023
- Vid2Avatar: 3D Avatar Reconstruction from Videos in the Wild via Self-supervised Scene DecompositionChen Guo, Tianjian Jiang, Xu Chen, Jie Song et al.CVPR 2023
- STaR: Self-Supervised Tracking and Reconstruction of Rigid Objects in Motion With Neural RenderingWentao Yuan, Zhaoyang Lv, Tanner Schmidt, Steven LovegroveCVPR 2021
