Cross-Modal Deep Face Normals With Deactivable Skip Connections
Victoria Fernández Abrevaya, Adnane Boukhayma, Philip H. S. Torr, Edmond Boyer
Abstract
We present an approach for estimating surface normals from in-the-wild color images of faces. While datadriven strategies have been proposed for single face images, limited available ground truth data makes this problem difficult. To alleviate this issue, we propose a method that can leverage all available image and normal data, whether paired or not, thanks to a novel cross-modal learning architecture. In particular, we enable additional training with single modality data, either color or normal, by using two encoder-decoder networks with a shared latent space. The proposed architecture also enables face details to be transferred between the image and normal domains, given paired data, through skip connections between the image encoder and normal decoder. Core to our approach is a novel module that we call deactivable skip connections, which allows integrating both the auto-encoded and imageto-normal branches within the same architecture that can be trained end-to-end. This allows learning of a rich latent space that can accurately capture the normal information. We compare against state-of-the-art methods and show that our approach can achieve significant improvements, both quantitative and qualitative, with natural face images.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers11
- Learning an animatable detailed 3D face model from in-the-wild imagesYao Feng, Haiwen Feng, Michael J. Black, Timo BolkartSIGGRAPH 2021 · 662 citations
- Neural Head Avatars from Monocular RGB VideosPhilip-William Grassal, Malte Prinzler, Titus Leistner, Carsten Rother et al.CVPR 2022 · 173 citations
- Ensembling Off-the-shelf Models for GAN TrainingNupur Kumari, Richard Zhang, Eli Shechtman, Jun-Yan ZhuCVPR 2022 · 75 citations
- Geometry-aware Two-scale PIFu Representation for Human ReconstructionZheng Dong, Ke Xu, Ziheng Duan, Hujun Bao et al.NeurIPS 2022 · 17 citations
- Physically-guided Disentangled Implicit Rendering for 3D Face ModelingZhenyu Zhang, Yanhao Ge, Ying Tai, Weijian Cao et al.CVPR 2022 · 5 citations
Builds on3
- Tex2Shape: Detailed Full Human Body Geometry From a Single ImageThiemo Alldieck, Gerard Pons-Moll, Christian Theobalt, Marcus A. MagnorICCV 2019 · 343 citations
- Photo-Realistic Facial Details Synthesis From Single ImageAnpei Chen, Zhang Chen, Guli Zhang, Kenny Mitchell et al.ICCV 2019 · 113 citations
- FACSIMILE: Fast and Accurate Scans From an Image in Less Than a SecondDavid Smith, Matthew Loper, Xiaochen Hu, Paris Mavroidis et al.ICCV 2019 · 59 citations
Related papers
- SharinGAN: Combining Synthetic and Real Data for Unsupervised Geometry EstimationKoutilya PNVR, Hao Zhou, David JacobsCVPR 2020
- Hi3dgen: High-Fidelity 3D Geometry Generation From Images Via Normal BridgingChongjie Ye, Yushuang Wu, Ziteng Lu, Jiahao Chang et al.ICCV 2025 · 11 citations
- MNSRNet: Multimodal Transformer Network for 3D Surface Super-ResolutionWuyuan Xie, Tengcong Huang, Miaohui WangCVPR 2022 · 9 citations
- Towards High-Fidelity Face Normal EstimationMeng Wang, Chaoyue Wang, Xiaojie Guo, Jiawan ZhangACM MM 2022 · 1 citation
- SHS-Net: Learning Signed Hyper Surfaces for Oriented Normal Estimation of Point CloudsQing Li, Huifang Feng, Kanle Shi, Yue Gao et al.CVPR 2023
