HUST: High-Fidelity Unbiased Skin Tone Estimation via Texture Quantization
Zimin Ran, Xingyu Ren, Xiang An, Kaicheng Yang, Ziyong Feng, Jing Yang, Rolandos Alexandros Potamias, Linchao Zhu, Jiankang Deng
Abstract
Recent 3D facial reconstruction methods have made significant progress in shape estimation, but high-fidelity unbiased facial albedo estimation remains challenging. Existing methods rely on expensive light-stage captured data, and although they have made progress in either high-fidelity reconstruction or unbiased skin tone estimation, no work has yet achieved optimal results in both aspects simultaneously.
In this paper, we present a novel high-fidelity unbiased facial diffuse albedo reconstruction method, HUST, which recovers the diffuse albedo map directly from a single image without the need for captured data. Our key insight is that the albedo map is the illumination-invariant texture map, which enables us to use inexpensive texture data for diffuse albedo estimation by eliminating illumination. To achieve this, we collect large-scale high-resolution facial images and train a VQGAN model in the image space. To adapt the pre-trained VQGAN model for UV texture generation, we fine-tune the encoder by using limited UV textures and our high-resolution faces under adversarial supervision in both image and latent space. Finally, we train a cross-attention module and utilize group identity loss to adapt the domain from texture to albedo. Extensive experiments demonstrate that HUST can predict high-fidelity facial albedos for inthe-wild images. On the FAIR benchmark, HUST achieves
This ICCV paper is the Open Access version, provided by the Computer Vision Foundation. Except for this watermark, it is identical to the accepted version; the final published version of the proceedings is available on IEEE Xplore.
the lowest average ITA error (11.20) and bias score (1.58), demonstrating superior accuracy and robust fairness across the entire spectrum of human skin tones. Our project page is https://ziminran.github.io/hust-iccv/.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 1adaba73-60e4-4966-9614-29fd2fde7c72Builds on23
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Learning an animatable detailed 3D face model from in-the-wild imagesYao Feng, Haiwen Feng, Michael J. Black, Timo BolkartSIGGRAPH 2021 · 662 citations
- Towards Robust Blind Face Restoration with Codebook Lookup TransformerShangchen Zhou, Kelvin C. K. Chan, Chongyi Li, Chen Change LoyNeurIPS 2022 · 431 citations
- Killing Two Birds with One Stone: Efficient and Robust Training of Face Recognition CNNs by Partial FCXiang An, Jiankang Deng, Jia Guo, Ziyong Feng et al.CVPR 2022 · 77 citations
- DreamFace: Progressive Generation of Animatable 3D Faces under Text GuidanceLongwen Zhang, Qiwei Qiu, Hongyang Lin, Qixuan Zhang et al.SIGGRAPH 2023 · 68 citations
Related papers
- Improving Fairness in Facial Albedo Estimation via Visual-Textual CuesXingyu Ren, Jiankang Deng, Chao Ma, Yichao Yan et al.CVPR 2023
- Relightify: Relightable 3D Faces from a Single Image via Diffusion ModelsFoivos Paraperas Papantoniou, Alexandros Lattas, Stylianos Moschoglou, Stefanos ZafeiriouICCV 2023 · 40 citations
- FreeUV: Ground-Truth-Free Realistic Facial UV Texture Recovery via Cross-Assembly Inference StrategyXingchao Yang, Takafumi Taketomi, Yuki Endo, Yoshihiro KanamoriCVPR 2025
- Monocular Facial Appearance Capture in the WildYingyan Xu, Kate Gadola, Prashanth Chandran, Sebastian Weiss et al.ICCV 2025
- Monocular Identity-Conditioned Facial Reflectance ReconstructionXingyu Ren, Jiankang Deng, Yuhao Cheng, Jia Guo et al.CVPR 2024 · 4 citations
