Normalized Avatar Synthesis Using StyleGAN and Perceptual Refinement
Huiwen Luo, Koki Nagano, Han-Wei Kung, Qingguo Xu, Zejian Wang, Lingyu Wei, Liwen Hu, Hao Li
Abstract
We introduce a highly robust GAN-based framework for digitizing a normalized 3D avatar of a person from a single unconstrained photo. While the input image can be of a smiling person or taken in extreme lighting conditions, our method can reliably produce a high-quality textured model of a person’s face in neutral expression and skin textures under diffuse lighting condition. Cutting-edge 3D face reconstruction methods use non-linear morphable face models combined with GAN-based decoders to capture the likeness and details of a person but fail to produce neutral head models with unshaded albedo textures which is critical for creating relightable and animation-friendly avatars for integration in virtual environments. The key challenges for existing methods to work is the lack of training and ground truth data containing normalized 3D faces. We propose a two-stage approach to address this problem. First, we adopt a highly robust normalized 3D face generator by embedding a non-linear morphable face model into a StyleGAN2 network. This allows us to generate detailed but normalized facial assets. This inference is then followed by a perceptual refinement step that uses the generated assets as regularization to cope with the limited available training samples of normalized faces. We further introduce a Normalized Face Dataset, which consists of a combination photogrammetry scans, carefully selected photographs, and generated fake people with neutral expressions in diffuse lighting conditions. While our prepared dataset contains two orders of magnitude less subjects than cutting edge GAN-based 3D facial reconstruction methods, we show that it is possible to produce high-quality normalized face models for very challenging unconstrained input images, and demonstrate superior performance to the current state-of-the-art.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext f88a8b14-96ed-42e5-9702-b0bb6311a268Cited by top-tier papers23
- ClipFace: Text-guided Editing of Textured 3D Morphable ModelsShivangi Aneja, Justus Thies, Angela Dai, Matthias NießnerSIGGRAPH 2023 · 40 citations
- Relightify: Relightable 3D Faces from a Single Image via Diffusion ModelsFoivos Paraperas Papantoniou, Alexandros Lattas, Stylianos Moschoglou, Stefanos ZafeiriouICCV 2023 · 40 citations
- HiFace: High-Fidelity 3D Face Reconstruction by Learning Static and Dynamic DetailsZenghao Chai, Tianke Zhang, Tianyu He, Xu Tan et al.ICCV 2023 · 33 citations
- 3D Facial Expressions through Analysis-by-Neural-SynthesisGeorge Retsinas, Panagiotis Paraskevas Filntisis, Radek Danecek, Victoria Fernández Abrevaya et al.CVPR 2024 · 26 citations
- HairMapper: Removing Hair from Portraits Using GANsYiqian Wu, Yong-Liang Yang, Xiaogang JinCVPR 2022 · 17 citations
Builds on13
- Image2StyleGAN: How to Embed Images Into the StyleGAN Latent Space?Rameen Abdal, Yipeng Qin, Peter WonkaICCV 2019 · 1,195 citations
- Probabilistic Face EmbeddingsYichun Shi, Anil K. JainICCV 2019 · 362 citations
- Photo-Realistic Facial Details Synthesis From Single ImageAnpei Chen, Zhang Chen, Guli Zhang, Kenny Mitchell et al.ICCV 2019 · 113 citations
- A Decoupled 3D Facial Shape Model by Adversarial TrainingVictoria Fernández Abrevaya, Adnane Boukhayma, Stefanie Wuhrer, Edmond BoyerICCV 2019 · 36 citations
- Image2StyleGAN++: How to Edit the Embedded Images?Rameen Abdal, Yipeng Qin, Peter WonkaCVPR 2020
Related papers
- StyleAvatar: Real-time Photo-realistic Portrait Avatar from a Single VideoLizhen Wang, Xiaochen Zhao, Jingxiang Sun, Yuxiang Zhang et al.SIGGRAPH 2023 · 49 citations
- MOST-GAN: 3D Morphable StyleGAN for Disentangled Face Image ManipulationSafa C. Medin, Bernhard Egger, Anoop Cherian, Ye Wang et al.AAAI 2022 · 38 citations
- Toonify3D: StyleGAN-based 3D Stylized Face GeneratorWonjong Jang, Yucheol Jung, Hyomin Kim, Gwangjin Ju et al.SIGGRAPH 2024 · 3 citations
- Self-Supervised Geometry-Aware Encoder for Style-Based 3D GAN InversionYushi Lan, Xuyi Meng, Shuai Yang, Chen Change Loy et al.CVPR 2023
- PhotoApp: photorealistic appearance editing of head portraitsMallikarjun B. R., Ayush Tewari, Abdallah Dib, Tim Weyrich et al.SIGGRAPH 2021 · 11 citations
