Fake it till you make it: face analysis in the wild using synthetic data alone
Erroll Wood, Tadas Baltrusaitis, Charlie Hewitt, Sebastian Dziadzio, Thomas J. Cashman, Jamie Shotton
Abstract
We demonstrate that it is possible to perform face-related computer vision in the wild using synthetic data alone. The community has long enjoyed the benefits of synthesizing training data with graphics, but the domain gap between real and synthetic data has remained a problem, especially for human faces. Researchers have tried to bridge this gap with data mixing, domain adaptation, and domain-adversarial training, but we show that it is possible to synthesize data with minimal domain gap, so that models trained on synthetic data generalize to real in-the-wild datasets. We describe how to combine a procedurally-generated parametric 3D face model with a comprehensive library of hand-crafted assets to render training images with unprecedented realism and diversity. We train machine learning systems for face-related tasks such as landmark localization and face parsing, showing that synthetic data can both match real data in accuracy as well as open up new approaches where manual labeling would be impossible. * Denotes equal contribution. https://microsoft.github.io/FaceSynthetics
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers67
- What does CLIP know about a red circle? Visual prompt engineering for VLMsAleksandar Shtedritski, Christian Rupprecht, Andrea VedaldiICCV 2023 · 262 citations
- Kubric: A scalable dataset generatorKlaus Greff, Francois Belletti, Lucas Beyer, Carl Doersch et al.CVPR 2022 · 183 citations
- IDiff-Face: Synthetic-based Face Recognition through Fizzy Identity-Conditioned Diffusion ModelsFadi Boutros, Jonas Henry Grebe, Arjan Kuijper, Naser DamerICCV 2023 · 106 citations
- Real-Time Radiance Fields for Single-Image Portrait View SynthesisAlex Trevithick, Matthew A. Chan, Michael Stengel, Eric R. Chan et al.SIGGRAPH 2023 · 69 citations
- GAIA: Zero-shot Talking Avatar GenerationTianyu He, Junliang Guo, Runyi Yu, Yuchi Wang et al.ICLR 2024 · 51 citations
Builds on8
- PIFu: Pixel-Aligned Implicit Function for High-Resolution Clothed Human DigitizationShunsuke Saito, Zeng Huang, Ryota Natsume, Shigeo Morishima et al.ICCV 2019 · 1,411 citations
- Adaptive Wing Loss for Robust Face Alignment via Heatmap RegressionXinyao Wang, Liefeng Bo, Fuxin LiICCV 2019 · 293 citations
- A New Dataset and Boundary-Attention Semantic Segmentation for Face ParsingYinglu Liu, Hailin Shi, Hao Shen, Yue Si et al.AAAI 2020 · 88 citations
- DF2Net: A Dense-Fine-Finer Network for Detailed 3D Face ReconstructionXiaoxing Zeng, Xiaojiang Peng, Yu QiaoICCV 2019 · 85 citations
- Deep Head Pose Estimation Using Synthetic Images and Partial Adversarial Domain Adaption for Continuous Label SpacesFelix Kuhnke, Jörn OstermannICCV 2019 · 51 citations
Related papers
- SynFace: Face Recognition with Synthetic DataHaibo Qiu, Baosheng Yu, Dihong Gong, Zhifeng Li et al.ICCV 2021 · 162 citations
- Analyzing the Synthetic-to-Real Domain Gap in 3D Hand Pose EstimationZhuoran Zhao, Linlin Yang, Pengzhan Sun, Pan Hui et al.CVPR 2025
- DAViD: Data-Efficient and Accurate Vision Models from Synthetic Data DAViD also references Michelangelo's David - an iconic symbol of anatomical precision-and the David vs. Goliath story, reflecting our small yet powerful dataset and modelsFatemeh Saleh, Sadegh Aliakbarian, Charlie Hewitt, Lohit Petikam et al.ICCV 2025 · 3 citations
- WildCAT3D: Appearance-Aware Multi-View Diffusion in the WildMorris Alper, David Novotný, Filippos Kokkinos, Hadar Averbuch-Elor et al.NeurIPS 2025 · 2 citations
- Demodalizing Face Recognition with Synthetic SamplesZhonghua Zhai, Pengju Yang, Xiaofeng Zhang, Maji Huang et al.AAAI 2021 · 9 citations
