Deep Head Pose Estimation Using Synthetic Images and Partial Adversarial Domain Adaption for Continuous Label Spaces
Felix Kuhnke, Jörn Ostermann
Abstract
Head pose estimation aims at predicting an accurate pose from an image. Current approaches rely on supervised deep learning, which typically requires large amounts of labeled data. Manual or sensor-based annotations of head poses are prone to errors. A solution is to generate synthetic training data by rendering 3D face models. However, the differences (domain gap) between rendered (source-domain) and real-world (target-domain) images can cause low performance. Advances in visual domain adaptation allow reducing the influence of domain differences using adversarial neural networks, which match the feature spaces between domains by enforcing domain-invariant features. While previous work on visual domain adaptation generally assumes discrete and shared label spaces, these assumptions are both invalid for pose estimation tasks. We are the first to present domain adaptation for head pose estimation with a focus on partially shared and continuous label spaces. More precisely, we adapt the predominant weighting approaches to continuous label spaces by applying a weighted resampling of the source domain during training. To evaluate our approach, we revise and extend existing datasets resulting in a new benchmark for visual domain adaption. Our experiments show that our method improves the accuracy of head pose estimation for real-world images despite using only labels from synthetic images.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 59b28862-26e3-476b-aa0c-7d881e0ca366Cited by top-tier papers9
- Fake it till you make it: face analysis in the wild using synthetic data aloneErroll Wood, Tadas Baltrusaitis, Charlie Hewitt, Sebastian Dziadzio et al.ICCV 2021 · 331 citations
- ArtiBoost: Boosting Articulated 3D Hand-Object Pose Estimation via Online Exploration and SynthesisLixin Yang, Kailin Li, Xinyu Zhan, Jun Lv et al.CVPR 2022 · 82 citations
- Generalizing Gaze Estimation with Outlier-guided Collaborative AdaptationYunfei Liu, Ruicong Liu, Haofei Wang, Feng LuICCV 2021 · 80 citations
- A Visual Analytics Approach to Facilitate the Proctoring of Online ExamsHaotian Li, Min Xu, Yong Wang, Huan Wei et al.CHI 2021 · 71 citations
- Interaction-aware Joint Attention Estimation Using People AttributesChihiro Nakatani, Hiroaki Kawashima, Norimichi UkitaICCV 2023 · 9 citations
Related papers
- Keypoint-Graph-Driven Learning Framework for Object Pose EstimationShaobo Zhang, Wanqing Zhao, Ziyu Guan, Xianlin Peng et al.CVPR 2021
- Adaptive Wasserstein Hourglass for Weakly Supervised RGB 3D Hand Pose EstimationYumeng Zhang, Li Chen, Yufeng Liu, Wen Zheng et al.ACM MM 2020 · 8 citations
- ONDA-Pose: Occlusion-Aware Neural Domain Adaptation for Self-Supervised 6D Object Pose EstimationTao Tan, Qiulei DongCVPR 2025
- Uncertainty-Aware Adaptation for Self-Supervised 3D Human Pose EstimationJogendra Nath Kundu, Siddharth Seth, Pradyumna YM, Varun Jampani et al.CVPR 2022 · 41 citations
- Lifelong Domain Adaptive 3D Human Pose EstimationQucheng Peng, Hongfei Xue, Pu Wang, Chen ChenAAAI 2026
