Self-Learning Transformations for Improving Gaze and Head Redirection
Yufeng Zheng, Seonwook Park, Xucong Zhang, Shalini De Mello, Otmar Hilliges
Abstract
Many computer vision tasks rely on labeled data. Rapid progress in generative modeling has led to the ability to synthesize photorealistic images. However, controlling specific aspects of the generation process such that the data can be used for supervision of downstream tasks remains challenging. In this paper we propose a novel generative model for images of faces, that is capable of producing high-quality images under fine-grained control over eye gaze and head orientation angles. This requires the disentangling of many appearance related factors including gaze and head orientation but also lighting, hue etc. We propose a novel architecture which learns to discover, disentangle and encode these extraneous variations in a self-learned manner. We further show that explicitly disentangling task-irrelevant factors results in more accurate modelling of gaze and head orientation. A novel evaluation scheme shows that our method improves upon the state-of-the-art in redirection accuracy and disentanglement between gaze direction and head orientation changes. Furthermore, we show that in the presence of limited amounts of real-world training data, our method allows for improvements in the downstream task of semi-supervised cross-dataset gaze estimation. Please check our project page at: this https URL
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 1d0f8c77-68ba-4f90-9050-8a59bee4dc05Cited by top-tier papers14
- PureGaze: Purifying Gaze Feature for Generalizable Gaze EstimationYihua Cheng, Yiwei Bao, Feng LuAAAI 2022 · 121 citations
- Generalizing Gaze Estimation with Rotation ConsistencyYiwei Bao, Yunfei Liu, Haofei Wang, Feng LuCVPR 2022 · 54 citations
- EyeNeRF: a hybrid representation for photorealistic synthesis, animation and relighting of human eyesGengyan Li, Abhimitra Meka, Franziska Mueller, Marcel C. Bühler et al.SIGGRAPH 2022 · 39 citations
- Gaze from Origin: Learning for Generalized Gaze Estimation by Embedding the Gaze Frontalization ProcessMingjie Xu, Feng LuAAAI 2024 · 10 citations
- PrivateGaze: Preserving User Privacy in Black-box Mobile Gaze Tracking ServicesLingyu Du, Jinyuan Jia, Xucong Zhang, Guohao LanUbiComp 2024 · 9 citations
Builds on7
- Few-Shot Adaptive Gaze EstimationSeonwook Park, Shalini De Mello, Pavlo Molchanov, Umar Iqbal et al.ICCV 2019 · 238 citations
- RelGAN: Multi-Domain Image-to-Image Translation via Relative AttributesYu-Jing Lin, Po-Wei Wu, Che-Han Chang, Edward Y. Chang et al.ICCV 2019 · 158 citations
- Understanding Human Gaze Communication by Spatio-Temporal Graph ReasoningLifeng Fan, Wenguan Wang, Song-Chun Zhu, Xinyu Tang et al.ICCV 2019 · 124 citations
- HoloGAN: Unsupervised Learning of 3D Representations From Natural ImagesThu Nguyen-Phuoc, Chuan Li, Lucas Theis, Christian Richardt et al.ICCV 2019 · 98 citations
- Monocular Neural Image Based Rendering With Continuous View ControlJie Song, Xu Chen, Otmar HilligesICCV 2019 · 85 citations
Related papers
- Unsupervised Gaze Representation Learning from Multi-view Face ImagesYiwei Bao, Feng LuCVPR 2024
- Controllable Continuous Gaze RedirectionWeihao Xia, Yujiu Yang, Jing-Hao Xue, Wensen FengACM MM 2020 · 13 citations
- What we Need is Explicit Controllability: Training 3D Gaze Estimator Using Only Facial ImagesTingwei Li, Jun Bao, Zhenzhong Kuang, Buyu LiuICCV 2025 · 1 citation
- Cross-Encoder for Unsupervised Gaze Representation LearningYunjia Sun, Jiabei Zeng, Shiguang Shan, Xilin ChenICCV 2021 · 40 citations
- RTGaze: Real-Time 3D-Aware Gaze Redirection from a Single ImageHengfei Wang, Zhongqun Zhang, Yihua Cheng, Hyung Jin ChangAAAI 2026
