Unsupervised Gaze Representation Learning from Multi-view Face Images
Yiwei Bao, Feng Lu
Abstract
Annotating gaze is an expensive and time-consuming endeavor, requiring costly eye-trackers or complex geometric calibration procedures. Although some eye-based unsupervised gaze representation learning methods have been proposed, the quality of gaze representation extracted by these methods degrades severely when the head pose is large. In this paper, we present the Multi-View Dual-Encoder (MV-DE), a framework designed to learn gaze representations from unlabeled multi-view face images. Through the proposed Dual-Encoder architecture and the multi-view gaze representation swapping strategy, the MV-DE successfully disentangles gaze from general facial information and derives gaze representations closely tied to the subject's eyeball rotation without gaze label. Experimental results illustrate that the gaze representations learned by the MV-DE can be used in downstream tasks, including gaze estimation and redirection. Gaze estimation results indicates that the proposed MV-DE displays notably higher robustness to uncontrolled head movements when compared to state-of-the-art (SOTA) unsupervised learning methods.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 6334e375-7304-4a9d-a0a8-3a2ff8ff7d18Cited by top-tier papers6
- OmniGaze: Reward-inspired Generalizable Gaze Estimation in the WildHongyu Qu, Jianan Wei, Xiangbo Shu, Yazhou Yao et al.NeurIPS 2025 · 15 citations
- Multi-View Gaze Target EstimationQiaomu Miao, Vivek Raju Golani, Jingyi Xu, Progga Paromita Dutta et al.ICCV 2025 · 4 citations
- A Generalized Label Shift Perspective for Cross-Domain Gaze EstimationHaoran Yang, Xiaohui Chen, Chuan-Xian RenNeurIPS 2025
- Semi-Supervised Gaze Estimation via Disentangled Subspace Contrastive LearningQida Tan, Hongyu Yang, Wenchao DuICML 2026
- 3D Prior Is All You Need: Cross-Task Few-shot 2D Gaze EstimationYihua Cheng, Hengfei Wang, Zhongqun Zhang, Yang Yue et al.CVPR 2025
Builds on13
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 24,064 citations
- Bootstrap Your Own Latent - A New Approach to Self-Supervised LearningJean-Bastien Grill, Florian Strub, Florent Altché, Corentin Tallec et al.NeurIPS 2020 · 9,171 citations
- Gaze360: Physically Unconstrained Gaze Estimation in the WildPetr Kellnhofer, Adrià Recasens, Simon Stent, Wojciech Matusik et al.ICCV 2019 · 469 citations
- A Coarse-to-Fine Adaptive Network for Appearance-Based Gaze EstimationYihua Cheng, Shiyao Huang, Fei Wang, Chen Qian et al.AAAI 2020 · 204 citations
- PureGaze: Purifying Gaze Feature for Generalizable Gaze EstimationYihua Cheng, Yiwei Bao, Feng LuAAAI 2022 · 121 citations
Related papers
- Cross-Encoder for Unsupervised Gaze Representation LearningYunjia Sun, Jiabei Zeng, Shiguang Shan, Xilin ChenICCV 2021 · 40 citations
- Unsupervised Representation Learning for Gaze EstimationYu Yu, Jean-Marc OdobezCVPR 2020
- UVAGaze: Unsupervised 1-to-2 Views Adaptation for Gaze EstimationRuicong Liu, Feng LuAAAI 2024 · 8 citations
- Self-Learning Transformations for Improving Gaze and Head RedirectionYufeng Zheng, Seonwook Park, Xucong Zhang, Shalini De Mello et al.NeurIPS 2020 · 50 citations
- Gaze from Origin: Learning for Generalized Gaze Estimation by Embedding the Gaze Frontalization ProcessMingjie Xu, Feng LuAAAI 2024 · 10 citations
