DVGaze: Dual-View Gaze Estimation
Yihua Cheng, Feng Lu
Abstract
Gaze estimation methods estimate gaze from facial appearance with a single camera. However, due to the limited view of a single camera, the captured facial appearance cannot provide complete facial information and thus complicate the gaze estimation problem. Recently, camera devices are rapidly updated. Dual cameras are affordable for users and have been integrated in many devices. This development suggests that we can further improve gaze estimation performance with dual-view gaze estimation. In this paper, we propose a dual-view gaze estimation network (DV-Gaze). DV-Gaze estimates dual-view gaze directions from a pair of images. We first propose a dual-view interactive convolution (DIC) block in DV-Gaze. DIC blocks exchange dual-view information during convolution in multiple feature scales. It fuses dual-view features along epipolar lines and compensates for the original feature with the fused feature. We further propose a dual-view transformer to estimate gaze from dual-view features. Camera poses are encoded to indicate the position information in the transformer. We also consider the geometric relation between dual-view gaze directions and propose a dual-view gaze consistency loss for DV-Gaze. DV-Gaze achieves state-of-the-art performance on ETH-XGaze and EVE datasets. Our experiments also prove the potential of dual-view gaze estimation. We release codes in https://github.com/yihuacheng/DVGaze .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 80ee5d14-14d3-4091-b7c2-09dacd75a18fCited by top-tier papers7
- UVAGaze: Unsupervised 1-to-2 Views Adaptation for Gaze EstimationRuicong Liu, Feng LuAAAI 2024 · 8 citations
- Multi-View Gaze Target EstimationQiaomu Miao, Vivek Raju Golani, Jingyi Xu, Progga Paromita Dutta et al.ICCV 2025 · 4 citations
- Bootstrap AutoEncoders With Contrastive Paradigm for Self-supervised Gaze EstimationYaoming Wang, Jin Li, Wenrui Dai, Bowen Shi et al.ICML 2024 · 2 citations
- What we Need is Explicit Controllability: Training 3D Gaze Estimator Using Only Facial ImagesTingwei Li, Jun Bao, Zhenzhong Kuang, Buyu LiuICCV 2025 · 1 citation
- Unsupervised Gaze Representation Learning from Multi-view Face ImagesYiwei Bao, Feng LuCVPR 2024
Builds on2
- A Coarse-to-Fine Adaptive Network for Appearance-Based Gaze EstimationYihua Cheng, Shiyao Huang, Fei Wang, Chen Qian et al.AAAI 2020 · 204 citations
- Kuiper Belt: Utilizing the "Out-of-natural Angle" Region in the Eye-gaze Interaction for Virtual RealityMyungguen Choi, Daisuke Sakamoto, Tetsuo OnoCHI 2022 · 14 citations
Related papers
- De^2Gaze: Deformable and Decoupled Representation Learning for 3D Gaze EstimationYunfeng Xiao, Xiaowei Bai, Baojun Chen, Hao Su et al.CVPR 2025
- Gaze-LLE: Gaze Target Estimation via Large-Scale Learned EncodersFiona Ryan, Ajay Bati, Sangmin Lee, Daniel Bolya et al.CVPR 2025
- Differential Contrastive Training for Gaze EstimationLin Zhang, Yi Tian, Xiyun Wang, Wanru Xu et al.ACM MM 2025 · 5 citations
- 3D Prior Is All You Need: Cross-Task Few-shot 2D Gaze EstimationYihua Cheng, Hengfei Wang, Zhongqun Zhang, Yang Yue et al.CVPR 2025
- FIFA: Fine-grained Inter-frame Attention for Driver's Video Gaze EstimationDaosong Hu, Mingyue Cui, Kai HuangCVPR 2025
