DGaze: CNN-Based Gaze Prediction in Dynamic Scenes
Zhiming Hu, Sheng Li, Congyi Zhang, Kangrui Yi, Guoping Wang, Dinesh Manocha
Abstract
We conduct novel analyses of users' gaze behaviors in dynamic virtual scenes and, based on our analyses, we present a novel CNN-based model called DGaze for gaze prediction in HMD-based applications. We first collect 43 users' eye tracking data in 5 dynamic scenes under free-viewing conditions. Next, we perform statistical analysis of our data and observe that dynamic object positions, head rotation velocities, and salient regions are correlated with users' gaze positions. Based on our analysis, we present a CNN-based model (DGaze) that combines object position sequence, head velocity sequence, and saliency features to predict users' gaze positions. Our model can be applied to predict not only realtime gaze positions but also gaze positions in the near future and can achieve better performance than prior method. In terms of realtime prediction, DGaze achieves a 22.0% improvement over prior method in dynamic scenes and obtains an improvement of 9.5% in static scenes, based on using the angular distance as the evaluation metric. We also propose a variant of our model called DGaze_ET that can be used to predict future gaze positions with higher precision by combining accurate past gaze data gathered using an eye tracker. We further analyze our CNN architecture and verify the effectiveness of each component in our model. We apply DGaze to gaze-contingent rendering and a game, and also present the evaluation results from a user study.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext cf78cc2c-5f40-4883-a8e6-a5d71086f34bCited by top-tier papers19
- When XR and AI Meet - A Scoping Review on Extended Reality and Artificial IntelligenceTeresa Hirzle, Florian Müller, Fiona Draxler, Martin Schmitz et al.CHI 2023 · 90 citations
- A privacy-preserving approach to streaming eye-tracking dataBrendan David-John, Diane Hosfelt, Kevin R. B. Butler, Eakta JainIEEE VR 2021 · 89 citations
- FixationNet: Forecasting Eye Fixations in Task-Oriented Virtual EnvironmentsZhiming Hu, Andreas Bulling, Sheng Li, Guoping WangIEEE VR 2021 · 76 citations
- ScanGAN360: A Generative Model of Realistic Scanpaths for 360° ImagesDaniel Martin, Ana Serrano, Alexander W. Bergman, Gordon Wetzstein et al.IEEE VR 2022 · 71 citations
- Digital Transformations of Classrooms in Virtual RealityHong Gao, Efe Bozkir, Lisa Hasenbein, Jens-Uwe Hahn et al.CHI 2021 · 69 citations
Related papers
- Deep-Saliency Foveated Ray Tracing For Real-time VR RenderingYang Gao, Wencan Li, Shiyu Liang, Weizichuan Feng et al.IEEE VR 2026
- vGaze: Implicit Saliency-Aware Calibration for Continuous Gaze Tracking on Mobile DevicesSongzhou Yang, Yuan He, Meng JinINFOCOM 2021 · 11 citations
- Eye Tracking-based LSTM for Locomotion Prediction in VRNiklas Stein, Gianni Bremer, Markus LappeIEEE VR 2022 · 39 citations
- A Spherical Convolution Approach for Learning Long Term Viewport Prediction in 360 Immersive VideoChenglei Wu, Ruixiao Zhang, Zhi Wang, Lifeng SunAAAI 2020 · 48 citations
- RTGaze: Real-Time 3D-Aware Gaze Redirection from a Single ImageHengfei Wang, Zhongqun Zhang, Yihua Cheng, Hyung Jin ChangAAAI 2026
