FixationNet: Forecasting Eye Fixations in Task-Oriented Virtual Environments
Zhiming Hu, Andreas Bulling, Sheng Li, Guoping Wang
Abstract
Fig. 1: Our model's eye fixation prediction performances in different scenes. The green dot represents the ground truth of eye fixation, the red dot denotes the result of our novel model, FixationNet, and the blue dot refers to the state-of-the-art method [21]. In practice, our model exhibits higher accuracy than the state-of-the-art method.
Abstract-Human visual attention in immersive virtual reality (VR) is key for many important applications, such as content design, gaze-contingent rendering, or gaze-based interaction. However, prior works typically focused on free-viewing conditions that have limited relevance for practical applications. We first collect eye tracking data of 27 participants performing a visual search task in four immersive VR environments. Based on this dataset, we provide a comprehensive analysis of the collected data and reveal correlations between users' eye fixations and other factors, i.e. users' historical gaze positions, task-related objects, saliency information of the VR content, and users' head rotation velocities. Based on this analysis, we propose FixationNet -a novel learning-based model to forecast users' eye fixations in the near future in VR. We evaluate the performance of our model for free-viewing and task-oriented settings and show that it outperforms the state of the art by a large margin of 19.8% (from a mean error of 2.93 • to 2.35 • ) in free-viewing and of 15.1% (from 2.05 • to 1.74 • ) in task-oriented situations. As such, our work provides new insights into task-oriented attention in virtual environments and guides future work on this important topic in VR research.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers14
- When XR and AI Meet - A Scoping Review on Extended Reality and Artificial IntelligenceTeresa Hirzle, Florian Müller, Fiona Draxler, Martin Schmitz et al.CHI 2023 · 90 citations
- ScanGAN360: A Generative Model of Realistic Scanpaths for 360° ImagesDaniel Martin, Ana Serrano, Alexander W. Bergman, Gordon Wetzstein et al.IEEE VR 2022 · 71 citations
- Eye Tracking-based LSTM for Locomotion Prediction in VRNiklas Stein, Gianni Bremer, Markus LappeIEEE VR 2022 · 39 citations
- PLUME: Record, Replay, Analyze and Share User Behavior in 6DoF XR ExperiencesCharles Javerliat, Sophie Villenave, Pierre Raimbaud, Guillaume LavouéIEEE VR 2024 · 35 citations
- Vergence Matching: Inferring Attention to Objects in 3D Environments for Gaze-Assisted SelectionLudwig Sidenmark, Christopher Clarke, Joshua Newn, Mathias N. Lystbæk et al.CHI 2023 · 29 citations
Builds on2
Related papers
- Fantastic Answers and Where to Find Them: Immersive Question-Directed Visual AttentionMing Jiang, Shi Chen, Jinhui Yang, Qi ZhaoCVPR 2020
- FovealNet: Advancing AI-Driven Gaze Tracking Solutions for Efficient Foveated Rendering in Virtual RealityWenxuan Liu, Budmonde Duinkharjav, Qi Sun, Sai Qian ZhangIEEE VR 2025 · 14 citations
- Looking but Not Focusing: Defining Gaze-Based Indices of Attention Lapses and Classifying Attentional StatesEugene Hwang, Jeongmi LeeCHI 2025 · 4 citations
- Deep-Saliency Foveated Ray Tracing For Real-time VR RenderingYang Gao, Wencan Li, Shiyu Liang, Weizichuan Feng et al.IEEE VR 2026
- Comparison of Visual Saliency for Dynamic Point Clouds: Task-free vs. Task-dependentXuemei Zhou, Irene Viola, Silvia Rossi, Pablo CésarIEEE VR 2025 · 5 citations
