Where Are You Looking?: A Large-Scale Dataset of Head and Gaze Behavior for 360-Degree Videos and a Pilot Study
Yili Jin, Junhua Liu, Fangxin Wang, Shuguang Cui
摘要
360° videos in recent years have experienced booming development. Compared to traditional videos, 360° videos are featured with uncertain user behaviors, bringing opportunities as well as challenges. Datasets are necessary for researchers and developers to explore new ideas and conduct reproducible analyses for fair comparisons among different solutions. However, existing related datasets mostly focused on users' field of view (FoV), ignoring the more important eye gaze information, not to mention the integrated extraction and analysis of both FoV and eye gaze. Besides, users' behavior patterns are highly related to videos, yet most existing datasets only contained videos with subjective and qualitative classification from video genres, which lack quantitative analysis and fail to characterize the intrinsic properties of a video scene.
To this end, we first propose a quantitative taxonomy for 360° videos that contains three objective technical metrics. Based on this taxonomy, we collect a dataset containing users' head and gaze behaviors simultaneously, which outperforms existing datasets with rich dimensions, large scale, strong diversity, and high frequency. Then we conduct a pilot study on users' behaviors and get some interesting findings such as user's head direction will follow his/her gaze direction with the most possible time interval. A case of application in tile-based 360° video streaming based on our dataset is later conducted, demonstrating a great performance improvement of existing works by leveraging our provided gaze information.
Our dataset is available at https://cuhksz-inml.github.io/head_ gaze_dataset/
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- NetLLM: Adapting Large Language Models for NetworkingDuo Wu, Xianda Wang, Yaqi Qiao, Zhi Wang 等SIGCOMM 2024 · 被引用 162 次
- Understanding User Behavior in Volumetric Video Watching: Dataset, Analysis and PredictionKaiyuan Hu, Haowen Yang, Yili Jin, Junhua Liu 等ACM MM 2023 · 被引用 37 次
- GMMFormer: Gaussian-Mixture-Model Based Transformer for Efficient Partially Relevant Video RetrievalYuting Wang, Jinpeng Wang, Bin Chen, Ziyun Zeng 等AAAI 2024 · 被引用 32 次
- HeadsetOff: Enabling Photorealistic Video Conferencing on Economical VR HeadsetsYili Jin, Xize Duan, Fangxin Wang, Xue LiuACM MM 2024 · 被引用 5 次
- Branch Explorer: Leveraging Branching Narratives to Support Interactive 360° Video Viewing for Blind and Low Vision UsersShuchang Xu, Xiaofu Jin, Wenshuo Zhang, Huamin Qu 等UIST 2025 · 被引用 4 次
相关 Paper
- Tile Classification Based Viewport Prediction with Multi-modal Fusion TransformerZhihao Zhang, Yiwei Chen, Weizhan Zhang, Caixia Yan 等ACM MM 2023 · 被引用 14 次
- EyeQoE: A Novel QoE Assessment Model for 360-degree Videos Using Ocular BehaviorsHuadi Zhu, Tianhao Li, Chaowei Wang, Wenqiang Jin 等UbiComp 2022 · 被引用 11 次
- Looking here or there? Gaze Following in 360-Degree ImagesYunhao Li, Wei Shen, Zhongpai Gao, Yucheng Zhu 等ICCV 2021 · 被引用 24 次
- The Eye-Head Mover Spectrum: Modelling Individual and Population Head Movement Tendencies in Virtual RealityJinghui Hu, Ludwig Sidenmark, Hock Siang Lee, Hans GellersenCHI 2026 · 被引用 2 次
- CaV3: Cache-assisted Viewport Adaptive Volumetric Video StreamingJunhua Liu, Boxiang Zhu, Fangxin Wang, Yili Jin 等IEEE VR 2023 · 被引用 40 次
