Understanding User Behavior in Volumetric Video Watching: Dataset, Analysis and Prediction
Kaiyuan Hu, Haowen Yang, Yili Jin, Junhua Liu, Yongting Chen, Miao Zhang, Fangxin Wang
Abstract
Volumetric video emerges as a new attractive video paradigm in recent years since it provides an immersive and interactive 3D viewing experience with six degree-of-freedom (DoF). Unlike traditional 2D or panoramic videos, volumetric videos require dense point clouds, voxels, meshes, or huge neural models to depict volumetric scenes, which results in a prohibitively high bandwidth burden for video delivery. Users' behavior analysis, especially the viewport and gaze analysis, then plays a significant role in prioritizing the content streaming within users' viewport and degrading the remaining content to maximize user QoE with limited bandwidth. Although understanding user behavior is crucial, to the best of our best knowledge, there are no available 3D volumetric video viewing datasets containing fine-grained user interactivity features, not to mention further analysis and behavior prediction.
In this paper, we for the first time release a volumetric video viewing behavior dataset, with a large scale, multiple dimensions, and diverse conditions. We conduct an in-depth analysis to understand user behaviors when viewing volumetric videos. Interesting findings on user viewport, gaze, and motion preference related to different videos and users are revealed. We finally design a transformerbased viewport prediction model that fuses the features of both gaze and motion, which is able to achieve high accuracy at various conditions. Our prediction model is expected to further benefit volumetric video streaming optimization.
Our dataset, along with the corresponding visualization tools is accessible at https://cuhksz-inml.github.io/user-behavior-in-vv-watching/
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 0fc26d68-5100-482d-91f7-6f964ff5b6e0Cited by top-tier papers1
Ask how each one uses itBuilds on6
- AI Choreographer: Music Conditioned 3D Dance Generation with AIST++Ruilong Li, Shan Yang, David A. Ross, Angjoo KanazawaICCV 2021 · 701 citations
- ViVo: visibility-aware mobile volumetric video streamingBo Han, Yu Liu, Feng QianMobiCom 2020 · 183 citations
- Using Reflexive Eye Movements for Fast Challenge-Response AuthenticationIvo Sluganovic, Marc Roeschlin, Kasper Bonne Rasmussen, Ivan MartinovicCCS 2016 · 93 citations
- Vues: practical mobile volumetric video streaming through multiview transcodingYu Liu, Bo Han, Feng Qian, Arvind Narayanan et al.MobiCom 2022 · 66 citations
- CaV3: Cache-assisted Viewport Adaptive Volumetric Video StreamingJunhua Liu, Boxiang Zhu, Fangxin Wang, Yili Jin et al.IEEE VR 2023 · 40 citations
Related papers
- Where Are You Looking?: A Large-Scale Dataset of Head and Gaze Behavior for 360-Degree Videos and a Pilot StudyYili Jin, Junhua Liu, Fangxin Wang, Shuguang CuiACM MM 2022 · 37 citations
- YuZu: Neural-Enhanced Volumetric Video StreamingAnlan Zhang, Chendong Wang, Bo Han, Feng QianNSDI 2022 · 115 citations
- Tile Classification Based Viewport Prediction with Multi-modal Fusion TransformerZhihao Zhang, Yiwei Chen, Weizhan Zhang, Caixia Yan et al.ACM MM 2023 · 14 citations
- PARIMA: Viewport Adaptive 360-Degree Video StreamingLovish Chopra, Sarthak Chakraborty, Abhijit Mondal, Sandip ChakrabortyWWW 2021 · 79 citations
- Kalman Filter-based Head Motion Prediction for Cloud-based Mixed RealitySerhan Gül, Sebastian Bosse, Dimitri Podborski, Thomas Schierl et al.ACM MM 2020 · 36 citations
