Understanding User Behavior in Volumetric Video Watching: Dataset, Analysis and Prediction
Kaiyuan Hu, Haowen Yang, Yili Jin, Junhua Liu, Yongting Chen, Miao Zhang, Fangxin Wang
摘要
Volumetric video emerges as a new attractive video paradigm in recent years since it provides an immersive and interactive 3D viewing experience with six degree-of-freedom (DoF). Unlike traditional 2D or panoramic videos, volumetric videos require dense point clouds, voxels, meshes, or huge neural models to depict volumetric scenes, which results in a prohibitively high bandwidth burden for video delivery. Users' behavior analysis, especially the viewport and gaze analysis, then plays a significant role in prioritizing the content streaming within users' viewport and degrading the remaining content to maximize user QoE with limited bandwidth. Although understanding user behavior is crucial, to the best of our best knowledge, there are no available 3D volumetric video viewing datasets containing fine-grained user interactivity features, not to mention further analysis and behavior prediction.
In this paper, we for the first time release a volumetric video viewing behavior dataset, with a large scale, multiple dimensions, and diverse conditions. We conduct an in-depth analysis to understand user behaviors when viewing volumetric videos. Interesting findings on user viewport, gaze, and motion preference related to different videos and users are revealed. We finally design a transformerbased viewport prediction model that fuses the features of both gaze and motion, which is able to achieve high accuracy at various conditions. Our prediction model is expected to further benefit volumetric video streaming optimization.
Our dataset, along with the corresponding visualization tools is accessible at https://cuhksz-inml.github.io/user-behavior-in-vv-watching/
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper6
- AI Choreographer: Music Conditioned 3D Dance Generation with AIST++Ruilong Li, Shan Yang, David A. Ross, Angjoo KanazawaICCV 2021 · 被引用 701 次
- ViVo: visibility-aware mobile volumetric video streamingBo Han, Yu Liu, Feng QianMobiCom 2020 · 被引用 183 次
- Using Reflexive Eye Movements for Fast Challenge-Response AuthenticationIvo Sluganovic, Marc Roeschlin, Kasper Bonne Rasmussen, Ivan MartinovicCCS 2016 · 被引用 93 次
- Vues: practical mobile volumetric video streaming through multiview transcodingYu Liu, Bo Han, Feng Qian, Arvind Narayanan 等MobiCom 2022 · 被引用 66 次
- CaV3: Cache-assisted Viewport Adaptive Volumetric Video StreamingJunhua Liu, Boxiang Zhu, Fangxin Wang, Yili Jin 等IEEE VR 2023 · 被引用 40 次
相关 Paper
- Where Are You Looking?: A Large-Scale Dataset of Head and Gaze Behavior for 360-Degree Videos and a Pilot StudyYili Jin, Junhua Liu, Fangxin Wang, Shuguang CuiACM MM 2022 · 被引用 37 次
- YuZu: Neural-Enhanced Volumetric Video StreamingAnlan Zhang, Chendong Wang, Bo Han, Feng QianNSDI 2022 · 被引用 115 次
- Tile Classification Based Viewport Prediction with Multi-modal Fusion TransformerZhihao Zhang, Yiwei Chen, Weizhan Zhang, Caixia Yan 等ACM MM 2023 · 被引用 14 次
- PARIMA: Viewport Adaptive 360-Degree Video StreamingLovish Chopra, Sarthak Chakraborty, Abhijit Mondal, Sandip ChakrabortyWWW 2021 · 被引用 79 次
- Kalman Filter-based Head Motion Prediction for Cloud-based Mixed RealitySerhan Gül, Sebastian Bosse, Dimitri Podborski, Thomas Schierl 等ACM MM 2020 · 被引用 36 次
