Frame Complexity-Aware Foveated Video Encoding for Real-time High-Quality Streaming
Ze Wu, Ahmad Alhilal, Yuk Hang Tsui, Matti Siekkinen, Pan Hui
摘要
VR streaming and VR cloud gaming require high-resolution video streaming to provide users with high quality visual experience and maximize their interaction and immersion. Consequently, the video streams have a high bitrate and require a large amount of available bandwidth. Foveated video encoding (FVE) reduces bandwidth demand by selectively allocating higher quality to perceptually relevant regions based on human visual characteristics. However, scenes with high spatial detail or rapid motion introduce spatial and temporal frame complexities. Conventional video encoding assesses complexity and manages quality and bitrate through an internal rate-control mechanism. However, the SoTA FVE methods perform the quality allocation process after the rate control has already run. This may lead to rate violations and, consequently, under-or overutilization of available bandwidth, which in turn causes increased latency and/or reduced visual quality. In this paper, we present a real-time complexity-adaptive FVE method that minimizes computational latency using a GPU-accelerated compute shader. By prioritizing key spatial and temporal complexity variables, our weighted complexity estimation optimizes quality assignment. Our method outperforms the complexity-agnostic FVE benchmark with 44-78% greater bitrate stability, 21% lower latency, and 7% higher perceptual quality. It also surpasses the partial complexity-aware FVE benchmark, delivering 18-94% better network utilization alongside a 7% latency reduction and a 4-5% gain in quality. Our method also ensures generalizability across diverse scenarios.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
相关 Paper
- Saliency-Guided Foveated Video Encoding for Low-Latency and Immersive Cloud VRZe Wu, Ahmad Alhilal, Yuk Hang Tsui, Wen Jye Chai 等IEEE VR 2026
- FovRL: Joint Foveation and Quality Control for Immersive VR Streaming Using Reinforcement LearningYuk Hang Tsui, Ze Wu, Ahmad Alhilal, Matti Siekkinen 等WWW 2026 · 被引用 1 次
- Gaze-Adaptive Foveation for Remote Rendered VRAdhi Widagdo, Teemu Kämäräinen, Ahmad Alhilal, Matti Siekkinen 等ACM MM 2025
- Deep-Saliency Foveated Ray Tracing For Real-time VR RenderingYang Gao, Wencan Li, Shiyu Liang, Weizichuan Feng 等IEEE VR 2026
- Instant Reality: Gaze-Contingent Perceptual Optimization for 3D Virtual Reality StreamingShaoyu Chen, Budmonde Duinkharjav, Xin Sun, Li-Yi Wei 等IEEE VR 2022 · 被引用 25 次
