SpotStream: Real-Time Video Transmission for Autonomous Driving via Small Object-Aware ROI
Zelin Song, Huanhuan Zhang, Pengcheng Zhang, Mingyue Zhao, Congkai An, Anfu Zhou, Liang Liu
摘要
Cloud-based autonomous driving relies on low-latency video transmission for accurate downstream perception tasks. However, low-latency video streaming solutions, optimized for human Quality of Experience (QoE), are misaligned with the needs of machine vision tasks. While Region-of-Interest (ROI) encoding aims to bridge this gap, existing methods suffer from two critical limitations: first, insufficient protection for small and vulnerable objects, which are often missed by detection mechanisms or given uniform, inadequate resource allocation; second, prohibitive encoding overhead from fine-grained partitioning, leading to poor system robustness in dynamic network environments. To address these issues, we propose SpotStream, a novel low-latency video transmission framework. SpotStream features a dual-stream importance prediction network to ensure comprehensive coverage and prioritized protection for small objects. It couples this with an efficient, multi-level ROI assignment strategy that significantly reduces encoding overhead by simplifying region structure while focusing resources on critical targets. Comprehensive evaluations demonstrate that SpotStream substantially outperforms state-of-the-art baselines. Notably, under challenging real-world network traces, Spot-Stream improves the F1 score for critical small objects by an average of 12.9% while maintaining excellent system robustness.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
相关 Paper
- Saliency-Guided Foveated Video Encoding for Low-Latency and Immersive Cloud VRZe Wu, Ahmad Alhilal, Yuk Hang Tsui, Wen Jye Chai 等IEEE VR 2026
- RL-RC-DoT: A Block-level RL agent for Task-Aware Video CompressionUri Gadot, Assaf Shocher, Shie Mannor, Gal Chechik 等CVPR 2025
- Frame Complexity-Aware Foveated Video Encoding for Real-time High-Quality StreamingZe Wu, Ahmad Alhilal, Yuk Hang Tsui, Matti Siekkinen 等IEEE VR 2026
- FovRL: Joint Foveation and Quality Control for Immersive VR Streaming Using Reinforcement LearningYuk Hang Tsui, Ze Wu, Ahmad Alhilal, Matti Siekkinen 等WWW 2026 · 被引用 1 次
- FOVEA: Foveated Image Magnification for Autonomous NavigationChittesh Thavamani, Mengtian Li, Nicolas Cebron, Deva RamananICCV 2021 · 被引用 45 次
