High-Throughput Ingestion for Video Warehouse: Comprehensive Configuration and Effective Exploration
Baiyan Zhang, Zepeng Li, Dongxiang Zhang, Huan Li, Kian-Lee Tan, Gang Chen
Abstract
The innovative concept of Video Extract-Transform-Load (V-ETL), recently proposed in Skyscraper, reinterprets large-scale video analytics as a data warehousing problem. In this study, we aim at enabling real-time and high-throughput ingestion of hundreds of video streams and maximizing the overall accuracy, by constructing a proper ingestion plan for each video stream. To achieve the goal, we construct a comprehensive configuration space that takes into account the configurable components in the entire ingestion pipeline, including numeric parameters and categorical options such as visual inference model selection. The new space is 10 7 times larger than existing approaches, rendering them as sub-optimal points in our space. To effectively explore the huge and heterogeneous configuration space, we devise an accuracy-aware search strategy based on graph embedding and reinforcement learning to establish the runtime-quality Pareto frontier. To reduce the configuration exploration cost for all video streams, we cluster video streams with similar contexts and adopt mixed integer programming to maximize the overall ingestion accuracy while ensuring the real-time ingestion requirement. In the experimental evaluation with one NVIDIA GeForce RTX 4090 GPU card, our Hippo can support real-time ingestion with 300 video streams and secures an ingestion accuracy that exceeds its competitors by more than 30%.
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Your agent calls
Lunesearch_papers
Free to start. No credit card required.
Terminal
Install the CLIlune papers get bd334038-5e32-4b42-94f8-c8e614e254baRelated papers
- Extract-Transform-Load for Video StreamsFerdinand Kossmann, Ziniu Wu, Eugenie Lai, Nesime Tatbul et al.VLDB 2023 · 21 citations
- CASVA: Configuration-Adaptive Streaming for Live Video AnalyticsMiao Zhang, Fangxin Wang, Jiangchuan LiuINFOCOM 2022 · 69 citations
- Tackling the Imbalance in Video Analytics Pipelines with Hierarchical Embodied IntelligenceWenhui Zhou, Lei Xie, Jingyi Ning, Shuyu Cao et al.INFOCOM 2026
- Gecko: Resource-Efficient and Accurate Queries in Real-Time Video Streams at the EdgeLiang Wang, Xiaoyang Qu, Jianzong Wang, Guokuan Li et al.INFOCOM 2024 · 11 citations
- Batch Adaptative Streaming for Video AnalyticsLei Zhang, Yuqing Zhang, Ximing Wu, Fangxin Wang et al.INFOCOM 2022 · 24 citations
