Optimizing Video Analytics with Declarative Model Relationships
Francisco Romero, Johann Hauswald, Aditi Partap, Daniel Kang, Matei Zaharia, Christos Kozyrakis
Abstract
The availability of vast video collections and the accuracy of ML models has generated significant interest in video analytics systems. Since naively processing all frames using expensive models is impractical, researchers have proposed optimizations such as selectively using faster but less accurate models to replace or filter frames for expensive models. However, these optimizations are difficult to apply on queries with multiple predicates and models, as users must manually explore a large optimization space. Without significant systems expertise or time investment, an analyst may manually create an execution plan that is unnecessarily expensive and/or terribly inaccurate.
We propose Relational Hints , a declarative interface that allows users to suggest ML model relationships based on domain knowledge. Users can express two key relationships: when a model can replace another (CAN REPLACE) and when a model can be used to filter frames for another (CAN FILTER). We aim to design an interface to express model relationships informed by domain specific knowledge and define the constraints by which these relationships hold. We then present the VIVA video analytics system that uses relational hints to optimize SQL queries on video datasets. VIVA automatically selects and validates the hints applicable to the query, generates possible query plans using a formal set of transformations, and finds the best performance plan that meets a user's accuracy requirements. VIVA relieves users from rewriting and manually optimizing video queries as new models become available and execution environments evolve. We evaluate VIVA implemented on top of Spark and show that hints improve performance up to 16.6X without sacrificing accuracy.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 5c3fd3f2-bd7b-44d3-bd9b-01cbe1adc72dCited by top-tier papers13
- Murakkab: Resource-Efficient Agentic Workflow Orchestration in Cloud PlatformsGohar Irfan Chaudhry, Esha Choukse, Haoran Qiu, Íñigo Goiri et al.OSDI 2026 · 26 citations
- Extract-Transform-Load for Video StreamsFerdinand Kossmann, Ziniu Wu, Eugenie Lai, Nesime Tatbul et al.VLDB 2023 · 21 citations
- Vulcan: Automatic Query Planning for Live ML AnalyticsYiwen Zhang, Xumiao Zhang, Ganesh Ananthanarayanan, Anand P. Iyer et al.NSDI 2024 · 17 citations
- VOCALExplore: Pay-as-You-Go Video Data Exploration and Model BuildingMaureen Daum, Enhao Zhang, Dong He, Stephen Mussmann et al.VLDB 2023 · 7 citations
- TVM: A Tile-based Video Management FrameworkTianxiong Zhong, Zhiwei Zhang, Guo Lu, Ye Yuan et al.VLDB 2024 · 6 citations
Builds on15
- INFaaS: Automated Model-less Inference ServingFrancisco Romero, Qian Li, Neeraja J. Yadwadkar, Christos KozyrakisUSENIX ATC 2021 · 325 citations
- BlazeIt: Optimizing Declarative Aggregation and Limit Queries for Neural Network-Based Video AnalyticsDaniel Kang, Peter Bailis, Matei ZahariaVLDB 2020 · 103 citations
- MIRIS: Fast Object Track Queries in VideoFavyen Bastani, Songtao He, Arjun Balasingam, Karthik Gopalakrishnan et al.SIGMOD 2020 · 68 citations
- Video Pose Distillation for Few-Shot, Fine-Grained Sports Action RecognitionJames Hong, Matthew Fisher, Michaël Gharbi, Kayvon FatahalianICCV 2021 · 54 citations
- Jointly Optimizing Preprocessing and Inference for DNN-based Visual AnalyticsDaniel Kang, Ankit Mathur, Teja Veeramacheneni, Peter Bailis et al.VLDB 2021 · 50 citations
Related papers
- EVA: A Symbolic Approach to Accelerating Exploratory Video Analytics with Materialized ViewsZhuangdi Xu, Gaurav Tarlok Kakkar, Joy Arulraj, Umakishore RamachandranSIGMOD 2022 · 26 citations
- A Method for Optimizing Opaque Filter QueriesWenjia He, Michael R. Anderson, Maxwell Strome, Michael J. CafarellaSIGMOD 2020 · 15 citations
- SketchQL: Video Moment Querying with a Visual Query InterfaceRenzhi Wu, Pramod Chunduri, Ali Payani, Xu Chu et al.SIGMOD 2025 · 6 citations
- FiGO: Fine-Grained Query Optimization in Video AnalyticsJiashen Cao, Karan Sarkar, Ramyad Hadidi, Joy Arulraj et al.SIGMOD 2022 · 37 citations
- Optimizing Dataflow Systems for Scalable Interactive VisualizationJunran Yang, Hyekang Kevin Joo, Sai S. Yerramreddy, Dominik Moritz et al.SIGMOD 2024 · 8 citations
