Panorama: A Data System for Unbounded Vocabulary Querying over Video
Yuhao Zhang, Arun Kumar
Abstract
Deep convolutional neural networks (CNNs) achieve state-of-the-art accuracy for many computer vision tasks. But using them for video monitoring applications incurs high computational cost and inference latency. Thus, recent works have studied how to improve system efficiency. But they largely focus on small "closed world" prediction vocabularies even though many applications in surveillance security, traffic analytics, etc. have an ever-growing set of target entities. We call this the "unbounded vocabulary" issue, and it is a key bottleneck for emerging video monitoring applications. We present the first data system for tacking this issue for video querying, Panorama. Our design philosophy is to build a unified and domain-agnostic system that lets application users generalize to unbounded vocabularies in an out-of-the-box manner without tedious manual re-training. To this end, we synthesize and innovate upon an array of techniques from the ML, vision, databases, and multimedia systems literature to devise a new system architecture. We also present techniques to ensure Panorama has high inference efficiency. Experiments with multiple real-world datasets show that Panorama can achieve between 2x to 20x higher efficiency than baseline approaches on in-vocabulary queries, while still yielding comparable accuracy and also generalizing well to unbounded vocabularies.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext c33acd1f-8122-4971-826c-dcee4b420d3eCited by top-tier papers14
- FiGO: Fine-Grained Query Optimization in Video AnalyticsJiashen Cao, Karan Sarkar, Ramyad Hadidi, Joy Arulraj et al.SIGMOD 2022 · 37 citations
- Accelerating Approximate Aggregation Queries with Expensive PredicatesDaniel Kang, John Guibas, Peter Bailis, Tatsunori Hashimoto et al.VLDB 2021 · 34 citations
- CoVA: Exploiting Compressed-Domain Analysis to Accelerate Video AnalyticsJinwoo Hwang, Minsu Kim, Daeun Kim, Seungho Nam et al.USENIX ATC 2022 · 29 citations
- SEIDEN: Revisiting Query Processing in Video Database SystemsJaeho Bang, Gaurav Tarlok Kakkar, Pramod Chunduri, Subrata Mitra et al.VLDB 2023 · 24 citations
- TASTI: Semantic Indexes for Machine Learning-based Queries over Unstructured DataDaniel Kang, John Guibas, Peter D. Bailis, Tatsunori Hashimoto et al.SIGMOD 2022 · 23 citations
Related papers
- Top-K Deep Video Analytics: A Probabilistic ApproachZiliang Lai, Chenxia Han, Chris Liu, Pengfei Zhang et al.SIGMOD 2021 · 7 citations
- Boggart: Towards General-Purpose Acceleration of Retrospective Video AnalyticsNeil Agarwal, Ravi NetravaliNSDI 2023
- Lava: Language Driven Scalable and Versatile Traffic Video AnalyticsYanrui Yu, Tianfei Zhou, Jiaxin Sun, Lianpeng Qiao et al.ACM MM 2025
- BlazeIt: Optimizing Declarative Aggregation and Limit Queries for Neural Network-Based Video AnalyticsDaniel Kang, Peter Bailis, Matei ZahariaVLDB 2020 · 103 citations
- Tracking Anything with Decoupled Video SegmentationHo Kei Cheng, Seoung Wug Oh, Brian L. Price, Alexander G. Schwing et al.ICCV 2023 · 240 citations
