TVM: A Tile-based Video Management Framework
Tianxiong Zhong, Zhiwei Zhang, Guo Lu, Ye Yuan, Yu-Ping Wang, Guoren Wang
Abstract
With the exponential growth of video data, there is a pressing need for ecient video analysis technology. Modern query frameworks aim to accelerate queries by reducing the frequency of calls to expensive deep neural networks, which often overlook the overhead associated with video decoding and retrieval. Furthermore, video storage frameworks optimize video retrieval through video partition or caching, often relying on prior information about the query workload. To further accelerate queries, this study introduces a novel tile-based video management framework, called TVM, which leverages the semantic information embedded in videos, without being dependent on specic query workloads. By constructing a tile-based semantic index for newly ingested videos, TVM eectively reduces the size of decoded and processed video data. To achieve this, TVM introduces an optimal index construction algorithm that utilizes cost function and pseudo-labels. Additionally, the framework proposes a query-driven tile parallel decoding algorithm and resource caching algorithms, which further expedite the retrieval of video frames. Experimental results demonstrate that TVM can signicantly enhance the throughput of various query tasks, achieving a notable speedup of more than 5.6⇥.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 966a0b2c-de2c-4fc7-a1fe-40fa9fc4e5b7Cited by top-tier papers4
- nsDB: Architecting the Next Generation Database by Integrating Neural and Symbolic Systems (Vision)Ye Yuan, Bo Tang, Tianfei Zhou, Zhiwei Zhang et al.VLDB 2024 · 3 citations
- Déjà Vu: Efficient Video-Language Query Engine with Learning-based Inter-Frame Computation ReuseJinwoo Hwang, Daeun Kim, Sangyeop Lee, Yoonsung Kim et al.VLDB 2025 · 2 citations
- LOVO: Efficient Complex Object Query in Large-Scale Video DatasetsYuxin Liu, Yuezhang Peng, Hefeng Zhou, Hongze Liu et al.ICDE 2025 · 2 citations
- How2Compress: Scalable and Efficient Edge Video Analytics via Adaptive Granular Video CompressionYuheng Wu, Thanh-Tung Nguyen, Lucas Liebe, Quang Tau et al.ACM MM 2025 · 1 citation
Builds on9
- FCOS: Fully Convolutional One-Stage Object DetectionZhi Tian, Chunhua Shen, Hao Chen, Tong HeICCV 2019 · 6,042 citations
- BlazeIt: Optimizing Declarative Aggregation and Limit Queries for Neural Network-Based Video AnalyticsDaniel Kang, Peter Bailis, Matei ZahariaVLDB 2020 · 103 citations
- MIRIS: Fast Object Track Queries in VideoFavyen Bastani, Songtao He, Arjun Balasingam, Karthik Gopalakrishnan et al.SIGMOD 2020 · 68 citations
- Jointly Optimizing Preprocessing and Inference for DNN-based Visual AnalyticsDaniel Kang, Ankit Mathur, Teja Veeramacheneni, Peter Bailis et al.VLDB 2021 · 50 citations
- Approximate Selection with Guarantees using ProxiesDaniel Kang, Edward Gan, Peter Bailis, Tatsunori Hashimoto et al.VLDB 2020 · 46 citations
Related papers
- TASM: A Tile-Based Storage Manager for Video AnalyticsMaureen Daum, Brandon Haynes, Dong He, Amrita Mazumdar et al.ICDE 2021 · 19 citations
- QaVA: Query-Aware Video Analysis Framework Based on Data Access PatternTianxiong Zhong, Zhiwei Zhang, Yihang Fu, Guo Lu et al.ICDE 2025 · 1 citation
- EVA: A Symbolic Approach to Accelerating Exploratory Video Analytics with Materialized ViewsZhuangdi Xu, Gaurav Tarlok Kakkar, Joy Arulraj, Umakishore RamachandranSIGMOD 2022 · 26 citations
- Optimizing Video Queries with Declarative CluesDaren Chao, Yueting Chen, Nick Koudas, Xiaohui YuVLDB 2024 · 5 citations
- Track Merging for Effective Video Query ProcessingDaren Chao, Yueting Chen, Nick Koudas, Xiaohui YuICDE 2023 · 5 citations
