Optimizing Machine Learning Inference Queries with Correlative Proxy Models
Zhihui Yang, Zuozhi Wang, Yicong Huang, Yao Lu, Chen Li, X. Sean Wang
摘要
We consider accelerating machine learning (ML) inference queries on unstructured datasets. Expensive operators such as feature extractors and classifiers are deployed as user-defined functions (UDFs), which are not penetrable with classic query optimization techniques such as predicate push-down. Recent optimization schemes (e.g., Probabilistic Predicates or PP) assume independence among the query predicates, build a proxy model for each predicate offline, and rewrite a new query by injecting these cheap proxy models in the front of the expensive ML UDFs. In such a manner, unlikely inputs that do not satisfy query predicates are filtered early to bypass the ML UDFs. We show that enforcing the independence assumption in this context may result in sub-optimal plans. In this paper, we propose CORE, a query optimizer that better exploits the predicate correlations and accelerates ML inference queries. Our solution builds the proxy models online for a new query and leverages a branch-and-bound search process to reduce the building costs. Results on three real-world text, image and video datasets show that CORE improves the query throughput by up to 63% compared to PP and up to 80% compared to running the queries as it is.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper12
- Optimizing Video Analytics with Declarative Model RelationshipsFrancisco Romero, Johann Hauswald, Aditi Partap, Daniel Kang 等VLDB 2023 · 被引用 37 次
- Serving and Optimizing Machine Learning Workflows on Heterogeneous InfrastructuresYongji Wu, Matthew Lentz, Danyang Zhuo, Yao LuVLDB 2023 · 被引用 31 次
- Extract-Transform-Load for Video StreamsFerdinand Kossmann, Ziniu Wu, Eugenie Lai, Nesime Tatbul 等VLDB 2023 · 被引用 21 次
- SmartLite: A DBMS-based Serving System for DNN Inference in Resource-constrained EnvironmentsQiuru Lin, Sai Wu, Junbo Zhao, Jian Dai 等VLDB 2024 · 被引用 17 次
- On Efficient Approximate Queries over Machine Learning ModelsDujian Ding, Sihem Amer-Yahia, Laks V. S. LakshmananVLDB 2023 · 被引用 11 次
它引用的顶会 Paper3
- BlazeIt: Optimizing Declarative Aggregation and Limit Queries for Neural Network-Based Video AnalyticsDaniel Kang, Peter Bailis, Matei ZahariaVLDB 2020 · 被引用 103 次
- Approximate Selection with Guarantees using ProxiesDaniel Kang, Edward Gan, Peter Bailis, Tatsunori Hashimoto 等VLDB 2020 · 被引用 46 次
- Model Slicing for Supporting Complex Analytics with Elastic Inference Cost and Resource ConstraintsShaofeng Cai, Gang Chen, Beng Chin Ooi, Jinyang GaoVLDB 2020 · 被引用 21 次
相关 Paper
- PLAQUE: Automated Predicate Learning at Query TimeYiming Lin, Sharad MehrotraSIGMOD 2024 · 被引用 2 次
- Predicate Pushdown for Data Science PipelinesCong Yan, Yin Lin, Yeye HeSIGMOD 2023 · 被引用 15 次
- A Method for Optimizing Opaque Filter QueriesWenjia He, Michael R. Anderson, Maxwell Strome, Michael J. CafarellaSIGMOD 2020 · 被引用 15 次
- Mitigating the Impedance Mismatch between Prediction Query Execution and Database EngineChenyang Zhang, Junxiong Peng, Chen Xu, Quanqing Xu 等SIGMOD 2025 · 被引用 7 次
- Aero: Adaptive Query Processing of ML QueriesGaurav Tarlok Kakkar, Jiashen Cao, Aubhro Sengupta, Joy Arulraj 等SIGMOD 2025 · 被引用 2 次
