Understanding and Detecting Query Performance Regression in Practical Index Tuning: [Experiments & Analysis]
Wentao Wu, Anshuman Dutt, Gaoxiang Xu, Vivek R. Narasayya, Surajit Chaudhuri
Abstract
Existing index tuners typically rely on the ''what if'' API provided by the query optimizer to estimate the execution cost of a query on top of an index configuration. Such cost estimates can be inaccurate and may therefore lead to significant query performance regression (QPR) once the recommended indexes are materialized. This becomes a serious problem for cloud database providers, such as Microsoft's Azure SQL Database, that offer index tuning as an automated service (a.k.a. ''auto-indexing''). Previous work has explored use of supervised machine learning (ML) to reduce the likelihood of QPR. However, the trained ML models have limited generalization capability when applied to new databases and workloads. We propose an alternative approach where we analyze the query plans with significant QPRs and look for structural changes due to the new index configuration that could explain the QPR. We perform such study for index tuning data across many benchmark and real-world database workloads, for multiple realistic index tuning scenarios. Our study reveals that most of the significant QPRs can be attributed to a small number of common ''regression patterns'' characterizing the structural plan changes, and we further propose a pattern-based QPR detector accordingly. Our experimental evaluation shows that the pattern-based QPR detector can significantly outperform existing ML-based QPR detectors.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers1
Ask how each one uses itBuilds on19
- An End-to-End Learning-based Cost EstimatorJi Sun, Guoliang LiVLDB 2020 · 251 citations
- Cardinality Estimation in DBMS: A Comprehensive Benchmark EvaluationYuxing Han, Ziniu Wu, Peizhi Wu, Rong Zhu et al.VLDB 2022 · 169 citations
- Are We Ready For Learned Cardinality Estimation?Xiaoying Wang, Changbo Qu, Weiyuan Wu, Jiannan Wang et al.VLDB 2021 · 156 citations
- QueryFormer: A Tree Transformer Model for Query Plan RepresentationYue Zhao, Gao Cong, Jiachen Shi, Chunyan MiaoVLDB 2022 · 117 citations
- Zero-Shot Cost Models for Out-of-the-box Learned Cost PredictionBenjamin Hilprecht, Carsten BinnigVLDB 2022 · 90 citations
Related papers
- DISTILL: Low-Overhead Data-Driven Techniques for Filtering and Costing Indexes for Scalable Index TuningTarique Siddiqui, Wentao Wu, Vivek R. Narasayya, Surajit ChaudhuriVLDB 2022 · 36 citations
- Refactoring Index Tuning Process with Benefit EstimationTao Yu, Zhaonian Zou, Weihua Sun, Yu YanVLDB 2024 · 13 citations
- Learned Index Benefits: Machine Learning Based Index Performance EstimationJiachen Shi, Gao Cong, Xiaoli LiVLDB 2022 · 42 citations
- Budget-aware Index Tuning with Reinforcement LearningWentao Wu, Chi Wang, Tarique Siddiqui, Junxiong Wang et al.SIGMOD 2022 · 33 citations
- HUNTER: An Online Cloud Database Hybrid Tuning System for Personalized RequirementsBaoqing Cai, Yu Liu, Ce Zhang, Guangyu Zhang et al.SIGMOD 2022 · 52 citations
