Magic mirror in my hand, which is the best in the land? An Experimental Evaluation of Index Selection Algorithms
Jan Kossmann, Stefan Halfpap, Marcel Jankrift, Rainer Schlosser
Abstract
Indexes are essential for the efficient processing of database workloads. Proposed solutions for the relevant and challenging index selection problem range from metadata-based simple heuristics, over sophisticated multi-step algorithms, to approaches that yield optimal results. The main challenges are (i) to accurately determine the effect of an index on the workload cost while considering the interaction of indexes and (ii) a large number of possible combinations resulting from workloads containing many queries and massive schemata with possibly thousands of attributes. In this work, we describe and analyze eight index selection algorithms that are based on different concepts and compare them along different dimensions, such as solution quality, runtime, multi-column support, solution granularity, and complexity. In particular, we analyze the solutions of the algorithms for the challenging analytical Join Order, TPC-H, and TPC-DS benchmarks. Afterward, we assess strengths and weaknesses, infer insights for index selection in general and each approach individually, before we give recommendations on when to use which approach.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 77fd861f-b300-4245-ab2a-bb644b1c2dd5Cited by top-tier papers32
- QueryFormer: A Tree Transformer Model for Query Plan RepresentationYue Zhao, Gao Cong, Jiachen Shi, Chunyan MiaoVLDB 2022 · 117 citations
- Are Updatable Learned Indexes Ready?Chaichon Wongkham, Baotong Lu, Chris Liu, Zhicong Zhong et al.VLDB 2022 · 66 citations
- Learned Index Benefits: Machine Learning Based Index Performance EstimationJiachen Shi, Gao Cong, Xiaoli LiVLDB 2022 · 42 citations
- DISTILL: Low-Overhead Data-Driven Techniques for Filtering and Costing Indexes for Scalable Index TuningTarique Siddiqui, Wentao Wu, Vivek R. Narasayya, Surajit ChaudhuriVLDB 2022 · 36 citations
- Budget-aware Index Tuning with Reinforcement LearningWentao Wu, Chi Wang, Tarique Siddiqui, Junxiong Wang et al.SIGMOD 2022 · 33 citations
Related papers
- MFIX: An Efficient and Reliable Index Advisor via Multi-Fidelity Bayesian OptimizationZhuo Chang, Xinyi Zhang, Yang Li, Xupeng Miao et al.ICDE 2024 · 5 citations
- Breaking It Down: An In-depth Study of Index AdvisorsWei Zhou, Chen Lin, Xuanhe Zhou, Guoliang LiVLDB 2024 · 21 citations
- ISUM: Efficiently Compressing Large and Complex Workloads for Scalable Index TuningTarique Siddiqui, Saehan Jo, Wentao Wu, Chi Wang et al.SIGMOD 2022 · 25 citations
- IDEBench: A Benchmark for Interactive Data ExplorationPhilipp Eichmann, Emanuel Zgraggen, Carsten Binnig, Tim KraskaSIGMOD 2020 · 57 citations
- Quantifying TPC-H Choke Points and Their OptimizationsMarkus Dreseler, Martin Boissier, Tilmann Rabl, Matthias UflackerVLDB 2020 · 91 citations
