A Study of Database Performance Sensitivity to Experiment Settings
Yang Wang, Miao Yu, Yujie Hui, Fang Zhou, Yuyang Huang, Rui Zhu, Xueyuan Ren, Tianxi Li, Xiaoyi Lu
Abstract
To allow performance comparison across different systems, our community has developed multiple benchmarks, such as TPC-C and YCSB, which are widely used. However, despite such effort, interpreting and comparing performance numbers is still a challenging task, because one can tune benchmark parameters, system features, and hardware settings, which can lead to very different system behaviors. Such tuning creates a long-standing question of whether the conclusion of a work can hold under different settings. This work tries to shed light on this question by reproducing 11 works evaluated under TPC-C and YCSB, measuring their performance under a wider range of settings, and investigating the reasons for the change of performance numbers. By doing so, this paper tries to motivate the discussion about whether and how we should address this problem. While this paper does not give a complete solution---this is beyond the scope of a single paper, it proposes concrete suggestions we can take to improve the state of the art.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 895a12cb-fb09-4309-8697-3faf6cd51aa4Cited by top-tier papers4
- NOC-NOC: Towards Performance-optimal Distributed TransactionsSi Liu, Luca Multazzu, Hengfeng Wei, David A. BasinSIGMOD 2024 · 6 citations
- Knock Out 2PC with Practicality Intact: a High-performance and General Distributed Transaction ProtocolZiliang Lai, Hua Fan, Wenchao Zhou, Zhanfeng Ma et al.ICDE 2023 · 6 citations
- A Hybrid Approach to Integrating Deterministic and Non-deterministic Concurrency Control in Database SystemsYinhao Hong, Hongyao Zhao, Wei Lu, Xiaoyong Du et al.VLDB 2025 · 3 citations
- On the Feasibility and Benefits of Extensive EvaluationYujie Hui, Miao Yu, Hao Qi, Yifan Gan et al.SIGMOD 2025 · 1 citation
Builds on6
- MLPerf Inference BenchmarkVijay Janapa Reddi, Christine Cheng, David Kanter, Peter Mattson et al.ISCA 2020 · 517 citations
- A large scale analysis of hundreds of in-memory cache clusters at TwitterJuncheng Yang, Yao Yue, K. V. RashmiOSDI 2020 · 245 citations
- Is Big Data Performance Reproducible in Modern Cloud Networks?Alexandru Uta, Alexandru Custura, Dmitry Duplyakin, Ivo Jimenez et al.NSDI 2020 · 74 citations
- Low-Latency Communication for Fast DBMS Using RDMA and Shared MemoryPhilipp Fent, Alexander van Renen, Andreas Kipf, Viktor Leis et al.ICDE 2020 · 44 citations
- Caracal: Contention Management with Deterministic Concurrency ControlDai Qin, Angela Demke Brown, Ashvin GoelSOSP 2021 · 37 citations
Related papers
- Redbench: Workload Synthesis From Cloud TracesJohannes Wehrstein, Roman Heinrich, Mihail Stoian, Skander Krid et al.VLDB 2026 · 7 citations
- Why TPC Is Not Enough: An Analysis of the Amazon Redshift FleetAlexander van Renen, Dominik Horn, Pascal Pfeil, Kapil Vaidya et al.VLDB 2024 · 65 citations
- Facilitating Database Tuning with Hyper-Parameter Optimization: A Comprehensive Experimental EvaluationXinyi Zhang, Zhuo Chang, Yang Li, Hong Wu et al.VLDB 2022 · 88 citations
- DSB: A Decision Support Benchmark for Workload-Driven and Traditional Database SystemsBailu Ding, Surajit Chaudhuri, Johannes Gehrke, Vivek R. NarasayyaVLDB 2021 · 62 citations
- Task bench: a parameterized benchmark for evaluating parallel runtime performanceElliott Slaughter, Wei Wu, Yuankun Fu, Legend Brandenburg et al.SC 2020 · 51 citations
