Self-Tuning Query Scheduling for Analytical Workloads
Benjamin Wagner, André Kohn, Thomas Neumann
Abstract
Most database systems delegate scheduling decisions to the operating system. While such an approach simplifies the overall database design, it also entails problems. Adaptive resource allocation becomes hard in the face of concurrent queries. Furthermore, incorporating domain knowledge to improve query scheduling is difficult.
To mitigate these problems, many modern systems employ forms of task-based parallelism. The execution of a single query is broken up into small, independent chunks of work (tasks). Now, fine-grained scheduling decisions based on these tasks are the responsibility of the database system. Despite being commonplace, little work has focused on the opportunities arising from this execution model.
In this paper, we show how task-based scheduling in database systems opens up new areas for optimization. We present a novel lock-free, self-tuning stride scheduler that optimizes query latencies for analytical workloads. By adaptively managing query priorities and task granularity, we provide high scheduling elasticity. By incorporating domain knowledge into the scheduling decisions, our system is able to cope with workloads that other systems struggle with. Even at high load, we retain near optimal latencies for short running queries. Compared to traditional database systems, our design often improves tail latencies by more than 10x.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 0034db97-5317-488b-bc2b-de1e7eaefba1Cited by top-tier papers11
- LSched: A Workload-Aware Learned Query Scheduler for Analytical Database SystemsIbrahim Sabek, Tenzin Samten Ukyab, Tim KraskaSIGMOD 2022 · 25 citations
- NyxCache: Flexible and Efficient Multi-tenant Persistent Memory CachingKan Wu, Kaiwei Tu, Yuvraj Patel, Rathijit Sen et al.FAST 2022 · 22 citations
- Accelerating database analytic query workloads using an associative processorHelena Caminal, Yannis Chronis, Tianshu Wu, Jignesh M. Patel et al.ISCA 2022 · 19 citations
- Blueprinting the Cloud: Unifying and Automatically Optimizing Cloud Data Infrastructures with BRADGeoffrey X. Yu, Ziniu Wu, Ferdinand Kossmann, Tianyu Li et al.VLDB 2024 · 11 citations
- On-Demand State Separation for Cloud Data WarehousingChristian Winter, Jana Giceva, Thomas Neumann, Alfons KemperVLDB 2022 · 9 citations
Builds on1
Related papers
- Low-Latency Transaction Scheduling via Userspace Interrupts: Why Wait or Yield When You Can Preempt?Kaisong Huang, Jiatang Zhou, Zhuoyue Zhao, Dong Xie et al.SIGMOD 2025 · 8 citations
- Improving DBMS Scheduling Decisions with Accurate Performance Prediction on Concurrent QueriesZiniu Wu, Markos Markakis, Chunwei Liu, Peter Baile Chen et al.VLDB 2025
- The Art of Latency Hiding in Modern Database EnginesKaisong Huang, Tianzheng Wang, Qingqing Zhou, Qingzhong MengVLDB 2024 · 23 citations
- Tao: Improving Resource Utilization while Guaranteeing SLO in Multi-tenant Relational Database-as-a-ServiceHaotian Liu, Runzhong Li, Ziyang Zhang, Bo TangSIGMOD 2025 · 2 citations
- An Efficient Transfer Learning Based Configuration Adviser for Database TuningXinyi Zhang, Hong Wu, Yang Li, Zhengju Tang et al.VLDB 2024 · 25 citations
