Moneyball: Proactive Auto-Scaling in Microsoft Azure SQL Database Serverless
Olga Poppe, Qun Guo, Willis Lang, Pankaj Arora, Morgan Oslake, Shize Xu, Ajay Kalhan
Abstract
Microsoft Azure SQL Database is among the leading relational database service providers in the cloud. Serverless compute automatically scales resources based on workload demand. When a database becomes idle its resources are reclaimed. When activity returns, resources are resumed. Customers pay only for resources they used. However, scaling is currently merely reactive, not proactive, according to customers' workloads. Therefore, resources may not be immediately available when a customer comes back online after a prolonged idle period. In this work, we focus on reducing this delay in resource availability by predicting the pause/resume patterns and proactively resuming resources for each database. Furthermore, we avoid taking away resources for short idle periods to relieve the back-end from ineffective pause/resume workflows. Results of this study are currently being used worldwide to find the middle ground between quality of service and cost of operation.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 6c9a1cb3-66e1-4cd2-b36d-ce4109db8fbbCited by top-tier papers7
- ElasticNotebook: Enabling Live Migration for Computational NotebooksZhaoheng Li, Pranav Gor, Rahul Prabhu, Hui Yu et al.VLDB 2024 · 12 citations
- Robust Auto-Scaling with Probabilistic Workload Forecasting for Cloud DatabasesHaitian Hang, Xiu Tang, Jianling Sun, Lingfeng Bao et al.ICDE 2024 · 11 citations
- CloudyBench: A Testbed for A Comprehensive Evaluation of Cloud-Native DatabasesChao Zhang, Guoliang Li, Leyao Liu, Tao Lv et al.ICDE 2025 · 4 citations
- Intelligent Pooling: Proactive Resource Provisioning in Large-scale Cloud ServiceDeepak Ravikumar, Alex Yeo, Yiwen Zhu, Aditya Lakra et al.VLDB 2024 · 3 citations
- Making LSM-Tree-based Key-Value Store Practical and Efficient for Multi-Tenant Serverless Cloud DatabasesYingjia Wang, Caixin Gong, Guoyun Zhu, Sheng Wang et al.SIGMOD 2026 · 1 citation
Builds on1
Related papers
- Flexible Resource Allocation for Relational Database-as-a-ServicePankaj Arora, Surajit Chaudhuri, Sudipto Das, Junfeng Dong et al.VLDB 2023 · 10 citations
- Tenant Placement in Over-subscribed Database-as-a-Service ClustersArnd Christian König, Yi Shan, Tobias Ziegler, Aarati Kakaraparthy et al.VLDB 2022 · 8 citations
- Serverless in the Wild: Characterizing and Optimizing the Serverless Workload at a Large Cloud ProviderMohammad Shahrad, Rodrigo Fonseca, Iñigo Goiri, Gohar Irfan Chaudhry et al.USENIX ATC 2020 · 946 citations
- Efficient Deep Learning Pipelines for Accurate Cost Estimations Over Large Scale Query WorkloadJohan Kok Zhi Kang, Gaurav, Sien Yi Tan, Feng Cheng et al.SIGMOD 2021 · 31 citations
- Understanding and Detecting Query Performance Regression in Practical Index Tuning: [Experiments & Analysis]Wentao Wu, Anshuman Dutt, Gaoxiang Xu, Vivek R. Narasayya et al.SIGMOD 2026
