Seagull: An Infrastructure for Load Prediction and Optimized Resource Allocation
Olga Poppe, Tayo Amuneke, Dalitso Banda, Aritra De, Ari Green, Manon Knoertzer, Ehi Nosakhare, Karthik Rajendran, Deepak Shankargouda, Meina Wang, Alan Au, Carlo Curino
Abstract
Microsoft Azure is dedicated to guarantee high quality of service to its customers, in particular, during periods of high customer activity, while controlling cost. We employ a Data Science (DS) driven solution to predict user load and leverage these predictions to optimize resource allocation. To this end, we built the SEAGULL infrastructure that processes per-server telemetry, validates the data, trains and deploys ML models. The models are used to predict customer load per server (24h into the future), and optimize service operations. SEAGULL continually re-evaluates accuracy of predictions, fallback to previously known good models and triggers alerts as appropriate. We deployed this infrastructure in production for PostgreSQL and MySQL servers across all Azure regions, and applied it to the problem of scheduling server backups during low-load time. This minimizes interference with user-induced load and improves customer experience. We built the SEAGULL infrastructure for load prediction and optimized resource allocation on the cloud. While the infrastructure is applicable to a wide range of use cases, we illustrated it by the backup scheduling scenario.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext a9160d38-b825-43be-9007-876a83bc7fcbCited by top-tier papers10
- Take it to the limit: peak prediction-driven resource overcommitment in datacentersNoman Bashir, Nan Deng, Krzysztof Rzadca, David Irwin et al.EuroSys 2021 · 60 citations
- Moneyball: Proactive Auto-Scaling in Microsoft Azure SQL Database ServerlessOlga Poppe, Qun Guo, Willis Lang, Pankaj Arora et al.VLDB 2022 · 40 citations
- The Holon Approach for Simultaneously Tuning Multiple Components in a Self-Driving Database Management System with Machine Learning via Synthesized Proto-ActionsWilliam Zhang, Wan Shen Lim, Matthew Butrovich, Andrew PavloVLDB 2024 · 13 citations
- Tiresias: Enabling Predictive Autonomous Storage and IndexingMichael Abebe, Horatiu Lazu, Khuzaima DaudjeeVLDB 2022 · 10 citations
- Tenant Placement in Over-subscribed Database-as-a-Service ClustersArnd Christian König, Yi Shan, Tobias Ziegler, Aarati Kakaraparthy et al.VLDB 2022 · 8 citations
Related papers
- Flexible Resource Allocation for Relational Database-as-a-ServicePankaj Arora, Surajit Chaudhuri, Sudipto Das, Junfeng Dong et al.VLDB 2023 · 10 citations
- End-to-end Optimization of Machine Learning Prediction QueriesKwanghyun Park, Karla Saur, Dalitso Banda, Rathijit Sen et al.SIGMOD 2022 · 50 citations
- A Resource-centric Analysis and Optimization of NoSQL Workloads using Distressed Resource Volume MetricGunika Verma, Aashutosh A V, Pooja Srinivas, Yogesh Simmhan et al.VLDB 2026
- Runtime Variation in Big Data AnalyticsYiwen Zhu, Rathijit Sen, Robert Horton, John Mark AgostaSIGMOD 2023 · 5 citations
- Prediction-Based Power Oversubscription in Cloud PlatformsAlok Gautam Kumbhare, Reza Azimi, Ioannis Manousakis, Anand Bonde et al.USENIX ATC 2021 · 90 citations
