Seagull: An Infrastructure for Load Prediction and Optimized Resource Allocation
Olga Poppe, Tayo Amuneke, Dalitso Banda, Aritra De, Ari Green, Manon Knoertzer, Ehi Nosakhare, Karthik Rajendran, Deepak Shankargouda, Meina Wang, Alan Au, Carlo Curino
摘要
Microsoft Azure is dedicated to guarantee high quality of service to its customers, in particular, during periods of high customer activity, while controlling cost. We employ a Data Science (DS) driven solution to predict user load and leverage these predictions to optimize resource allocation. To this end, we built the SEAGULL infrastructure that processes per-server telemetry, validates the data, trains and deploys ML models. The models are used to predict customer load per server (24h into the future), and optimize service operations. SEAGULL continually re-evaluates accuracy of predictions, fallback to previously known good models and triggers alerts as appropriate. We deployed this infrastructure in production for PostgreSQL and MySQL servers across all Azure regions, and applied it to the problem of scheduling server backups during low-load time. This minimizes interference with user-induced load and improves customer experience. We built the SEAGULL infrastructure for load prediction and optimized resource allocation on the cloud. While the infrastructure is applicable to a wide range of use cases, we illustrated it by the backup scheduling scenario.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper10
- Take it to the limit: peak prediction-driven resource overcommitment in datacentersNoman Bashir, Nan Deng, Krzysztof Rzadca, David Irwin 等EuroSys 2021 · 被引用 60 次
- Moneyball: Proactive Auto-Scaling in Microsoft Azure SQL Database ServerlessOlga Poppe, Qun Guo, Willis Lang, Pankaj Arora 等VLDB 2022 · 被引用 40 次
- The Holon Approach for Simultaneously Tuning Multiple Components in a Self-Driving Database Management System with Machine Learning via Synthesized Proto-ActionsWilliam Zhang, Wan Shen Lim, Matthew Butrovich, Andrew PavloVLDB 2024 · 被引用 13 次
- Tiresias: Enabling Predictive Autonomous Storage and IndexingMichael Abebe, Horatiu Lazu, Khuzaima DaudjeeVLDB 2022 · 被引用 10 次
- Tenant Placement in Over-subscribed Database-as-a-Service ClustersArnd Christian König, Yi Shan, Tobias Ziegler, Aarati Kakaraparthy 等VLDB 2022 · 被引用 8 次
相关 Paper
- Flexible Resource Allocation for Relational Database-as-a-ServicePankaj Arora, Surajit Chaudhuri, Sudipto Das, Junfeng Dong 等VLDB 2023 · 被引用 10 次
- End-to-end Optimization of Machine Learning Prediction QueriesKwanghyun Park, Karla Saur, Dalitso Banda, Rathijit Sen 等SIGMOD 2022 · 被引用 50 次
- A Resource-centric Analysis and Optimization of NoSQL Workloads using Distressed Resource Volume MetricGunika Verma, Aashutosh A V, Pooja Srinivas, Yogesh Simmhan 等VLDB 2026
- Runtime Variation in Big Data AnalyticsYiwen Zhu, Rathijit Sen, Robert Horton, John Mark AgostaSIGMOD 2023 · 被引用 5 次
- Prediction-Based Power Oversubscription in Cloud PlatformsAlok Gautam Kumbhare, Reza Azimi, Ioannis Manousakis, Anand Bonde 等USENIX ATC 2021 · 被引用 90 次
