Slowpoke: End-to-end Throughput Optimization Modeling for Microservice Applications
Yizheng Xie, Di Jin, Oguzhan Çölkesen, Vasiliki Kalavri, John Liagouris, Nikos Vasilakis
摘要
SLOWPOKE is a new system to accurately quantify the effects of hypothetical optimizations on end-to-end throughput for microservice applications, without relying on tracing or a priori knowledge of the call graph. Microservice operators can use SLOWPOKE to ask what-if performance analysis questions of the form "What throughput could my retail application sustain if I optimized the shopping cart service from 10K req/s to 20K req/s?". Given a target service and its hypothetical optimization, SLOWPOKE employs a performance model that determines how to selectively slow down non-target services to preserve the relative effect of the optimization. It then performs profiling experiments to predict the end-to-end throughput, as if the optimization had been implemented. Applied to four real-world microservice applications, SLOWPOKE accurately quantifies optimization effects with a root mean squared error of only 2.07%. It is also effective in more complex scenarios, e.g., predicting throughput after scaling optimizations or when bottlenecks arise from mutex contention. Evaluated in large-scale deployments of 45 nodes and 108 synthetic benchmarks, SLOWPOKE further demonstrates its scalability and coverage of a wide range of microservice characteristics.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper14
- FIRM: An Intelligent Fine-grained Resource Management Framework for SLO-Oriented MicroservicesHaoran Qiu, Subho S. Banerjee, Saurabh Jha, Zbigniew T. Kalbarczyk 等OSDI 2020 · 被引用 350 次
- Autopilot: workload autoscaling at GoogleKrzysztof Rzadca, Pawel Findeisen, Jacek Swiderski, Przemyslaw Zych 等EuroSys 2020 · 被引用 299 次
- Sinan: ML-based and QoS-aware resource management for cloud microservicesYanqi Zhang, Weizhe Hua, Zhuangzhuang Zhou, G. Edward Suh 等ASPLOS 2021 · 被引用 226 次
- ORION and the Three Rights: Sizing, Bundling, and Prewarming for Serverless DAGsAshraf Mahgoub, Edgardo Barsallo Yi, Karthick Shankar, Sameh Elnikety 等OSDI 2022 · 被引用 111 次
- Accelerometer: Understanding Acceleration Opportunities for Data Center Overheads at HyperscaleAkshitha Sriraman, Abhishek DhanotiaASPLOS 2020 · 被引用 78 次
相关 Paper
- FlowScope: Non-Intrusive Distributed Tracing with Method-Level Delay Estimation for Microservices TroubleshootingYantao Geng, Han Zhang, Zhiheng Wu, Yahui Li 等ICSE 2026
- FastPERT: Towards Fast Microservice Application Latency Prediction via Structural Inductive Bias over PERT NetworksDa Sun Handason Tam, Huanle Xu, Yang Liu, Siyue Xie 等AAAI 2025 · 被引用 5 次
- MicroRank: End-to-End Latency Issue Localization with Extended Spectrum Analysis in Microservice EnvironmentsGuangba Yu, Pengfei Chen, Hongyang Chen, Zijie Guan 等WWW 2021 · 被引用 152 次
- MuCache: A General Framework for Caching in Microservice GraphsHaoran Zhang, Konstantinos Kallas, Spyros Pavlatos, Rajeev Alur 等NSDI 2024 · 被引用 14 次
- Critical Path Guided Decision Making with CALLIGATORMeghna Pancholi, Lee Baugh, Olaf Schnapauff, David E. Culler 等SIGCOMM 2026
