Autothrottle: A Practical Bi-Level Approach to Resource Management for SLO-Targeted Microservices
Zibo Wang, Pinghe Li, Chieh-Jan Mike Liang, Feng Wu, Francis Y. Yan
摘要
Achieving resource efficiency while preserving end-user experience is non-trivial for cloud application operators. As cloud applications progressively adopt microservices, resource managers are faced with two distinct levels of system behavior: end-to-end application latency and per-service resource usage. Translating between the two levels, however, is challenging because user requests traverse heterogeneous services that collectively (but unevenly) contribute to the endto-end latency. We present Autothrottle, a bi-level resource management framework for microservices with latency SLOs (service-level objectives). It architecturally decouples application SLO feedback from service resource control, and bridges them through the notion of performance targets. Specifically, an application-wide learning-based controller is employed to periodically set performance targets-expressed as CPU throttle ratios-for per-service heuristic controllers to attain. We evaluate Autothrottle on three microservice applications, with workload traces from production scenarios. Results show superior CPU savings, up to 26.21% over the best-performing baseline and up to 93.84% over all baselines.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper11
- Rajomon: Decentralized and Coordinated Overload Control for Latency-Sensitive MicroservicesJiali Xing, Akis Giannoukos, Paul Loh, Shuyue Wang 等NSDI 2025 · 被引用 12 次
- TopFull: An Adaptive Top-Down Overload Control for SLO-Oriented MicroservicesJinwoo Park, Jaehyeong Park, Youngmok Jung, Hwijoon Lim 等SIGCOMM 2024 · 被引用 10 次
- How Soon is Now? Preloading Images for Virtual Disks with ThinkAheadXinqi Chen, Yu Zhang, Erci Xu, Changhong Wang 等FAST 2026 · 被引用 1 次
- Uber's Failover Architecture: Reconciling Reliability and Efficiency in Hyperscale Microservice InfrastructureMayank Bansal, Milind Chabbi, Kenneth Bogh, Srikanth Prodduturi 等NSDI 2026 · 被引用 1 次
- DistRS: Disaggregated Reward Service for RLVR with Batch-Level ConstraintRuidong Zhu, Mingcong Han, Yinmin Zhong, Wencong Xiao 等NSDI 2026 · 被引用 1 次
它引用的顶会 Paper8
- Learning in situ: a randomized experiment in video streamingFrancis Y. Yan, Hudson Ayers, Chenzhi Zhu, Sadjad Fouladi 等NSDI 2020 · 被引用 360 次
- FIRM: An Intelligent Fine-grained Resource Management Framework for SLO-Oriented MicroservicesHaoran Qiu, Subho S. Banerjee, Saurabh Jha, Zbigniew T. Kalbarczyk 等OSDI 2020 · 被引用 350 次
- Autopilot: workload autoscaling at GoogleKrzysztof Rzadca, Pawel Findeisen, Jacek Swiderski, Przemyslaw Zych 等EuroSys 2020 · 被引用 299 次
- Sinan: ML-based and QoS-aware resource management for cloud microservicesYanqi Zhang, Weizhe Hua, Zhuangzhuang Zhou, G. Edward Suh 等ASPLOS 2021 · 被引用 226 次
- Faster and Cheaper Serverless Computing on Harvested ResourcesYanqi Zhang, Iñigo Goiri, Gohar Irfan Chaudhry, Rodrigo Fonseca 等SOSP 2021 · 被引用 131 次
相关 Paper
- Erlang: Application-Aware Autoscaling for Cloud MicroservicesVighnesh Sachidananda, Anirudh SivaramanEuroSys 2024 · 被引用 7 次
- ANT-man: towards agile power management in the microservice eraXiaofeng Hou, Chao Li, Jiacheng Liu, Lu Zhang 等SC 2020 · 被引用 35 次
- Towards Performance Robustness for MicroservicesDivyanshu Saxena, Gaurav Vipat, Jiaxin Lin, Jingbo Wang 等NSDI 2026
- Grad: Intelligent Microservice Scaling by Harnessing Resource FungibilityLiao Chen, Chenyu Lin, Shutian Luo, Huanle Xu 等HPCA 2025 · 被引用 5 次
- Practical Efficient Microservice Autoscaling with QoS AssuranceMd Rajib Hossen, Mohammad A. Islam, Kishwar AhmedHPDC 2022 · 被引用 45 次
