MerKury: Adaptive Resource Allocation to Enhance the Kubernetes Performance for Large-Scale Clusters
Jiayin Luo, Xinkui Zhao, Yuxin Ma, Shengye Pang, Jianwei Yin
摘要
As a prevalent paradigm of modern web applications, cloud computing has experienced a surge in adoption. The deployment of vast and various workloads encapsulated within containers has become ubiquitous across cloud platforms, imposing substantial demands on the supporting infrastructure. However, Kubernetes (k8s), the de-facto standard for container orchestration, struggles with low scheduling throughput and high latency in large-scale clusters. The primary challenges are identified as excessive loads of read requests and resource contention among co-located components. In response to these challenges, in this paper, we present MerKury, a lightweight framework to enhance the Kubernetes performance for large-scale clusters. It employs a dual strategy: first, it preprocesses specific requests to alleviate unnecessary load, and second, it introduces an adaptive resource allocation algorithm to mitigate resource contention. Evaluations under different scenarios of varying cluster scale have demonstrated that MerKury notably augments cluster scheduling throughput up to 16.4× and reduces request latency by up to 39.3%, outperforming vanilla Kubernetes and baseline resource allocation methods. CCS CONCEPTS • Computer systems organization → Cloud computing.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper12
- StepConf: SLO-Aware Dynamic Resource Configuration for Serverless Function WorkflowsZhaojie Wen, Yishuo Wang, Fangming LiuINFOCOM 2022 · 被引用 72 次
- Parcae: Proactive, Liveput-Optimized DNN Training on Preemptible InstancesJiangfei Duan, Ziang Song, Xupeng Miao, Xiaoli Xi 等NSDI 2024 · 被引用 54 次
- Automatic Reliability Testing For Cluster Management ControllersXudong Sun, Wenqing Luo, Jiawei Tyler Gu, Aishwarya Ganesan 等OSDI 2022 · 被引用 44 次
- Cilantro: Performance-Aware Resource Allocation for General Objectives via Online FeedbackRomil Bhardwaj, Kirthevasan Kandasamy, Asim Biswal, Wenshuo Guo 等OSDI 2023 · 被引用 41 次
- Understanding and Optimizing Workloads for Unified Resource Management in Large Cloud PlatformsChengzhi Lu, Huanle Xu, Kejiang Ye, Guoyao Xu 等EuroSys 2023 · 被引用 34 次
相关 Paper
- Tailored Learning-Based Scheduling for Kubernetes-Oriented Edge-Cloud SystemYiwen Han, Shihao Shen, Xiaofei Wang, Shiqiang Wang 等INFOCOM 2021 · 被引用 93 次
- Oakestra: A Lightweight Hierarchical Orchestration Framework for Edge ComputingGiovanni Bartolomeo, Mehdi Yosofie, Simon Bäurle, Oliver Haluszczynski 等USENIX ATC 2023 · 被引用 51 次
- KubeShare: A Framework to Manage GPUs as First-Class and Shared Resources in Container CloudTing-An Yeh, Hung-Hsin Chen, Jerry ChouHPDC 2020 · 被引用 56 次
- Pyramid: A Secure, Resource-Efficient, and Pluggable Kubernetes for Multi-TenancyXiang Li, Weijie Liu, Fabing Li, Hongliang Tian 等EuroSys 2026
- vBOIDs: Taming Chaos via Coarse-Grained Scheduling Abstraction for ContainersKaesi Manakkal, Nathan Daughety, Yu Sun, Marcus Pendleton 等OSDI 2026
