Karma: Resource Allocation for Dynamic Demands
Midhul Vuppalapati, Giannis Fikioris, Rachit Agarwal, Asaf Cidon, Anurag Khandelwal, Éva Tardos
摘要
We consider the problem of fair resource allocation in a system where user demands are dynamic, that is, where user demands vary over time. Our key observation is that the classical max-min fairness algorithm for resource allocation provides many desirable properties (e.g., Pareto efficiency, strategy-proofness, and fairness), but only under the strong assumption of user demands being static over time. For the realistic case of dynamic user demands, the max-min fairness algorithm loses one or more of these properties. We present Karma, a new resource allocation mechanism for dynamic user demands. The key technical contribution in Karma is a credit-based resource allocation algorithm: in each quantum, users donate their unused resources and are assigned credits when other users borrow these resources; Karma carefully orchestrates the exchange of credits across users (based on their instantaneous demands, donated resources and borrowed resources), and performs prioritized resource allocation based on users' credits. We theoretically establish Karma guarantees related to Pareto efficiency, strategy-proofness, and fairness for dynamic user demands. Empirical evaluations over production workloads show that these properties translate well into practice: Karma is able to reduce disparity in performance across users to a bare minimum while maintaining Pareto-optimal system-wide performance.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- Fair-CO2: Fair Attribution for Cloud Carbon EmissionsLeo Han, Jash Kakadia, Benjamin C. Lee, Udit GuptaISCA 2025 · 被引用 10 次
- TopFull: An Adaptive Top-Down Overload Control for SLO-Oriented MicroservicesJinwoo Park, Jaehyeong Park, Youngmok Jung, Hwijoon Lim 等SIGCOMM 2024 · 被引用 10 次
- MerKury: Adaptive Resource Allocation to Enhance the Kubernetes Performance for Large-Scale ClustersJiayin Luo, Xinkui Zhao, Yuxin Ma, Shengye Pang 等WWW 2025
- Mitigating Application Resource Overload with Targeted Task CancellationYigong Hu, Zeyin Zhang, Yicheng Liu, Yile Gu 等SOSP 2025
它引用的顶会 Paper13
- Serverless in the Wild: Characterizing and Optimizing the Serverless Workload at a Large Cloud ProviderMohammad Shahrad, Rodrigo Fonseca, Iñigo Goiri, Gohar Irfan Chaudhry 等USENIX ATC 2020 · 被引用 946 次
- Heterogeneity-Aware Cluster Scheduling Policies for Deep Learning WorkloadsDeepak Narayanan, Keshav Santhanam, Fiodar Kazhamiaka, Amar Phanishayee 等OSDI 2020 · 被引用 286 次
- A large scale analysis of hundreds of in-memory cache clusters at TwitterJuncheng Yang, Yao Yue, K. V. RashmiOSDI 2020 · 被引用 245 次
- The CacheLib Caching Engine: Design and Experiences at ScaleBenjamin Berg, Daniel S. Berger, Sara McAllister, Isaac Grosof 等OSDI 2020 · 被引用 145 次
- Building An Elastic Query Engine on Disaggregated StorageMidhul Vuppalapati, Justin Miron, Rachit Agarwal, Dan Truong 等NSDI 2020 · 被引用 142 次
相关 Paper
- Quota Marketplace: Dynamic Pricing for Efficient Allocation of ML Training ResourcesBalasubramanian Sivan, Renato Paes Leme, Mihai Tiuca, Ian McFarlane 等OSDI 2026 · 被引用 1 次
- Time Fairness in Online Knapsack ProblemsAdam Lechowicz, Rik Sengupta, Bo Sun, Shahin Kamali 等ICLR 2024 · 被引用 8 次
- Fair and Efficient Allocations with Limited DemandsSushirdeep Narayana, Ian A. KashAAAI 2021 · 被引用 3 次
- Fair Scheduling for Time-dependent ResourcesBo Li, Minming Li, Ruilong ZhangNeurIPS 2021 · 被引用 23 次
- Experiential Fairness: Bridging the Gap Between User Experience and Resource-Centric Fairness in Online LLM ServicesJiahua Huang, Wentai Wu, Yongheng Liu, Guozhi Liu 等AAAI 2026
