ORION and the Three Rights: Sizing, Bundling, and Prewarming for Serverless DAGs
Ashraf Mahgoub, Edgardo Barsallo Yi, Karthick Shankar, Sameh Elnikety, Somali Chaterji, Saurabh Bagchi
摘要
Serverless applications represented as DAGs have been growing in popularity. For many of these applications, it would be useful to estimate the end-to-end (E2E) latency and to allocate resources to individual functions so as to meet probabilistic guarantees for the E2E latency. This goal has not been met till now due to three fundamental challenges. The first is the high variability and correlation in the execution time of individual functions, the second is the skew in execution times of the parallel invocations, and the third is the incidence of cold starts. In this paper, we introduce ORION to achieve this goal. We first analyze traces from a production FaaS infrastructure to identify three characteristics of serverless DAGs. We use these to motivate and design three features. The first is a performance model that accounts for runtime variabilities and dependencies among functions in a DAG. The second is a method for co-locating multiple parallel invocations within a single VM thus mitigating contentbased skew among these invocations. The third is a method for pre-warming VMs for subsequent functions in a DAG with the right look-ahead time. We integrate these three innovations and evaluate ORION on AWS Lambda with three serverless DAG applications. Our evaluation shows that compared to three competing approaches, ORION achieves up to 90% lower
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper22
- SkyPilot: An Intercloud Broker for Sky ComputingZongheng Yang, Zhanghao Wu, Michael Luo, Wei-Lin Chiang 等NSDI 2023 · 被引用 135 次
- ServerlessLLM: Low-Latency Serverless Inference for Large Language ModelsYao Fu, Leyang Xue, Yeqi Huang, Andrei-Octavian Brabete 等OSDI 2024 · 被引用 125 次
- Parrot: Efficient Serving of LLM-based Applications with Semantic VariableChaofan Lin, Zhenhua Han, Chengruidong Zhang, Yuqing Yang 等OSDI 2024 · 被引用 112 次
- Toppings: CPU-Assisted, Rank-Aware Adapter Serving for LLM InferenceSuyi Li, Hanfeng Lu, Tianyuan Wu, Minchen Yu 等USENIX ATC 2025 · 被引用 22 次
- Automated Verification of Idempotence for Stateful Serverless ApplicationsHaoran Ding, Zhaoguo Wang, Zhuohao Shen, Rong Chen 等OSDI 2023 · 被引用 16 次
它引用的顶会 Paper6
- Serverless in the Wild: Characterizing and Optimizing the Serverless Workload at a Large Cloud ProviderMohammad Shahrad, Rodrigo Fonseca, Iñigo Goiri, Gohar Irfan Chaudhry 等USENIX ATC 2020 · 被引用 946 次
- FaasCache: keeping serverless computing alive with greedy-dual cachingAlexander Fuerst, Prateek SharmaASPLOS 2021 · 被引用 223 次
- SONIC: Application-aware Data Passing for Chained Serverless ApplicationsAshraf Mahgoub, Karthick Shankar, Subrata Mitra, Ana Klimovic 等USENIX ATC 2021 · 被引用 170 次
- Lambada: Interactive Data Analytics on Cold Data Using Serverless Cloud InfrastructureIngo Müller, Renato Marroquín, Gustavo AlonsoSIGMOD 2020 · 被引用 135 次
- COSE: Configuring Serverless Functions using Statistical LearningNabeel Akhtar, Ali Raza, Vatche Ishakian, Ibrahim MattaINFOCOM 2020 · 被引用 97 次
相关 Paper
- SPES: Towards Optimizing Performance-Resource Trade-Off for Serverless FunctionsCheryl Lee, Zhouruixin Zhu, Tianyi Yang, Yintong Huo 等ICDE 2024 · 被引用 13 次
- Jolteon: Unleashing the Promise of Serverless for Serverless WorkflowsZili Zhang, Chao Jin, Xin JinNSDI 2024 · 被引用 15 次
- Fork in the Road: Reflections and Optimizations for Cold Start Latency in Production Serverless SystemsXiaohu Chai, Tianyu Zhou, Keyang Hu, Jianfeng Tan 等OSDI 2025 · 被引用 7 次
- Metronome: Differentiated Delay Scheduling for Serverless FunctionsZhuangbin Chen, Juzheng Zheng, Zibin ZhengICSE 2026 · 被引用 1 次
- Concurrency-Informed Orchestration for Serverless FunctionsQichang Liu, Yue Cheng, Haiying Shen, Ao Wang 等ASPLOS 2025 · 被引用 7 次
