Ditto: Efficient Serverless Analytics with Elastic Parallelism
Chao Jin, Zili Zhang, Xingyu Xiang, Songyun Zou, Gang Huang, Xuanzhe Liu, Xin Jin
Abstract
Serverless computing provides fine-grained resource elasticity for data analytics-a job can flexibly scale its resources for each stage, instead of sticking to a fixed pool of resources throughout its lifetime. Due to different data dependencies and different shuffling overheads caused by intra-and inter-server communication, the best degree of parallelism (DoP) for each stage varies based on runtime conditions.
We present Ditto, a job scheduler for serverless analytics that leverages fine-grained resource elasticity to optimize for job completion time (JCT) and cost. The key idea of Ditto is to use a new scheduling granularity-stage group-to decouple parallelism configuration from function placement. Ditto bundles stages into stage groups based on their data dependencies and IO characteristics. It exploits the parallelized time characteristics of the stages to determine the parallelism configuration, and prioritizes the placement of stage groups with large shuffling traffic, so that the stages in these groups can leverage zero-copy intra-server communication for efficient shuffling. We build a system prototype of Ditto and evaluate it with a variety of benchmarking workloads. Experimental results show that Ditto outperforms existing solutions by up to 2.5× on JCT and up to 1.8× on cost.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers8
- Jolteon: Unleashing the Promise of Serverless for Serverless WorkflowsZili Zhang, Chao Jin, Xin JinNSDI 2024 · 15 citations
- Making Serverless Pay-For-Use a Reality with LeopardTingjia Cao, Andrea C. Arpaci-Dusseau, Remzi H. Arpaci-Dusseau, Tyler Caraza-HarterNSDI 2025 · 11 citations
- Optimizing Distributed Deployment of Mixture-of-Experts Model Inference in Serverless ComputingMengfan Liu, Wei Wang, Chuan WuINFOCOM 2025 · 6 citations
- Online Container Caching with Late-Warm for IoT Data ProcessingGuopeng Li, Haisheng Tan, Xuan Zhang, Chi Zhang et al.ICDE 2024 · 4 citations
- Burst Computing: Quick, Sudden, Massively Parallel Processing on Serverless ResourcesDaniel Barcelona Pons, Aitor Arjona, Pedro García López, Enrique Molina-Giménez et al.USENIX ATC 2025 · 3 citations
Builds on6
- Nightcore: efficient and scalable serverless computing for latency-sensitive, interactive microservicesZhipeng Jia, Emmett WitchelASPLOS 2021 · 218 citations
- Batch: machine learning inference serving on serverless platforms with adaptive batchingAhsan Ali, Riccardo Pinciroli, Feng Yan, Evgenia SmirniSC 2020 · 184 citations
- SONIC: Application-aware Data Passing for Chained Serverless ApplicationsAshraf Mahgoub, Karthick Shankar, Subrata Mitra, Ana Klimovic et al.USENIX ATC 2021 · 170 citations
- Faastlane: Accelerating Function-as-a-Service WorkflowsSwaroop Kotni, Ajay Nayak, Vinod Ganapathy, Arkaprava BasuUSENIX ATC 2021 · 142 citations
- SPRIGHT: extracting the server from serverless computing! high-performance eBPF-based event-driven, shared-memory processingShixiong Qi, Leslie Monis, Ziteng Zeng, Ian-Chin Wang et al.SIGCOMM 2022 · 85 citations
Related papers
- Caerus: NIMBLE Task Scheduling for Serverless AnalyticsHong Zhang, Yupeng Tang, Anurag Khandelwal, Jingrong Chen et al.NSDI 2021 · 75 citations
- MinFlow: High-performance and Cost-efficient Data Passing for I/O-intensive Stateful Serverless AnalyticsTao Li, Yongkun Li, Wenzhe Zhu, Yinlong Xu et al.FAST 2024 · 5 citations
- AQUATOPE: QoS-and-Uncertainty-Aware Resource Management for Multi-stage Serverless WorkflowsZhuangzhuang Zhou, Yanqi Zhang, Christina DelimitrouASPLOS 2023 · 78 citations
- Demeter: Fine-grained Function Orchestration for Geo-distributed Serverless AnalyticsXiaofei Yue, Song Yang, Liehuang Zhu, Stojan Trajanovski et al.INFOCOM 2024 · 13 citations
- Dilu: Enabling GPU Resourcing-on-Demand for Serverless DL Serving via Introspective ElasticityCunchi Lv, Xiao Shi, Zhengyu Lei, Jinyue Huang et al.ASPLOS 2025 · 10 citations
