GRANNY: Granular Management of Compute-Intensive Applications in the Cloud
Carlos Segarra, Simon Shillaker, Guo Li, Eleftheria Mappoura, Rodrigo Bruno, Lluís Vilanova, Peter R. Pietzuch
摘要
Parallel applications are typically implemented using multithreading (with shared memory, e.g., OpenMP) or multiprocessing (with message passing, e.g., MPI). While it seems attractive to deploy such applications in cloud virtual machines (VMs), existing cloud schedulers fail to manage such applications efficiently: they cannot scale multi-threaded applications dynamically when more CPU cores in a VM become available, and they cause fragmentation over time due to the static allocation of multi-process applications to VMs.
We describe GRANNY, a new distributed runtime that enables the fine-granular management of multi-threaded/process applications in cloud environments. GRANNY supports the vertical scaling of multi-threaded applications within a VM and the horizontal migration of multi-process applications between VMs. GRANNY achieves both through a single WebAssembly-based execution abstraction: Granules can execute application code with thread or process semantics and allow for efficient snapshotting. GRANNY scales up applications by adding more Granules, and de-fragments applications by migrating Granules between VMs. In both cases, it launches new Granules from snapshots efficiently. We evaluate GRANNY with dynamic scheduling policies and show that, compared to current schedulers, it reduces the makespan for OpenMP workloads by up to 60% and the fragmentation for MPI workloads by up to 25%.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Burst Computing: Quick, Sudden, Massively Parallel Processing on Serverless ResourcesDaniel Barcelona Pons, Aitor Arjona, Pedro García López, Enrique Molina-Giménez 等USENIX ATC 2025 · 被引用 3 次
- Hierarchical Integration of WebAssembly in Serverless for Efficiency and InteroperabilityMohammadamin Baqershahi, Changyuan Lin, Visal Saosuo, Paul Chen 等NSDI 2026 · 被引用 2 次
它引用的顶会 Paper6
- Faasm: Lightweight Isolation for Efficient Stateful Serverless ComputingSimon Shillaker, Peter R. PietzuchUSENIX ATC 2020 · 被引用 382 次
- No Provisioned Concurrency: Fast RDMA-codesigned Remote Fork for Serverless ComputingXingda Wei, Fangming Lu, Tianxia Wang, Jinyu Gu 等OSDI 2023 · 被引用 78 次
- MigrOS: Transparent Live-Migration Support for Containerised RDMA ApplicationsMaksym Planeta, Jan Bierbaum, Leo Sahaya Daphne Antony, Torsten Hoefler 等USENIX ATC 2021 · 被引用 27 次
- Going beyond the Limits of SFI: Flexible and Secure Hardware-Assisted In-Process Isolation with HFIShravan Narayan, Tal Garfinkel, Mohammadkazem Taram, Joey Rudek 等ASPLOS 2023 · 被引用 27 次
- RIBBON: cost-effective and qos-aware deep learning model inference using a diverse pool of cloud computing instancesBaolin Li, Rohan Basu Roy, Tirthak Patel, Vijay Gadepally 等SC 2021 · 被引用 16 次
相关 Paper
- Advanced synchronization techniques for task-based runtime systemsDavid Álvarez, Kevin Sala, Marcos Maroñas, Aleix Roca 等PPoPP 2021 · 被引用 24 次
- Emma: Elastic Multi-Resource Management for Realtime Stream ProcessingRengan Dou, Xin Wang, Richard T. B. MaINFOCOM 2024 · 被引用 2 次
- Ditto: Efficient Serverless Analytics with Elastic ParallelismChao Jin, Zili Zhang, Xingyu Xiang, Songyun Zou 等SIGCOMM 2023 · 被引用 27 次
- Nu: Achieving Microsecond-Scale Resource Fungibility with Logical ProcessesZhenyuan Ruan, Seo Jin Park, Marcos K. Aguilera, Adam Belay 等NSDI 2023
- GRIT: Enhancing Multi-GPU Performance with Fine-Grained Dynamic Page PlacementYueqi Wang, Bingyao Li, Aamer Jaleel, Jun Yang 等HPCA 2024 · 被引用 19 次
