Emma: Elastic Multi-Resource Management for Realtime Stream Processing
Rengan Dou, Xin Wang, Richard T. B. Ma
摘要
In stream processing applications, an operator is often instantiated into multiple parallel execution instances, referred to as executors, to facilitate large-scale data processing. Due to unpredictable changes in executor workloads, data tuples processed by different executors may exhibit varying latency. In particular, within the same operator, the executor with the maximum latency significantly impacts the end-to-end (E2E) latency of the application. Existing solutions, such as load balancing and horizontal scaling, which involve workload migration, often incur substantial time overhead induced by state migration and synchronization. In contrast, elastically scaling up/down resources of executors rather than moving workloads can not only effectively handle workload fluctuations but also offer rapid adjustments; however, prior works only considered CPU scaling with the assumption of sufficient memory.In this paper, we propose Emma, an elastic multi-resource manager. Emma leverages the resource elasticity of lightweight virtualization containers, e.g., Linux containers, to resize the resource of executors at runtime. The core of Emma is a multi-resource provisioning plan that conducts performance analysis and resource adjustment in real-time. We explore the relationship between resources and performance experimentally and theoretically, guiding the plan to adaptively allocate the appropriate combination of resources to each executor to 1) accommodate the dynamic workload; 2) efficiently utilize resources to enhance the performance of as many executors as possible. Additionally, we propose an online learning method that makes the manager seamlessly adapt to diverse stream applications. We integrate Emma with Apache Samza, and our experiments show that compared to existing solutions, Emma can significantly reduce latency by orders of magnitude in real-world applications.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
相关 Paper
- StreamSwitch: Fulfilling Latency Service-Layer Agreement for Stateful StreamingZhaochen She, Yancan Mao, Hailin Xiang, Xin Wang 等INFOCOM 2023 · 被引用 5 次
- Enjima: A Resource-Adaptive Stream Processing SystemLasantha Fernando, Taebin Kim, Khuzaima Daudjee, Tilmann RablSIGMOD 2026
- Move Fast and Meet Deadlines: Fine-grained Real-time Stream Processing with CameoLe Xu, Shivaram Venkataraman, Indranil Gupta, Luo Mai 等NSDI 2021 · 被引用 38 次
- Latency-Oriented Elastic Memory Management at Task-Granularity for Stateful Streaming ProcessingRengan Dou, Richard T. B. MaINFOCOM 2023 · 被引用 2 次
- Meces: Latency-efficient Rescaling via Prioritized State Migration for Stateful Distributed Stream Processing SystemsRong Gu, Han Yin, Weichang Zhong, Chunfeng Yuan 等USENIX ATC 2022 · 被引用 22 次
