SASPAR: Shared Adaptive Stream Partitioning
Jeyhun Karimov, Hans-Arno Jacobsen
摘要
Data partitioning induces network transfers and dominates the cost of stream data analytics. Moreover, partitioning streaming data for multiple stream queries in the same cluster can easily saturate the network bandwidth and lead to high end-to-end latencies.The goal of this paper is to share the partition operation in streaming workloads and maximize the sharing opportunities for multiple stream queries. However, there are several challenges, such as minimizing data copy, optimizing the partitioning strategy for multiple queries, and minimizing latency.We propose SASPAR, Shared Adaptive Stream Partitioner, which is able to share data partitioning among multiple stream queries. Our contributions are threefold. First, we propose a new technique to optimize the partitioning strategy for multiple stream queries. Second, we present an adaptive query execution framework that performs optimizations at run-time, without stopping the query execution plan. Third, we utilize meta-heuristics and machine learning when solving the underlying optimization problem takes more time than expected.SASPAR is designed as a versatile layer to sit on top of a stream processing engine (SPE). We operate SASPAR on top of three state-of-the-art SPEs with hundreds of stream queries. Our experimental results show that SASPAR improves the performance (throughput and latency) of all underlying SPEs by up to 3x.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
相关 Paper
- SaSPartitioner: A Self-Adaptive Streaming Partitioner Using Deep Reinforcement LearningShenghao Gong, Liu Liu, Ziquan Fang, Yunjun Gao 等ICDE 2026
- Process Faster, Pay Less: Functional Isolation for Stream ProcessingEleni Zapridou, Michael Koepf, Panagiotis Sioulas, Ioannis Mytilinis 等ICDE 2026
- Accelerating Stream Processing Engines via Hardware OffloadingZhengyan Guo, Mingxing Zhang, Yingdi Shan, Kang Chen 等SIGMOD 2026
- AJoin: Ad-hoc Stream Joins at ScaleJeyhun Karimov, Tilmann Rabl, Volker MarklVLDB 2020 · 被引用 14 次
- Resource-efficient Shared Query Execution via Exploiting Time SlacknessDixin Tang, Zechao Shang, William W. Ma, Aaron J. Elmore 等SIGMOD 2021 · 被引用 4 次
