Lune

INFOCOM2026顶会

Sluice: End-to-End Latency Guarantee for Long-running Dataflow Systems

Zhaochen She, Yancan Mao, Richard T. B. Ma

2026年份

摘要

End-to-end latency is a key performance metric for long-running dataflow systems, as it directly impacts the timeliness and quality of results in real-time applications. In cloud environments, dynamic resource scaling is essential for handling workload fluctuations efficiently. However, maintaining consistent latency remains challenging due to resource contention and scaling-induced disruptions. Existing scaling techniques improve efficiency but fall short in guaranteeing latency, as they lack accurate latency estimation and timely scaling decisions. We present Sluice, a general framework for guaranteeing end-to-end latency. The core insight behind Sluice is to decompose latency into intrinsic and extrinsic components, isolating predictable, controllable delays from external disruptions like scaling and scheduling. Sluice dynamically reserves a buffer for extrinsic latency and scales resources to tightly control intrinsic latency. At its core, Sluice integrates (1) a novel Latency Estimation Model (LEM) that uses instantaneous backlog sizes along the critical path, and (2) a scaling controller that leverages LEM to make latency-aware, resource-efficient scaling decisions. We implement Sluice on Apache Flink and evaluate it using diverse real-world workloads. Results show that Sluice consistently enforces end-to-end latency guarantees under dynamic conditions while maintaining competitive resource efficiency, outperforming existing approaches in both latency control and task usage.

问问这篇 Paper

问问你的智能体。

Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。

可以从这些问题问起

智能体调用

Lunesearch_papers

在 Lune 里问

免费开始,无需绑卡

相关 Paper

黄昏的海面,两侧是细线勾勒的悬崖