Scarf: Self-Adaptive Tuning via Multi-Objective Reinforcement Learning for Apache Flink
Liu Liu, Shenghao Gong, Ziquan Fang, Yunjun Gao
摘要
Distributed stream processing systems (DSPSs) such as Apache Flink have become omnipresent for real-time data processing in e-commerce, finance, telecommunications, etc. The execution behavior of Flink is controlled by a vast and complex space of configuration knobs, necessitating automatic knob tuning to economize resource usage while maintaining sufficient processing capabilities for a given workload. Existing automatic methods largely adjust limited configuration knobs, respond slowly to dynamic workloads, and have difficulty transferring knowledge between heterogeneous jobs with diverse knob spaces.
To solve these problems, we present Scarf , a self-adaptive configuration tuning framework using multi-objective reinforcement learning (RL) for Apache Flink. Specifically, (1) we accelerate job-specific knob selection by clustering historical workloads according to their sensitivity to knob changes, dramatically reducing redundant sampling; (2) we formulate tuning as a multi-objective RL problem that jointly optimizes throughput and resource usage, learning a forest of RL models offline representing the Pareto front of the configurations, and dynamically selecting configurations from the Pareto front under fluctuating online workloads; (3) we enable rapid adaptation to new job topologies via a transferable actor-critic architecture based on graph neural networks (GNNs), complemented with a progressive neural-network (PNN) warm-up strategy. We implement Scarf on Apache Flink and evaluate it on a diverse range of streaming applications. Our framework significantly outperforms state-of-the-art DSPS tuning approaches, achieving up to 62.5% savings in CPU resources, 68.3% savings in memory usage, 77.1% reduction in online tuning time, while maintaining sufficient processing abilities throughout workload fluctuations.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper11
- AWARE: Automate Workload Autoscaling with Reinforcement Learning in Production Cloud SystemsHaoran Qiu, Weichao Mao, Chen Wang, Hubertus Franke 等USENIX ATC 2023 · 被引用 95 次
- Facilitating Database Tuning with Hyper-Parameter Optimization: A Comprehensive Experimental EvaluationXinyi Zhang, Zhuo Chang, Yang Li, Hong Wu 等VLDB 2022 · 被引用 88 次
- LlamaTune: Sample-Efficient DBMS Configuration TuningKonstantinos Kanellis, Cong Ding, Brian Kroth, Andreas Müller 等VLDB 2022 · 被引用 73 次
- DB-BERT: A Database Tuning Tool that "Reads the Manual"Immanuel TrummerSIGMOD 2022 · 被引用 71 次
- UDO: Universal Database Optimization using Reinforcement LearningJunxiong Wang, Immanuel Trummer, Debabrota BasuVLDB 2021 · 被引用 53 次
相关 Paper
- SaSPartitioner: A Self-Adaptive Streaming Partitioner Using Deep Reinforcement LearningShenghao Gong, Liu Liu, Ziquan Fang, Yunjun Gao 等ICDE 2026
- Learning from the Past: Adaptive Parallelism Tuning for Stream Processing SystemsYuxing Han, Lixiang Chen, Haoyu Wang, Zhanghao Chen 等ICDE 2025 · 被引用 2 次
- Adaptive Code Learning for Spark Configuration TuningChen Lin, Junqing Zhuang, Jiadong Feng, Hui Li 等ICDE 2022 · 被引用 28 次
- ZERoTuNE: Learned Zero-Shot Cost Models for Parallelism Tuning in Stream ProcessingPratyush Agnihotri, Boris Koldehofe, Paul Stiegele, Roman Heinrich 等ICDE 2024 · 被引用 13 次
- Alzo: Auto-Tuning with Reinforcement Learning for DAG-based BlockchainsQiuyu Ding, Rongkai Zhang, Qinnan Zhang, Zhen Xiao 等WWW 2026
