Generalizable and interpretable learning for configuration extrapolation
Yi Ding, Ahsan Pervaiz, Michael Carbin, Henry Hoffmann
摘要
Modern software applications are increasingly configurable, which puts a burden on users to tune these configurations for their target hardware and workloads. To help users, machine learning techniques can model the complex relationships between software configuration parameters and performance. While powerful, these learners have two major drawbacks: (1) they rarely incorporate prior knowledge and (2) they produce outputs that are not interpretable by users. These limitations make it difficult to ( 1) leverage information a user has already collected (e.g., tuning for new hardware using the best configurations from old hardware) and ( 2) gain insights into the learner's behavior (e.g., understanding why the learner chose different configurations on different hardware or for different workloads). To address these issues, this paper presents two configuration extrapolation tools, Gil and Gil+, using the proposed generalizable and interpretable learning approaches. To incorporate prior knowledge, the proposed tools (1) start from known configurations, (2) iteratively construct a new linear model, (3) extrapolate better performance configurations from that model, and (4) repeat. Since the base learners are linear models, these tools are inherently interpretable. We enhance this property with a graphical representation of how they arrived at the highest performance configuration. We evaluate Gil and Gil+ by using them to configure Apache Spark workloads on different hardware platforms and find that, compared to prior work, Gil and Gil+ produce comparable, and sometimes even better performance configurations, but with interpretable results.
• Software and its engineering → Software configuration management and version control systems; • Computing methodologies → Machine learning approaches.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- Unicorn: reasoning about configurable system performance through the lens of causalityMd Shahriar Iqbal, Rahul Krishna, Mohammad Ali Javidian, Baishakhi Ray 等EuroSys 2022 · 被引用 60 次
- CAFQA: A Classical Simulation Bootstrap for Variational Quantum AlgorithmsGokul Subramanian Ravi, Pranav Gokhale, Yi Ding, William M. Kirby 等ASPLOS 2023 · 被引用 39 次
- Bayesian Multi-Level Performance Models for Multi-Factor Variability of Configurable Software SystemsJohannes Dorn, Stefan Mühlbauer, Stefan Jahns, Sven Apel 等ICSE 2026
它引用的顶会 Paper4
- Lessons Learned from the Chameleon TestbedKate Keahey, Jason Anderson, Zhuo Zhen, Pierre Riteau 等USENIX ATC 2020 · 被引用 398 次
- Statically inferring performance properties of software configurationsChi Li, Shu Wang, Henry Hoffmann, Shan LuEuroSys 2020 · 被引用 25 次
- ALERT: Accurate Learning for Energy and TimelinessChengcheng Wan, Muhammad Husni Santriaji, Eri Rogers, Henry Hoffmann 等USENIX ATC 2020 · 被引用 15 次
- White-Box Analysis over Machine Learning: Modeling Performance of Configurable SystemsMiguel Velez, Pooyan Jamshidi, Norbert Siegmund, Sven Apel 等ICSE 2021 · 被引用 5 次
相关 Paper
- Analysing the Impact of Workloads on Modeling the Performance of Configurable Software SystemsStefan Mühlbauer, Florian Sattler, Christian Kaltenecker, Johannes Dorn 等ICSE 2023 · 被引用 20 次
- Adaptive Code Learning for Spark Configuration TuningChen Lin, Junqing Zhuang, Jiadong Feng, Hui Li 等ICDE 2022 · 被引用 28 次
- White-Box Performance-Influence Models: A Profiling and Learning ApproachMax Weber, Sven Apel, Norbert SiegmundICSE 2021 · 被引用 2 次
- LOCAT: Low-Overhead Online Configuration Auto-Tuning of Spark SQL ApplicationsJinhan Xin, Kai Hwang, Zhibin YuSIGMOD 2022 · 被引用 34 次
- Explainable Database Management System Configuration Tuning through CounterfactualsXinyue Shao, Hongzhi Wang, Xiao Zhu, Tianyu Mu 等ICDE 2024
