BaCO: A Fast and Portable Bayesian Compiler Optimization Framework
Erik Orm Hellsten, Artur L. F. Souza, Johannes Lenfers, Rubens Lacouture, Olivia Hsu, Adel Ejjeh, Fredrik Kjolstad, Michel Steuwer, Kunle Olukotun, Luigi Nardi
Abstract
We introduce the Bayesian Compiler Optimization framework (BaCO), a general purpose autotuner for modern compilers targeting CPUs, GPUs, and FPGAs. BaCO provides the flexibility needed to handle the requirements of modern autotuning tasks. Particularly, it deals with permutation, ordered, and continuous parameter types along with both known and unknown parameter constraints. To reason about these parameter types and efficiently deliver high-quality code, BaCO uses Bayesian optimization algorithms specialized towards the autotuning domain. We demonstrate BaCO's effectiveness on three modern compiler systems: TACO, RISE & ELEVATE, and HPVM2FPGA for CPUs, GPUs, and FPGAs respectively. For these domains, BaCO outperforms current state-of-the-art auto-tuners by delivering on average 1.36X--1.56X faster code with a tiny search budget, and BaCO is able to reach expert-level performance 2.9X--3.9X faster.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 57864944-7c80-4e73-a94f-937db53a6bfdCited by top-tier papers9
- Bounce: Reliable High-Dimensional Bayesian Optimization for Combinatorial and Mixed SpacesLeonard Papenmeier, Luigi Nardi, Matthias PoloczekNeurIPS 2023 · 40 citations
- CATO: End-to-End Optimization of ML-Based Traffic Analysis PipelinesGerry Wan, Shinan Liu, Francesco Bronzino, Nick Feamster et al.NSDI 2025 · 16 citations
- SCOOT: SLO-Oriented Performance Tuning for LLM Inference EnginesKe Cheng, Zhi Wang, Wen Hu, Tiannuo Yang et al.WWW 2025 · 13 citations
- Kareus: Joint Reduction of Dynamic and Static Energy in Large Model TrainingRuofan Wu, Jae-Won Chung, Mosharaf ChowdhuryOSDI 2026 · 8 citations
- REASONING COMPILER: LLM-Guided Optimizations for Efficient Model ServingAnnabelle Sujun Tang, Christopher Priebe, Rohan Mahapatra, Lianhui Qin et al.NeurIPS 2025 · 7 citations
Builds on4
- Sparse GPU kernels for deep learningTrevor Gale, Matei Zaharia, Cliff Young, Erich ElsenSC 2020 · 170 citations
- A sparse iteration space transformation framework for sparse tensor algebraRyan Senanayake, Changwan Hong, Ziheng Wang, Amalee Wilson et al.OOPSLA 2020 · 51 citations
- GPTune: multitask learning for autotuning exascale applicationsYang Liu, Wissam M. Sid-Lakhdar, Osni Marques, Xinran Zhu et al.PPoPP 2021 · 45 citations
- Bliss: auto-tuning complex applications using a pool of diverse lightweight learning modelsRohan Basu Roy, Tirthak Patel, Vijay Gadepally, Devesh TiwariPLDI 2021 · 41 citations
Related papers
- Efficient Compiler Autotuning via Bayesian OptimizationJunjie Chen, Ningxin Xu, Peiqi Chen, Hongyu ZhangICSE 2021 · 73 citations
- CARBS: Compiler Autotuning via Randomized Biased SearchWei Li, Bin Gao, Weng-Fai WongHPDC 2026
- Bayesian Code Diffusion for Efficient Automatic Deep Learning Program OptimizationIsu Jeong, Seulki LeeOSDI 2025
- CoffeeBoost: Gradient Boosting Native Conformal Inference for Bayesian OptimizationYuanhao Lai, Pengfei Zheng, Chenpeng Ji, Cheng Qiu et al.AAAI 2025 · 1 citation
- FPBOXer: Efficient Input-Generation for Targeting Floating-Point Exceptions in GPU ProgramsAnh Tran, Ignacio Laguna, Ganesh GopalakrishnanHPDC 2024 · 3 citations
