MFES-HB: Efficient Hyperband with Multi-Fidelity Quality Measurements
Yang Li, Yu Shen, Jiawei Jiang, Jinyang Gao, Ce Zhang, Bin Cui
Abstract
Hyperparameter optimization (HPO) is a fundamental problem in automatic machine learning (AutoML). However, due to the expensive evaluation cost of models (e.g., training deep learning models or training models on large datasets), vanilla Bayesian optimization (BO) is typically computationally infeasible. To alleviate this issue, Hyperband (HB) utilizes the early stopping mechanism to speed up configuration evaluations by terminating those badly-performing configurations in advance. This leads to two kinds of quality measurements: (1) many low-fidelity measurements for configurations that get early-stopped, and (2) few high-fidelity measurements for configurations that are evaluated without being early stopped. The state-of-the-art HB-style method, BOHB, aims to combine the benefits of both BO and HB. Instead of sampling configurations randomly in HB, BOHB samples configurations based on a BO surrogate model, which is constructed with the high-fidelity measurements only. However, the scarcity of high-fidelity measurements greatly hampers the efficiency of BO to guide the configuration search.
In this paper, we present MFES-HB, an efficient Hyperband method that is capable of utilizing both the high-fidelity and low-fidelity measurements to accelerate the convergence of HPO tasks. Designing MFES-HB is not trivial as the low-fidelity measurements can be biased yet informative to guide the configuration search. Thus we propose to build a Multi-Fidelity Ensemble Surrogate (MFES) based on the generalized Product of Experts framework, which can integrate useful information from multi-fidelity measurements effectively. The empirical studies on the real-world AutoML tasks demonstrate that MFES-HB can achieve 3.3-8.9x speedups over the state-of-the-art approach --- BOHB.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers15
- Facilitating Database Tuning with Hyper-Parameter Optimization: A Comprehensive Experimental EvaluationXinyi Zhang, Zhuo Chang, Yang Li, Hong Wu et al.VLDB 2022 · 88 citations
- PaSca: A Graph Neural Architecture Search System under the Scalable ParadigmWentao Zhang, Yu Shen, Zheyu Lin, Yang Li et al.WWW 2022 · 69 citations
- VolcanoML: Speeding up End-to-End AutoML via Scalable Search Space DecompositionYang Li, Yu Shen, Wentao Zhang, Jiawei Jiang et al.VLDB 2021 · 55 citations
- Hyper-Tune: Towards Efficient Hyper-parameter Tuning at ScaleYang Li, Yu Shen, Huaijun Jiang, Wentao Zhang et al.VLDB 2022 · 32 citations
- Hydro: Surrogate-Based Hyperparameter Tuning Service in DatacentersQinghao Hu, Zhisheng Ye, Meng Zhang, Qiaoling Chen et al.OSDI 2023 · 16 citations
Builds on1
Related papers
- Efficient Hyperparameter Optimization with Adaptive Fidelity IdentificationJiantong Jiang, Zeyi Wen, Atif Bin Mansoor, Ajmal MianCVPR 2024
- LAMDA: Two-Phase HPO via Learning Prior from Low-Fidelity DataFan Li, Shengbo Wang, Ke LiAAAI 2026
- HyperJump: Accelerating HyperBand via Risk ModellingPedro Mendes, Maria Casimiro, Paolo Romano, David GarlanAAAI 2023 · 11 citations
- Cost-Sensitive Freeze-thaw Bayesian Optimization for Efficient Hyperparameter TuningDong Bok Lee, Aoxuan Silvia Zhang, Byungjoo Kim, Junhyeon Park et al.NeurIPS 2025 · 2 citations
- Bayesian Optimization for Simultaneous Selection of Machine Learning Algorithms and Hyperparameters on Shared Latent SpaceKazuki Ishikawa, Ryota Ozaki, Yohei Kanzaki, Ichiro Takeuchi et al.KDD 2025 · 1 citation
