LAMDA: Two-Phase HPO via Learning Prior from Low-Fidelity Data
Fan Li, Shengbo Wang, Ke Li
Abstract
Hyperparameter Optimization (HPO) is crucial in machine learning, aiming to optimize hyperparameters to enhance model performance. Although existing methods that leverage prior knowledge—drawn from either previous experiments or expert insights—can accelerate optimization, acquiring a correct prior for a specific HPO task is non-trivial. In this work, we propose to relieve the reliance on external knowledge by learning a reliable prior directly from low-fidelity (LF) problems. We introduce Lamda, an algorithm-agnostic framework designed to boost any baseline HPO algorithm. Specifically, Lamda operates in two phases: (1) it learns a reliable prior by exploring the LF landscape under limited computational budgets, and (2) it leverages this learned prior to guide the HPO process. We showcase how the Lamda framework can be integrated with various HPO algorithms to boost their performance, and further conduct theoretical analysis towards the integrated Bayesian optimization and bandit-based Hyperband. We conduct experiments on 56 HPO problems spanning diverse domains and model scales. Results show that Lamda consistently enhances its baseline algorithms. Compared to nine state-of-the-art HPO algorithms, our Lamda variant achieves the best performance in 51 out of 56 HPO tasks while it is the second best algorithm in the other 5 cases.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext c4b5aaef-e026-4aae-af08-775ffbe3878aBuilds on10
- Multi-fidelity Bayesian Optimization with Max-value Entropy Search and its ParallelizationShion Takeno, Hitoshi Fukuoka, Yuhki Tsukada, Toshiyuki Koyama et al.ICML 2020 · 83 citations
- Multi-Fidelity Bayesian Optimization via Deep Neural NetworksShibo Li, Wei W. Xing, Robert M. Kirby, Shandian ZheNeurIPS 2020 · 74 citations
- PFNs4BO: In-Context Learning for Bayesian OptimizationSamuel Müller, Matthias Feurer, Noah Hollmann, Frank HutterICML 2023 · 71 citations
- PriorBand: Practical Hyperparameter Optimization in the Age of Deep LearningNeeratyoy Mallik, Edward Bergman, Carl Hvarfner, Danny Stoll et al.NeurIPS 2023 · 50 citations
- MFES-HB: Efficient Hyperband with Multi-Fidelity Quality MeasurementsYang Li, Yu Shen, Jiawei Jiang, Jinyang Gao et al.AAAI 2021 · 32 citations
Related papers
- BO: Augmenting Acquisition Functions with User Beliefs for Bayesian OptimizationCarl Hvarfner, Danny Stoll, Artur L. F. Souza, Marius Lindauer et al.ICLR 2022 · 93 citations
- HyperSTAR: Task-Aware Hyperparameters for Deep NetworksGaurav Mittal, Chang Liu, Nikolaos Karianakis, Victor Fragoso et al.CVPR 2020
- Efficient Automatic CASH via Rising BanditsYang Li, Jiawei Jiang, Jinyang Gao, Yingxia Shao et al.AAAI 2020 · 45 citations
- HyperJump: Accelerating HyperBand via Risk ModellingPedro Mendes, Maria Casimiro, Paolo Romano, David GarlanAAAI 2023 · 11 citations
- Deep Ranking Ensembles for Hyperparameter OptimizationAbdus Salam Khazi, Sebastian Pineda-Arango, Josif GrabockaICLR 2023 · 1 citation
