LAMDA: Two-Phase HPO via Learning Prior from Low-Fidelity Data
Fan Li, Shengbo Wang, Ke Li
摘要
Hyperparameter Optimization (HPO) is crucial in machine learning, aiming to optimize hyperparameters to enhance model performance. Although existing methods that leverage prior knowledge—drawn from either previous experiments or expert insights—can accelerate optimization, acquiring a correct prior for a specific HPO task is non-trivial. In this work, we propose to relieve the reliance on external knowledge by learning a reliable prior directly from low-fidelity (LF) problems. We introduce Lamda, an algorithm-agnostic framework designed to boost any baseline HPO algorithm. Specifically, Lamda operates in two phases: (1) it learns a reliable prior by exploring the LF landscape under limited computational budgets, and (2) it leverages this learned prior to guide the HPO process. We showcase how the Lamda framework can be integrated with various HPO algorithms to boost their performance, and further conduct theoretical analysis towards the integrated Bayesian optimization and bandit-based Hyperband. We conduct experiments on 56 HPO problems spanning diverse domains and model scales. Results show that Lamda consistently enhances its baseline algorithms. Compared to nine state-of-the-art HPO algorithms, our Lamda variant achieves the best performance in 51 out of 56 HPO tasks while it is the second best algorithm in the other 5 cases.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper10
- Multi-fidelity Bayesian Optimization with Max-value Entropy Search and its ParallelizationShion Takeno, Hitoshi Fukuoka, Yuhki Tsukada, Toshiyuki Koyama 等ICML 2020 · 被引用 83 次
- Multi-Fidelity Bayesian Optimization via Deep Neural NetworksShibo Li, Wei W. Xing, Robert M. Kirby, Shandian ZheNeurIPS 2020 · 被引用 74 次
- PFNs4BO: In-Context Learning for Bayesian OptimizationSamuel Müller, Matthias Feurer, Noah Hollmann, Frank HutterICML 2023 · 被引用 71 次
- PriorBand: Practical Hyperparameter Optimization in the Age of Deep LearningNeeratyoy Mallik, Edward Bergman, Carl Hvarfner, Danny Stoll 等NeurIPS 2023 · 被引用 50 次
- MFES-HB: Efficient Hyperband with Multi-Fidelity Quality MeasurementsYang Li, Yu Shen, Jiawei Jiang, Jinyang Gao 等AAAI 2021 · 被引用 32 次
相关 Paper
- BO: Augmenting Acquisition Functions with User Beliefs for Bayesian OptimizationCarl Hvarfner, Danny Stoll, Artur L. F. Souza, Marius Lindauer 等ICLR 2022 · 被引用 93 次
- HyperSTAR: Task-Aware Hyperparameters for Deep NetworksGaurav Mittal, Chang Liu, Nikolaos Karianakis, Victor Fragoso 等CVPR 2020
- Efficient Automatic CASH via Rising BanditsYang Li, Jiawei Jiang, Jinyang Gao, Yingxia Shao 等AAAI 2020 · 被引用 45 次
- HyperJump: Accelerating HyperBand via Risk ModellingPedro Mendes, Maria Casimiro, Paolo Romano, David GarlanAAAI 2023 · 被引用 11 次
- Deep Ranking Ensembles for Hyperparameter OptimizationAbdus Salam Khazi, Sebastian Pineda-Arango, Josif GrabockaICLR 2023 · 被引用 1 次
