Meta-learning Hyperparameter Performance Prediction with Neural Processes
Ying Wei, Peilin Zhao, Junzhou Huang
Abstract
The surrogate that predicts the performance of hyperparameters has been a key component for sequential model-based hyperparameter optimization. In practical applications, a trial of a hyperparameter configuration may be so costly that a surrogate is expected to return an optimal configuration with as few trials as possible. Observing that human experts draw on their expertise in a machine learning model by trying configurations that once performed well on other datasets, we are inspired to build a trial-efficient surrogate by transferring the meta-knowledge learned from historical trials on other datasets. We propose an end-to-end surrogate named as Transfer Neural Processes (TNP) that learns a comprehensive set of meta-knowledge, including the parameters of historical surrogates, historical trials, and initial configurations for other datasets. Experiments on extensive OpenML datasets and three computer vision datasets demonstrate that the proposed algorithm achieves state-of-the-art performance in at least one order of magnitude less trials.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers8
- Meta Discovery: Learning to Discover Novel Classes given Very Limited DataHaoang Chi, Feng Liu, Wenjing Yang, Long Lan et al.ICLR 2022 · 52 citations
- Episodic Multi-Task Learning with Heterogeneous Neural ProcessesJiayi Shen, Xiantong Zhen, Qi Wang, Marcel WorringNeurIPS 2023 · 21 citations
- What Matters For Meta-Learning Vision Regression Tasks?Ning Gao, Hanna Ziesche, Ngo Anh Vien, Michael Volpp et al.CVPR 2022 · 14 citations
- TransBO: Hyperparameter Optimization via Two-Phase Transfer LearningYang Li, Yu Shen, Huaijun Jiang, Wentao Zhang et al.KDD 2022 · 12 citations
- Monte Carlo Tree Search based Space Transfer for Black Box OptimizationShukuan Wang, Ke Xue, Lei Song, Xiaobin Huang et al.NeurIPS 2024 · 11 citations
Builds on1
Related papers
- Transfer NAS with Meta-learned Bayesian SurrogatesGresa Shala, Thomas Elsken, Frank Hutter, Josif GrabockaICLR 2023
- Neural Architecture and Hyperparameter Selection Through Meta-Learning on Time SeriesErfan Moeini, Christopher Vox, Marie Anastacio, Wadie Skaf et al.AAAI 2026 · 1 citation
- Few-Shot Bayesian Optimization with Deep Kernel SurrogatesMartin Wistuba, Josif GrabockaICLR 2021 · 87 citations
- Deep Ranking Ensembles for Hyperparameter OptimizationAbdus Salam Khazi, Sebastian Pineda-Arango, Josif GrabockaICLR 2023 · 1 citation
- AutoTransfer: AutoML with Knowledge Transfer - An Application to Graph Neural NetworksKaidi Cao, Jiaxuan You, Jiaju Liu, Jure LeskovecICLR 2023
