Meta-learning Hyperparameter Performance Prediction with Neural Processes
Ying Wei, Peilin Zhao, Junzhou Huang
摘要
The surrogate that predicts the performance of hyperparameters has been a key component for sequential model-based hyperparameter optimization. In practical applications, a trial of a hyperparameter configuration may be so costly that a surrogate is expected to return an optimal configuration with as few trials as possible. Observing that human experts draw on their expertise in a machine learning model by trying configurations that once performed well on other datasets, we are inspired to build a trial-efficient surrogate by transferring the meta-knowledge learned from historical trials on other datasets. We propose an end-to-end surrogate named as Transfer Neural Processes (TNP) that learns a comprehensive set of meta-knowledge, including the parameters of historical surrogates, historical trials, and initial configurations for other datasets. Experiments on extensive OpenML datasets and three computer vision datasets demonstrate that the proposed algorithm achieves state-of-the-art performance in at least one order of magnitude less trials.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper8
- Meta Discovery: Learning to Discover Novel Classes given Very Limited DataHaoang Chi, Feng Liu, Wenjing Yang, Long Lan 等ICLR 2022 · 被引用 52 次
- Episodic Multi-Task Learning with Heterogeneous Neural ProcessesJiayi Shen, Xiantong Zhen, Qi Wang, Marcel WorringNeurIPS 2023 · 被引用 21 次
- What Matters For Meta-Learning Vision Regression Tasks?Ning Gao, Hanna Ziesche, Ngo Anh Vien, Michael Volpp 等CVPR 2022 · 被引用 14 次
- TransBO: Hyperparameter Optimization via Two-Phase Transfer LearningYang Li, Yu Shen, Huaijun Jiang, Wentao Zhang 等KDD 2022 · 被引用 12 次
- Monte Carlo Tree Search based Space Transfer for Black Box OptimizationShukuan Wang, Ke Xue, Lei Song, Xiaobin Huang 等NeurIPS 2024 · 被引用 11 次
它引用的顶会 Paper1
相关 Paper
- Transfer NAS with Meta-learned Bayesian SurrogatesGresa Shala, Thomas Elsken, Frank Hutter, Josif GrabockaICLR 2023
- Neural Architecture and Hyperparameter Selection Through Meta-Learning on Time SeriesErfan Moeini, Christopher Vox, Marie Anastacio, Wadie Skaf 等AAAI 2026 · 被引用 1 次
- Few-Shot Bayesian Optimization with Deep Kernel SurrogatesMartin Wistuba, Josif GrabockaICLR 2021 · 被引用 87 次
- Deep Ranking Ensembles for Hyperparameter OptimizationAbdus Salam Khazi, Sebastian Pineda-Arango, Josif GrabockaICLR 2023 · 被引用 1 次
- AutoTransfer: AutoML with Knowledge Transfer - An Application to Graph Neural NetworksKaidi Cao, Jiaxuan You, Jiaju Liu, Jure LeskovecICLR 2023
