Modeling All Response Surfaces in One for Conditional Search Spaces
Jiaxing Li, Wei Liu, Chao Xue, Yibing Zhan, Xiaoxing Wang, Weifeng Liu, Dacheng Tao
Abstract
Bayesian Optimization (BO) is a sample-efficient black-box optimizer commonly used in search spaces where hyperparameters are independent. However, in many practical AutoML scenarios, there will be dependencies among hyperparameters, forming a conditional search space, which can be partitioned into structurally distinct subspaces. The structure and dimensionality of hyperparameter configurations vary across these subspaces, challenging the application of BO. Some previous BO works have proposed solutions to develop multiple Gaussian Process models in these subspaces. However, these approaches tend to be inefficient as they require a substantial number of observations to guarantee each GP's performance and cannot capture relationships between hyperparameters across different subspaces. To address these issues, this paper proposes a novel approach to model the response surfaces of all subspaces in one, which can model the relationships between hyperparameters elegantly via a self-attention mechanism. Concretely, we design a structure-aware hyperparameter embedding to preserve the structural information. Then, we introduce an attention-based deep feature extractor, capable of projecting configurations with different structures from various subspaces into a unified feature space, where the response surfaces can be formulated using a single standard Gaussian Process. The empirical results on a simulation function, various real-world tasks, and HPO-B benchmark demonstrate that our proposed approach improves the efficacy and efficiency of BO within conditional search spaces.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 48c26444-19d1-43b5-b4d3-13b97fb1e0cdBuilds on8
- Unexpected Improvements to Expected Improvement for Bayesian OptimizationSebastian Ament, Samuel Daulton, David Eriksson, Maximilian Balandat et al.NeurIPS 2023 · 280 citations
- Sample-Efficient Optimization in the Latent Space of Deep Generative Models via Weighted RetrainingAustin Tripp, Erik A. Daxberger, José Miguel Hernández-LobatoNeurIPS 2020 · 186 citations
- Few-Shot Bayesian Optimization with Deep Kernel SurrogatesMartin Wistuba, Josif GrabockaICLR 2021 · 87 citations
- Increasing the Scope as You Learn: Adaptive Bayesian Optimization in Nested SubspacesLeonard Papenmeier, Luigi Nardi, Matthias PoloczekNeurIPS 2022 · 76 citations
- Combining Latent Space and Structured Kernels for Bayesian Optimization over Combinatorial SpacesAryan Deshwal, Janardhan Rao DoppaNeurIPS 2021 · 65 citations
Related papers
- Meta-Learning Acquisition Functions for Transfer Learning in Bayesian OptimizationMichael Volpp, Lukas P. Fröhlich, Kirsten Fischer, Andreas Doerr et al.ICLR 2020 · 104 citations
- High-Dimensional Bayesian Optimization via Nested Riemannian ManifoldsNoémie Jaquier, Leonel Dario RozoNeurIPS 2020 · 33 citations
- Bayesian Optimization for Simultaneous Selection of Machine Learning Algorithms and Hyperparameters on Shared Latent SpaceKazuki Ishikawa, Ryota Ozaki, Yohei Kanzaki, Ichiro Takeuchi et al.KDD 2025 · 1 citation
- Scalable First-Order Bayesian Optimization via Structured Automatic DifferentiationSebastian E. Ament, Carla P. GomesICML 2022 · 12 citations
- Deep Pipeline Embeddings for AutoMLSebastian Pineda-Arango, Josif GrabockaKDD 2023 · 6 citations
