Weighted Sampling for Combined Model Selection and Hyperparameter Tuning
Dimitrios Sarigiannis, Thomas P. Parnell, Haralampos Pozidis
摘要
The combined algorithm selection and hyperparameter tuning (CASH) problem is characterized by large hierarchical hyperparameter spaces. Model-free hyperparameter tuning methods can explore such large spaces efficiently since they are highly parallelizable across multiple machines. When no prior knowledge or meta-data exists to boost their performance, these methods commonly sample random configurations following a uniform distribution. In this work, we propose a novel sampling distribution as an alternative to uniform sampling and prove theoretically that it has a better chance of finding the best configuration in a worst-case setting. In order to compare competing methods rigorously in an experimental setting, one must perform statistical hypothesis testing. We show that there is little-to-no agreement in the automated machine learning literature regarding which methods should be used. We contrast this disparity with the methods recommended by the broader statistics literature, and identify a suitable approach. We then select three popular model-free solutions to CASH and evaluate their performance, with uniform sampling as well as the proposed sampling scheme, across 67 datasets from the OpenML platform. We investigate the trade-off between exploration and exploitation across the three algorithms, and verify empirically that the proposed sampling distribution improves performance in all cases.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
相关 Paper
- Efficient Automatic CASH via Rising BanditsYang Li, Jiawei Jiang, Jinyang Gao, Yingxia Shao 等AAAI 2020 · 被引用 45 次
- Put CASH on Bandits: A Max K-Armed Problem for Automated Machine LearningAmir Rezaei Balef, Claire Vernade, Katharina EggenspergerNeurIPS 2025 · 被引用 4 次
- TSC-AutoML: Meta-learning for Automatic Time Series Classification Algorithm SelectionTianyu Mu, Hongzhi Wang, Shenghe Zheng, Zhiyu Liang 等ICDE 2023 · 被引用 12 次
- DivBO: Diversity-aware CASH for Ensemble LearningYu Shen, Yupeng Lu, Yang Li, Yaofeng Tu 等NeurIPS 2022 · 被引用 15 次
- PSEO: Optimizing Post-hoc Stacking Ensemble Through Hyperparameter TuningBeicheng Xu, Wei Liu, Keyao Ding, Yupeng Lu 等AAAI 2026 · 被引用 2 次
