CASH via Optimal Diversity for Ensemble Learning
Pranav Poduval, Sanjay Kumar Patnala, Gaurav Oberoi, Nitish Srivasatava, Siddhartha Asthana
Abstract
The Combined Algorithm Selection and Hyperparameter Optimization (CASH) problem is pivotal in Automatic Machine Learning (AutoML). Most leading approaches combine Bayesian optimization with post-hoc ensemble building to create advanced AutoML systems. Bayesian optimization (BO) typically focuses on identifying a singular algorithm and its hyperparameters that outperform all other configurations. Recent developments have highlighted an oversight in prior CASH methods: the lack of consideration for diversity among the base learners of the ensemble. This oversight was overcome by explicitly injecting the search for diversity into the traditional CASH problem. However, despite recent developments, BO's limitation lies in its inability to directly optimize ensemble generalization error, offering no theoretical assurance that increased diversity correlates with enhanced ensemble performance. Our research addresses this gap by establishing a theoretical foundation that integrates diversity into the core of BO for direct ensemble learning. We explore a theoretically sound framework that describes the relationship between pair-wise diversity and ensemble performance, which allows our Bayesian optimization framework Optimal Diversity Bayesian Optimization (OptDivBO) to directly and efficiently minimize ensemble generalization error. OptDivBO guarantees an optimal balance between pairwise diversity and individual model performance, setting a new precedent in ensemble learning within CASH. Empirical results on 20 public datasets show that OptDivBO achieves the best average test ranks of 1.57 and 1.4 in classification and regression tasks.
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Your agent calls
Lunesearch_papers
Free to start. No credit card required.
Terminal
Install the CLIlune papers get 35ab53b1-2f6c-42f8-84ef-ae0b573bd44eCited by top-tier papers2
- PSEO: Optimizing Post-hoc Stacking Ensemble Through Hyperparameter TuningBeicheng Xu, Wei Liu, Keyao Ding, Yupeng Lu et al.AAAI 2026 · 2 citations
- CoFEH: LLM-driven Feature Engineering Empowered by Collaborative Bayesian Hyperparameter OptimizationBeicheng Xu, Keyao Ding, Wei Liu, Yupeng Lu et al.KDD 2026
Related papers
- DivBO: Diversity-aware CASH for Ensemble LearningYu Shen, Yupeng Lu, Yang Li, Yaofeng Tu et al.NeurIPS 2022 · 15 citations
- Efficient Automatic CASH via Rising BanditsYang Li, Jiawei Jiang, Jinyang Gao, Yingxia Shao et al.AAAI 2020 · 45 citations
- Put CASH on Bandits: A Max K-Armed Problem for Automated Machine LearningAmir Rezaei Balef, Claire Vernade, Katharina EggenspergerNeurIPS 2025 · 4 citations
- Weighted Sampling for Combined Model Selection and Hyperparameter TuningDimitrios Sarigiannis, Thomas P. Parnell, Haralampos PozidisAAAI 2020 · 3 citations
- Joint Training of Deep Ensembles Fails Due to Learner CollusionAlan Jeffares, Tennison Liu, Jonathan Crabbé, Mihaela van der SchaarNeurIPS 2023 · 34 citations
