HyperSTAR: Task-Aware Hyperparameters for Deep Networks
Gaurav Mittal, Chang Liu, Nikolaos Karianakis, Victor Fragoso, Mei Chen, Yun Fu
Abstract
While deep neural networks excel in solving visual recognition tasks, they require significant effort to find hyperparameters that make them work optimally. Hyperparameter Optimization (HPO) approaches have automated the process of finding good hyperparameters but they do not adapt to a given task (task-agnostic), making them computationally inefficient. To reduce HPO time, we present Hy-perSTAR (System for Task Aware Hyperparameter Recommendation), a task-aware method to warm-start HPO for deep neural networks. HyperSTAR ranks and recommends hyperparameters by predicting their performance conditioned on a joint dataset-hyperparameter space. It learns a dataset (task) representation along with the performance predictor directly from raw images in an end-to-end fashion. The recommendations, when integrated with an existing HPO method, make it task-aware and significantly reduce the time to achieve optimal performance. We conduct extensive experiments on 10 publicly available large-scale image classification datasets over two different network architectures, validating that HyperSTAR evaluates 50% less configurations to achieve the best performance compared to existing methods. We further demonstrate that HyperSTAR makes Hyperband (HB) task-aware, achieving the optimal accuracy in just 25% of the budget required by both vanilla HB and Bayesian Optimized HB (BOHB).
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext c745d5b0-06b4-4dcc-be80-750ee9aff1a9Cited by top-tier papers7
- T-AutoML: Automated Machine Learning for Lesion Segmentation using Transformers in 3D Medical ImagingDong Yang, Andriy Myronenko, Xiaosong Wang, Ziyue Xu et al.ICCV 2021 · 28 citations
- Learning to Learn across Diverse Data Biases in Deep Face RecognitionChang Liu, Xiang Yu, Yi-Hsuan Tsai, Masoud Faraki et al.CVPR 2022 · 22 citations
- Domain Generalization via Feature Variation DecorrelationChang Liu, Lichen Wang, Kai Li, Yun FuACM MM 2021 · 21 citations
- NASOA: Towards Faster Task-oriented Online Fine-tuning with a Zoo of ModelsHang Xu, Ning Kang, Gengwei Zhang, Chuanlong Xie et al.ICCV 2021 · 10 citations
- Privacy-preserving Online AutoML for Domain-Specific Face DetectionChenqian Yan, Yuge Zhang, Quanlu Zhang, Yaming Yang et al.CVPR 2022 · 9 citations
Builds on1
Related papers
- PriorBand: Practical Hyperparameter Optimization in the Age of Deep LearningNeeratyoy Mallik, Edward Bergman, Carl Hvarfner, Danny Stoll et al.NeurIPS 2023 · 50 citations
- LAMDA: Two-Phase HPO via Learning Prior from Low-Fidelity DataFan Li, Shengbo Wang, Ke LiAAAI 2026
- HyperJump: Accelerating HyperBand via Risk ModellingPedro Mendes, Maria Casimiro, Paolo Romano, David GarlanAAAI 2023 · 11 citations
- Neural Architecture and Hyperparameter Selection Through Meta-Learning on Time SeriesErfan Moeini, Christopher Vox, Marie Anastacio, Wadie Skaf et al.AAAI 2026 · 1 citation
- AUTOMATA: Gradient Based Data Subset Selection for Compute-Efficient Hyper-parameter TuningKrishnaTeja Killamsetty, Guttu Sai Abhishek, Aakriti, Ganesh Ramakrishnan et al.NeurIPS 2022 · 37 citations
