HyperSTAR: Task-Aware Hyperparameters for Deep Networks
Gaurav Mittal, Chang Liu, Nikolaos Karianakis, Victor Fragoso, Mei Chen, Yun Fu
摘要
While deep neural networks excel in solving visual recognition tasks, they require significant effort to find hyperparameters that make them work optimally. Hyperparameter Optimization (HPO) approaches have automated the process of finding good hyperparameters but they do not adapt to a given task (task-agnostic), making them computationally inefficient. To reduce HPO time, we present Hy-perSTAR (System for Task Aware Hyperparameter Recommendation), a task-aware method to warm-start HPO for deep neural networks. HyperSTAR ranks and recommends hyperparameters by predicting their performance conditioned on a joint dataset-hyperparameter space. It learns a dataset (task) representation along with the performance predictor directly from raw images in an end-to-end fashion. The recommendations, when integrated with an existing HPO method, make it task-aware and significantly reduce the time to achieve optimal performance. We conduct extensive experiments on 10 publicly available large-scale image classification datasets over two different network architectures, validating that HyperSTAR evaluates 50% less configurations to achieve the best performance compared to existing methods. We further demonstrate that HyperSTAR makes Hyperband (HB) task-aware, achieving the optimal accuracy in just 25% of the budget required by both vanilla HB and Bayesian Optimized HB (BOHB).
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper7
- T-AutoML: Automated Machine Learning for Lesion Segmentation using Transformers in 3D Medical ImagingDong Yang, Andriy Myronenko, Xiaosong Wang, Ziyue Xu 等ICCV 2021 · 被引用 28 次
- Learning to Learn across Diverse Data Biases in Deep Face RecognitionChang Liu, Xiang Yu, Yi-Hsuan Tsai, Masoud Faraki 等CVPR 2022 · 被引用 22 次
- Domain Generalization via Feature Variation DecorrelationChang Liu, Lichen Wang, Kai Li, Yun FuACM MM 2021 · 被引用 21 次
- NASOA: Towards Faster Task-oriented Online Fine-tuning with a Zoo of ModelsHang Xu, Ning Kang, Gengwei Zhang, Chuanlong Xie 等ICCV 2021 · 被引用 10 次
- Privacy-preserving Online AutoML for Domain-Specific Face DetectionChenqian Yan, Yuge Zhang, Quanlu Zhang, Yaming Yang 等CVPR 2022 · 被引用 9 次
它引用的顶会 Paper1
相关 Paper
- PriorBand: Practical Hyperparameter Optimization in the Age of Deep LearningNeeratyoy Mallik, Edward Bergman, Carl Hvarfner, Danny Stoll 等NeurIPS 2023 · 被引用 50 次
- LAMDA: Two-Phase HPO via Learning Prior from Low-Fidelity DataFan Li, Shengbo Wang, Ke LiAAAI 2026
- HyperJump: Accelerating HyperBand via Risk ModellingPedro Mendes, Maria Casimiro, Paolo Romano, David GarlanAAAI 2023 · 被引用 11 次
- Neural Architecture and Hyperparameter Selection Through Meta-Learning on Time SeriesErfan Moeini, Christopher Vox, Marie Anastacio, Wadie Skaf 等AAAI 2026 · 被引用 1 次
- AUTOMATA: Gradient Based Data Subset Selection for Compute-Efficient Hyper-parameter TuningKrishnaTeja Killamsetty, Guttu Sai Abhishek, Aakriti, Ganesh Ramakrishnan 等NeurIPS 2022 · 被引用 37 次
