Hyper-Tune: Towards Efficient Hyper-parameter Tuning at Scale
Yang Li, Yu Shen, Huaijun Jiang, Wentao Zhang, Jixiang Li, Ji Liu, Ce Zhang, Bin Cui
摘要
The ever-growing demand and complexity of machine learning are putting pressure on hyper-parameter tuning systems: while the evaluation cost of models continues to increase, the scalability of state-of-the-arts starts to become a crucial bottleneck. In this paper, inspired by our experience when deploying hyper-parameter tuning in a real-world application in production and the limitations of existing systems, we propose Hyper-Tune, an efficient and robust distributed hyper-parameter tuning framework. Compared with existing systems, Hyper-Tune highlights multiple system optimizations, including (1) automatic resource allocation, (2) asynchronous scheduling, and (3) multi-fidelity optimizer. We conduct extensive evaluations on benchmark datasets and a large-scale real-world dataset in production. Empirically, with the aid of these optimizations, Hyper-Tune outperforms competitive hyper-parameter tuning systems on a wide range of scenarios, including XGBoost, CNN, RNN, and some architectural hyper-parameters for neural networks. Compared with the state-of-the-art BOHB and A-BOHB, Hyper-Tune achieves up to 11.2X and 5.1X speedups, respectively.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- Optimizing Hyperparameters with Conformal Quantile RegressionDavid Salinas, Jacek Golebiowski, Aaron Klein, Matthias W. Seeger 等ICML 2023 · 被引用 13 次
- Efficient Hyperparameter Optimization with Adaptive Fidelity IdentificationJiantong Jiang, Zeyi Wen, Atif Bin Mansoor, Ajmal MianCVPR 2024
- GR-Gauge: Cost-efficient Training Configuration By Gauging the Gradient RedundancyGuanjie Wang, Chen ChenCVPR 2026
- MFTune: An Efficient Multi-fidelity Framework for Spark SQL Configuration TuningBeicheng Xu, Lingching Tung, Yuchen Wang, Yupeng Lu 等VLDB 2026
- LAMDA: Two-Phase HPO via Learning Prior from Low-Fidelity DataFan Li, Shengbo Wang, Ke LiAAAI 2026
它引用的顶会 Paper10
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah 等NeurIPS 2020 · 被引用 64,255 次
- NAS-Bench-201: Extending the Scope of Reproducible Neural Architecture SearchXuanyi Dong, Yi YangICLR 2020 · 被引用 825 次
- PC-DARTS: Partial Channel Connections for Memory-Efficient Architecture SearchYuhui Xu, Lingxi Xie, Xiaopeng Zhang, Xin Chen 等ICLR 2020 · 被引用 691 次
- BANANAS: Bayesian Optimization with Neural Architectures for Neural Architecture SearchColin White, Willie Neiswanger, Yash SavaniAAAI 2021 · 被引用 401 次
- ResTune: Resource Oriented Tuning Boosted by Meta-Learning for Cloud DatabasesXinyi Zhang, Hong Wu, Zhuo Chang, Shuowei Jin 等SIGMOD 2021 · 被引用 113 次
相关 Paper
- RubberBand: cloud-based hyperparameter tuningUjval Misra, Richard Liaw, Lisa Dunlap, Romil Bhardwaj 等EuroSys 2021 · 被引用 21 次
- Hydro: Surrogate-Based Hyperparameter Tuning Service in DatacentersQinghao Hu, Zhisheng Ye, Meng Zhang, Qiaoling Chen 等OSDI 2023 · 被引用 16 次
- PASHA: Efficient HPO and NAS with Progressive Resource AllocationOndrej Bohdal, Lukas Balles, Martin Wistuba, Beyza Ermis 等ICLR 2023 · 被引用 4 次
- AutoByte: Automatic Configuration for Optimal Communication Scheduling in DNN TrainingYiqing Ma, Hao Wang, Yiming Zhang, Kai ChenINFOCOM 2022 · 被引用 11 次
- Resource-Guided Configuration Space Reduction for Deep Learning ModelsYanjie Gao, Yonghao Zhu, Hongyu Zhang, Haoxiang Lin 等ICSE 2021 · 被引用 17 次
