PASHA: Efficient HPO and NAS with Progressive Resource Allocation
Ondrej Bohdal, Lukas Balles, Martin Wistuba, Beyza Ermis, Cédric Archambeau, Giovanni Zappella
摘要
Hyperparameter optimization (HPO) and neural architecture search (NAS) are methods of choice to obtain the best-in-class machine learning models, but in practice they can be costly to run. When models are trained on large datasets, tuning them with HPO or NAS rapidly becomes prohibitively expensive for practitioners, even when efficient multi-fidelity methods are employed. We propose an approach to tackle the challenge of tuning machine learning models trained on large datasets with limited computational resources. Our approach, named PASHA, extends ASHA and is able to dynamically allocate maximum resources for the tuning procedure depending on the need. The experimental comparison shows that PASHA identifies well-performing hyperparameter configurations and architectures while consuming significantly fewer computational resources than ASHA.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- FairTune: Optimizing Parameter Efficient Fine Tuning for Fairness in Medical Image AnalysisRaman Dutt, Ondrej Bohdal, Sotirios A. Tsaftaris, Timothy M. HospedalesICLR 2024 · 被引用 30 次
- pTNAS: Progressive Neural Architecture Search for Tabular DataNaili Xing, Shaofeng Cai, Lingze Zeng, Jiaqi Zhu 等ICML 2026 · 被引用 4 次
- Enhancing the Performance of Bandit-based Hyperparameter OptimizationYile Chen, Zeyi Wen, Jian Chen, Jin HuangICDE 2024 · 被引用 3 次
- Meta Omnium: A Benchmark for General-Purpose Learning-to-LearnOndrej Bohdal, Yinbing Tian, Yongshuo Zong, Ruchika Chavhan 等CVPR 2023
- Efficient Hyperparameter Optimization with Adaptive Fidelity IdentificationJiantong Jiang, Zeyi Wen, Atif Bin Mansoor, Ajmal MianCVPR 2024
它引用的顶会 Paper4
- NAS-Bench-201: Extending the Scope of Reproducible Neural Architecture SearchXuanyi Dong, Yi YangICLR 2020 · 被引用 825 次
- PC-DARTS: Partial Channel Connections for Memory-Efficient Architecture SearchYuhui Xu, Lingxi Xie, Xiaopeng Zhang, Xin Chen 等ICLR 2020 · 被引用 691 次
- Small Data, Big Decisions: Model Selection in the Small-Data RegimeJörg Bornschein, Francesco Visin, Simon OsinderoICML 2020 · 被引用 48 次
- EcoNAS: Finding Proxies for Economical Neural Architecture SearchDongzhan Zhou, Xinchi Zhou, Wenwei Zhang, Chen Change Loy 等CVPR 2020
相关 Paper
- Supervising the Multi-Fidelity Race of Hyperparameter ConfigurationsMartin Wistuba, Arlind Kadra, Josif GrabockaNeurIPS 2022 · 被引用 24 次
- RubberBand: cloud-based hyperparameter tuningUjval Misra, Richard Liaw, Lisa Dunlap, Romil Bhardwaj 等EuroSys 2021 · 被引用 21 次
- Optimizer Benchmarking Needs to Account for Hyperparameter TuningPrabhu Teja Sivaprasad, Florian Mai, Thijs Vogels, Martin Jaggi 等ICML 2020 · 被引用 60 次
- Hyper-Tune: Towards Efficient Hyper-parameter Tuning at ScaleYang Li, Yu Shen, Huaijun Jiang, Wentao Zhang 等VLDB 2022 · 被引用 32 次
- AUTOMATA: Gradient Based Data Subset Selection for Compute-Efficient Hyper-parameter TuningKrishnaTeja Killamsetty, Guttu Sai Abhishek, Aakriti, Ganesh Ramakrishnan 等NeurIPS 2022 · 被引用 37 次
