Learning to Rank Learning Curves
Martin Wistuba, Tejaswini Pedapati
Abstract
Many automated machine learning methods, such as those for hyperparameter and neural architecture optimization, are computationally expensive because they involve training many different model configurations. In this work, we present a new method that saves computational budget by terminating poor configurations early on in the training. In contrast to existing methods, we consider this task as a ranking and transfer learning problem. We qualitatively show that by optimizing a pairwise ranking loss and leveraging learning curves from other datasets, our model is able to effectively rank learning curves without having to observe many or very long learning curves. We further demonstrate that our method can be used to accelerate a neural architecture search by a factor of up to 100 without a significant performance degradation of the discovered architecture. In further experiments we analyze the quality of ranking, the influence of different model components as well as the predictive behavior of the model.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 869b834d-146a-4e7c-90d4-8d7123bea002Cited by top-tier papers8
- T-AutoML: Automated Machine Learning for Lesion Segmentation using Transformers in 3D Medical ImagingDong Yang, Andriy Myronenko, Xiaosong Wang, Ziyue Xu et al.ICCV 2021 · 28 citations
- Supervising the Multi-Fidelity Race of Hyperparameter ConfigurationsMartin Wistuba, Arlind Kadra, Josif GrabockaNeurIPS 2022 · 24 citations
- Scaling Laws for Hyperparameter OptimizationArlind Kadra, Maciej Janowski, Martin Wistuba, Josif GrabockaNeurIPS 2023 · 23 citations
- Optimizing Hyperparameters with Conformal Quantile RegressionDavid Salinas, Jacek Golebiowski, Aaron Klein, Matthias W. Seeger et al.ICML 2023 · 13 citations
- Private Rank Aggregation in Central and Local ModelsDaniel Alabi, Badih Ghazi, Ravi Kumar, Pasin ManurangsiAAAI 2022 · 12 citations
Builds on1
Related papers
- Architecture-Aware Learning Curve Extrapolation via Graph Ordinary Differential EquationYanna Ding, Zijie Huang, Xiao Shou, Yihang Guo et al.AAAI 2025 · 4 citations
- RankNAS: Efficient Neural Architecture Search by Pairwise RankingChi Hu, Chenglong Wang, Xiangnan Ma, Xia Meng et al.EMNLP 2021 · 10 citations
- RANK-NOSH: Efficient Predictor-Based Architecture Search via Non-Uniform Successive HalvingRuochen Wang, Xiangning Chen, Minhao Cheng, Xiaocheng Tang et al.ICCV 2021 · 14 citations
- Transfer Learning based Search Space Design for Hyperparameter TuningYang Li, Yu Shen, Huaijun Jiang, Tianyi Bai et al.KDD 2022 · 8 citations
- PASHA: Efficient HPO and NAS with Progressive Resource AllocationOndrej Bohdal, Lukas Balles, Martin Wistuba, Beyza Ermis et al.ICLR 2023 · 4 citations
