NAS-Bench-x11 and the Power of Learning Curves
Shen Yan, Colin White, Yash Savani, Frank Hutter
Abstract
While early research in neural architecture search (NAS) required extreme computational resources, the recent releases of tabular and surrogate benchmarks have greatly increased the speed and reproducibility of NAS research. However, two of the most popular benchmarks do not provide the full training information for each architecture. As a result, on these benchmarks it is not possible to run many types of multi-fidelity techniques, such as learning curve extrapolation, that require evaluating architectures at arbitrary epochs. In this work, we present a method using singular value decomposition and noise modeling to create surrogate benchmarks, NAS-Bench-111, NAS-Bench-311, and NAS-Bench-NLP11, that output the full training information for each architecture, rather than just the final validation accuracy. We demonstrate the power of using the full training information by introducing a learning curve extrapolation framework to modify single-fidelity algorithms, showing that it leads to improvements over popular single-fidelity algorithms which claimed to be state-of-the-art upon release. Our code and pretrained models are available at https://github.com/automl/nas-bench-x11 .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext f9e78ac9-6a44-4d28-be87-093225f39290Cited by top-tier papers7
- NAS-Bench-Suite: NAS Evaluation is (Now) Surprisingly EasyYash Mehta, Colin White, Arber Zela, Arjun Krishnakumar et al.ICLR 2022 · 54 citations
- On Redundancy and Diversity in Cell-based Neural Architecture SearchXingchen Wan, Binxin Ru, Pedro M. Esperança, Zhenguo LiICLR 2022 · 27 citations
- DiffusionNAG: Predictor-guided Neural Architecture Generation with Diffusion ModelsSohyun An, Hayeon Lee, Jaehyeong Jo, Seanie Lee et al.ICLR 2024 · 21 citations
- TA-GATES: An Encoding Scheme for Neural Network ArchitecturesXuefei Ning, Zixuan Zhou, Junbo Zhao, Tianchen Zhao et al.NeurIPS 2022 · 19 citations
- Database Native Model Selection: Harnessing Deep Neural Networks in Database SystemsNaili Xing, Shaofeng Cai, Gang Chen, Zhaojing Luo et al.VLDB 2024 · 14 citations
Builds on18
- NAS-Bench-201: Extending the Scope of Reproducible Neural Architecture SearchXuanyi Dong, Yi YangICLR 2020 · 825 citations
- PC-DARTS: Partial Channel Connections for Memory-Efficient Architecture SearchYuhui Xu, Lingxi Xie, Xiaopeng Zhang, Xin Chen et al.ICLR 2020 · 691 citations
- Understanding and Robustifying Differentiable Architecture SearchArber Zela, Thomas Elsken, Tonmoy Saikia, Yassine Marrakchi et al.ICLR 2020 · 408 citations
- BANANAS: Bayesian Optimization with Neural Architectures for Neural Architecture SearchColin White, Willie Neiswanger, Yash SavaniAAAI 2021 · 401 citations
- Evaluating The Search Phase of Neural Architecture SearchKaicheng Yu, Christian Sciuto, Martin Jaggi, Claudiu Musat et al.ICLR 2020 · 370 citations
Related papers
- Surrogate NAS Benchmarks: Going Beyond the Limited Search Spaces of Tabular NAS BenchmarksArber Zela, Julien Niklas Siems, Lucas Zimmer, Jovita Lukasik et al.ICLR 2022 · 100 citations
- How Powerful are Performance Predictors in Neural Architecture Search?Colin White, Arber Zela, Robin Ru, Yang Liu et al.NeurIPS 2021 · 168 citations
- Efficient Hyperparameter Optimization with Adaptive Fidelity IdentificationJiantong Jiang, Zeyi Wen, Atif Bin Mansoor, Ajmal MianCVPR 2024
- Accel-NASBench: Sustainable Benchmarking for Accelerator-Aware NASAfzal Ahmad, Linfeng Du, Zhiyao Xie, Wei ZhangDAC 2024 · 1 citation
- EA-HAS-Bench: Energy-aware Hyperparameter and Architecture Search BenchmarkShuguang Dou, Xinyang Jiang, Cairong Zhao, Dongsheng LiICLR 2023
