Geometry-Aware Gradient Algorithms for Neural Architecture Search
Liam Li, Mikhail Khodak, Nina Balcan, Ameet Talwalkar
Abstract
Recent state-of-the-art methods for neural architecture search (NAS) exploit gradient-based optimization by relaxing the problem into continuous optimization over architectures and shared-weights, a noisy process that remains poorly understood. We argue for the study of single-level empirical risk minimization to understand NAS with weight-sharing, reducing the design of NAS methods to devising optimizers and regularizers that can quickly obtain high-quality solutions to this problem. Invoking the theory of mirror descent, we present a geometry-aware framework that exploits the underlying structure of this optimization to return sparse architectural parameters, leading to simple yet novel algorithms that enjoy fast convergence guarantees and achieve state-of-the-art accuracy on the latest NAS benchmarks in computer vision. Notably, we exceed the best published results for both CIFAR and ImageNet on both the DARTS search space and NAS-Bench-201; on the latter we achieve near-oracle-optimal performance on CIFAR-10 and CIFAR-100. Together, our theory and experiments demonstrate a principled way to co-design optimizers and continuous relaxations of discrete NAS search spaces.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 1c874c85-3434-415e-8615-597a52c4aaf6Cited by top-tier papers17
- How Powerful are Performance Predictors in Neural Architecture Search?Colin White, Arber Zela, Robin Ru, Yang Liu et al.NeurIPS 2021 · 168 citations
- Federated Hyperparameter Tuning: Challenges, Baselines, and Connections to Weight-SharingMikhail Khodak, Renbo Tu, Tian Li, Liam Li et al.NeurIPS 2021 · 111 citations
- Efficient Architecture Search for Diverse TasksJunhong Shen, Mikhail Khodak, Ameet TalwalkarNeurIPS 2022 · 42 citations
- NAS-Bench-x11 and the Power of Learning CurvesShen Yan, Colin White, Yash Savani, Frank HutterNeurIPS 2021 · 36 citations
- Generalizing Few-Shot NAS with Gradient MatchingShoukang Hu, Ruochen Wang, Lanqing Hong, Zhenguo Li et al.ICLR 2022 · 29 citations
Builds on9
- NAS-Bench-201: Extending the Scope of Reproducible Neural Architecture SearchXuanyi Dong, Yi YangICLR 2020 · 825 citations
- Progressive Differentiable Architecture Search: Bridging the Depth Gap Between Search and EvaluationXin Chen, Lingxi Xie, Jun Wu, Qi TianICCV 2019 · 725 citations
- PC-DARTS: Partial Channel Connections for Memory-Efficient Architecture SearchYuhui Xu, Lingxi Xie, Xiaopeng Zhang, Xin Chen et al.ICLR 2020 · 691 citations
- Understanding and Robustifying Differentiable Architecture SearchArber Zela, Thomas Elsken, Tonmoy Saikia, Yassine Marrakchi et al.ICLR 2020 · 408 citations
- Evaluating The Search Phase of Neural Architecture SearchKaicheng Yu, Christian Sciuto, Martin Jaggi, Claudiu Musat et al.ICLR 2020 · 370 citations
Related papers
- IS-DARTS: Stabilizing DARTS through Precise Measurement on Candidate ImportanceHongyi He, Longjun Liu, Haonan Zhang, Nanning ZhengAAAI 2024 · 21 citations
- -DARTS: Mitigating Performance Collapse by Harmonizing Operation Selection among CellsSajad Movahedi, Melika Adabinejad, Ayyoob Imani, Arezou Keshavarz et al.ICLR 2023
- iDARTS: Differentiable Architecture Search with Stochastic Implicit GradientsMiao Zhang, Steven W. Su, Shirui Pan, Xiaojun Chang et al.ICML 2021 · 81 citations
- BaLeNAS: Differentiable Architecture Search via the Bayesian Learning RuleMiao Zhang, Shirui Pan, Xiaojun Chang, Steven Su et al.CVPR 2022
- MiLeNAS: Efficient Neural Architecture Search via Mixed-Level ReformulationChaoyang He, Haishan Ye, Li Shen, Tong ZhangCVPR 2020
