Geometry-Aware Gradient Algorithms for Neural Architecture Search
Liam Li, Mikhail Khodak, Nina Balcan, Ameet Talwalkar
摘要
Recent state-of-the-art methods for neural architecture search (NAS) exploit gradient-based optimization by relaxing the problem into continuous optimization over architectures and shared-weights, a noisy process that remains poorly understood. We argue for the study of single-level empirical risk minimization to understand NAS with weight-sharing, reducing the design of NAS methods to devising optimizers and regularizers that can quickly obtain high-quality solutions to this problem. Invoking the theory of mirror descent, we present a geometry-aware framework that exploits the underlying structure of this optimization to return sparse architectural parameters, leading to simple yet novel algorithms that enjoy fast convergence guarantees and achieve state-of-the-art accuracy on the latest NAS benchmarks in computer vision. Notably, we exceed the best published results for both CIFAR and ImageNet on both the DARTS search space and NAS-Bench-201; on the latter we achieve near-oracle-optimal performance on CIFAR-10 and CIFAR-100. Together, our theory and experiments demonstrate a principled way to co-design optimizers and continuous relaxations of discrete NAS search spaces.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper17
- How Powerful are Performance Predictors in Neural Architecture Search?Colin White, Arber Zela, Robin Ru, Yang Liu 等NeurIPS 2021 · 被引用 168 次
- Federated Hyperparameter Tuning: Challenges, Baselines, and Connections to Weight-SharingMikhail Khodak, Renbo Tu, Tian Li, Liam Li 等NeurIPS 2021 · 被引用 111 次
- Efficient Architecture Search for Diverse TasksJunhong Shen, Mikhail Khodak, Ameet TalwalkarNeurIPS 2022 · 被引用 42 次
- NAS-Bench-x11 and the Power of Learning CurvesShen Yan, Colin White, Yash Savani, Frank HutterNeurIPS 2021 · 被引用 36 次
- Generalizing Few-Shot NAS with Gradient MatchingShoukang Hu, Ruochen Wang, Lanqing Hong, Zhenguo Li 等ICLR 2022 · 被引用 29 次
它引用的顶会 Paper9
- NAS-Bench-201: Extending the Scope of Reproducible Neural Architecture SearchXuanyi Dong, Yi YangICLR 2020 · 被引用 825 次
- Progressive Differentiable Architecture Search: Bridging the Depth Gap Between Search and EvaluationXin Chen, Lingxi Xie, Jun Wu, Qi TianICCV 2019 · 被引用 725 次
- PC-DARTS: Partial Channel Connections for Memory-Efficient Architecture SearchYuhui Xu, Lingxi Xie, Xiaopeng Zhang, Xin Chen 等ICLR 2020 · 被引用 691 次
- Understanding and Robustifying Differentiable Architecture SearchArber Zela, Thomas Elsken, Tonmoy Saikia, Yassine Marrakchi 等ICLR 2020 · 被引用 408 次
- Evaluating The Search Phase of Neural Architecture SearchKaicheng Yu, Christian Sciuto, Martin Jaggi, Claudiu Musat 等ICLR 2020 · 被引用 370 次
相关 Paper
- IS-DARTS: Stabilizing DARTS through Precise Measurement on Candidate ImportanceHongyi He, Longjun Liu, Haonan Zhang, Nanning ZhengAAAI 2024 · 被引用 21 次
- -DARTS: Mitigating Performance Collapse by Harmonizing Operation Selection among CellsSajad Movahedi, Melika Adabinejad, Ayyoob Imani, Arezou Keshavarz 等ICLR 2023
- iDARTS: Differentiable Architecture Search with Stochastic Implicit GradientsMiao Zhang, Steven W. Su, Shirui Pan, Xiaojun Chang 等ICML 2021 · 被引用 81 次
- BaLeNAS: Differentiable Architecture Search via the Bayesian Learning RuleMiao Zhang, Shirui Pan, Xiaojun Chang, Steven Su 等CVPR 2022
- MiLeNAS: Efficient Neural Architecture Search via Mixed-Level ReformulationChaoyang He, Haishan Ye, Li Shen, Tong ZhangCVPR 2020
