ZARTS: On Zero-order Optimization for Neural Architecture Search
Xiaoxing Wang, Wenxuan Guo, Jianlin Su, Xiaokang Yang, Junchi Yan
Abstract
Differentiable architecture search (DARTS) has been a popular one-shot paradigm for NAS due to its high efficiency. It introduces trainable architecture parameters to represent the importance of candidate operations and proposes first/secondorder approximation to estimate their gradients, making it possible to solve NAS by gradient descent algorithm. However, our in-depth empirical results show that the approximation often distorts the loss landscape, leading to the biased objective to optimize and, in turn, inaccurate gradient estimation for architecture parameters. This work turns to zero-order optimization and proposes a novel NAS scheme, called ZARTS, to search without enforcing the above approximation. Specifically, three representative zero-order optimization methods are introduced: RS, MGS, and GLD, among which MGS performs best by balancing the accuracy and speed. Moreover, we explore the connections between RS/MGS and gradient descent algorithm and show that our ZARTS can be seen as a robust gradient-free counterpart to DARTS. Extensive experiments on multiple datasets and search spaces show the remarkable performance of our method. In particular, results on 12 benchmarks verify the outstanding robustness of ZARTS, where the performance of DARTS collapses due to its known instability issue. Also, we search on the search space of DARTS to compare with peer methods, and our discovered architecture achieves 97.54% accuracy on CIFAR-10 and 75.7% top-1 accuracy on Im-ageNet. Finally, we combine our ZARTS with three orthogonal variants of DARTS for faster search speed and better performance.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 27b9deb2-c6ac-4964-bb79-92dd6636855fCited by top-tier papers7
- MeCo: Zero-Shot NAS with One Data and Single Forward Pass via Minimum Eigenvalue of CorrelationTangyu Jiang, Haodi Wang, Rongfang BieNeurIPS 2023 · 32 citations
- PROTES: Probabilistic Optimization with Tensor SamplingAnastasia Batsheva, Andrei Chertkov, Gleb V. Ryzhakov, Ivan V. OseledetsNeurIPS 2023 · 27 citations
- Towards Efficient Low-Order Hybrid Optimizer for Language Model Fine-TuningMinping Chen, You-Liang Huang, Zeyi WenAAAI 2025 · 6 citations
- ROME: Robustifying Memory-Efficient NAS via Topology Disentanglement and Gradient AccumulationXiaoxing Wang, Xiangxiang Chu, Yuda Fan, Zhexi Zhang et al.ICCV 2023 · 5 citations
- MobiZO: Enabling Efficient LLM Fine-Tuning at the Edge via Inference EnginesLei Gao, Amir Ziashahabi, Yue Niu, Salman Avestimehr et al.EMNLP 2025 · 1 citation
Builds on14
- Progressive Differentiable Architecture Search: Bridging the Depth Gap Between Search and EvaluationXin Chen, Lingxi Xie, Jun Wu, Qi TianICCV 2019 · 725 citations
- PC-DARTS: Partial Channel Connections for Memory-Efficient Architecture SearchYuhui Xu, Lingxi Xie, Xiaopeng Zhang, Xin Chen et al.ICLR 2020 · 691 citations
- Understanding and Robustifying Differentiable Architecture SearchArber Zela, Thomas Elsken, Tonmoy Saikia, Yassine Marrakchi et al.ICLR 2020 · 408 citations
- Evaluating The Search Phase of Neural Architecture SearchKaicheng Yu, Christian Sciuto, Martin Jaggi, Claudiu Musat et al.ICLR 2020 · 370 citations
- Stabilizing Differentiable Architecture Search via Perturbation-based RegularizationXiangning Chen, Cho-Jui HsiehICML 2020 · 235 citations
Related papers
- Zero-Cost Operation Scoring in Differentiable Architecture SearchLichuan Xiang, Lukasz Dudziak, Mohamed S. Abdelfattah, Thomas C. P. Chau et al.AAAI 2023 · 13 citations
- iDARTS: Differentiable Architecture Search with Stochastic Implicit GradientsMiao Zhang, Steven W. Su, Shirui Pan, Xiaojun Chang et al.ICML 2021 · 81 citations
- Differentiable Architecture Search with Random FeaturesXuanyang Zhang, Yonggang Li, Xiangyu Zhang, Yongtao Wang et al.CVPR 2023
- Rethinking Bi-Level Optimization in Neural Architecture Search: A Gibbs Sampling PerspectiveChao Xue, Xiaoxing Wang, Junchi Yan, Yonggang Hu et al.AAAI 2021 · 33 citations
- MiLeNAS: Efficient Neural Architecture Search via Mixed-Level ReformulationChaoyang He, Haishan Ye, Li Shen, Tong ZhangCVPR 2020
