Towards Improving the Consistency, Efficiency, and Flexibility of Differentiable Neural Architecture Search
Yibo Yang, Shan You, Hongyang Li, Fei Wang, Chen Qian, Zhouchen Lin
Abstract
Most differentiable neural architecture search methods construct a super-net for search and derive a target-net as its sub-graph for evaluation. There exists a significant gap between the architectures in search and evaluation. As a result, current methods suffer from an inconsistent, inefficient, and inflexible search process. In this paper, we introduce EnTranNAS that is composed of Engine-cells and Transit-cells. The Engine-cell is differentiable for architecture search, while the Transit-cell only transits a sub-graph by architecture derivation. Consequently, the gap between the architectures in search and evaluation is significantly reduced. Our method also spares much memory and computation cost, which speeds up the search process. A feature sharing strategy is introduced for more balanced optimization and more efficient search. Furthermore, we develop an architecture derivation method to replace the traditional one that is based on a hand-crafted rule. Our method enables differentiable sparsification, and keeps the derived architecture equivalent to that of Engine-cell, which further improves the consistency between search and evaluation. More importantly, it supports the search for topology where a node can be connected to prior nodes with any number of connections, so that the searched architectures could be more flexible. Our search on CIFAR-10 has an error rate of 2.22% with only 0.07 GPU-day. We can also directly perform the search on ImageNet with topology learnable and achieve a top-1 error rate of 23.8% in 2.1 GPU-day.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext b3c5e732-5e59-4c02-9654-f19beb4d63c3Cited by top-tier papers11
- Locally Free Weight Sharing for Network Width SearchXiu Su, Shan You, Tao Huang, Fei Wang et al.ICLR 2021 · 45 citations
- K-shot NAS: Learnable Weight-Sharing for NAS with K-shot SupernetsXiu Su, Shan You, Mingkai Zheng, Fei Wang et al.ICML 2021 · 38 citations
- Analyzing and Mitigating Interference in Neural Architecture SearchJin Xu, Xu Tan, Kaitao Song, Renqian Luo et al.ICML 2022 · 30 citations
- ShiftAddNAS: Hardware-Inspired Search for More Accurate and Efficient Neural NetworksHaoran You, Baopu Li, Huihong Shi, Yonggan Fu et al.ICML 2022 · 20 citations
- ZiCo: Zero-shot NAS via inverse Coefficient of Variation on GradientsGuihong Li, Yuedong Yang, Kartikeya Bhardwaj, Radu MarculescuICLR 2023 · 19 citations
Builds on9
- Once-for-All: Train One Network and Specialize it for Efficient DeploymentHan Cai, Chuang Gan, Tianzhe Wang, Zhekai Zhang et al.ICLR 2020 · 1,522 citations
- Progressive Differentiable Architecture Search: Bridging the Depth Gap Between Search and EvaluationXin Chen, Lingxi Xie, Jun Wu, Qi TianICCV 2019 · 725 citations
- Evaluating The Search Phase of Neural Architecture SearchKaicheng Yu, Christian Sciuto, Martin Jaggi, Claudiu Musat et al.ICLR 2020 · 370 citations
- AtomNAS: Fine-Grained End-to-End Neural Architecture SearchJieru Mei, Yingwei Li, Xiaochen Lian, Xiaojie Jin et al.ICLR 2020 · 110 citations
- Efficient Neural Architecture Search via Proximal IterationsQuanming Yao, Ju Xu, Wei-Wei Tu, Zhanxing ZhuAAAI 2020 · 108 citations
Related papers
- ISTA-NAS: Efficient and Consistent Neural Architecture Search by Sparse CodingYibo Yang, Hongyang Li, Shan You, Fei Wang et al.NeurIPS 2020 · 66 citations
- Unchain the Search Space with Hierarchical Differentiable Architecture SearchGuanting Liu, Yujie Zhong, Sheng Guo, Matthew R. Scott et al.AAAI 2021 · 3 citations
- DrNAS: Dirichlet Neural Architecture SearchXiangning Chen, Ruochen Wang, Minhao Cheng, Xiaocheng Tang et al.ICLR 2021 · 7 citations
- DOTS: Decoupling Operation and Topology in Differentiable Architecture SearchYuchao Gu, Lijuan Wang, Yun Liu, Yi Yang et al.CVPR 2021
- Block-Wisely Supervised Neural Architecture Search With Knowledge DistillationChanglin Li, Jiefeng Peng, Liuchun Yuan, Guangrun Wang et al.CVPR 2020
