Multinomial Distribution Learning for Effective Neural Architecture Search
Xiawu Zheng, Rongrong Ji, Lang Tang, Baochang Zhang, Jianzhuang Liu, Qi Tian
Abstract
Architectures obtained by Neural Architecture Search (NAS) have achieved highly competitive performance in various computer vision tasks. However, the prohibitive computation demand of forward-backward propagation in deep neural networks and searching algorithms makes it difficult to apply NAS in practice. In this paper, we propose a Multinomial Distribution Learning for extremely effective NAS, which considers the search space as a joint multinomial distribution, i.e., the operation between two nodes is sampled from this distribution, and the optimal network structure is obtained by the operations with the most likely probability in this distribution. Therefore, NAS can be transformed to a multinomial distribution learning problem, i.e., the distribution is optimized to have high expectation of the performance. Besides, a hypothesis that the performance ranking is consistent in every training epoch is proposed and demonstrated to further accelerate the learning process. Experiments on CIFAR-10 and Im-ageNet demonstrate the effectiveness of our method. On CIFAR-10, the structure searched by our method achieves 2.55% test error, while being 6.0× (only 4 GPU hours on GTX1080Ti) faster compared with state-of-the-art NAS algorithms. On ImageNet, our model achieves 75.2% top-1 accuracy under MobileNet settings (MobileNet V1/V2), while being 1.2× faster with measured GPU latency. Test code with pre-trained models are available at https: //github.com/tanglang96/MDENAS
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext f08d1163-dfc9-4a19-9922-0cc410de566fCited by top-tier papers23
- FairNAS: Rethinking Evaluation Fairness of Weight Sharing Neural Architecture SearchXiangxiang Chu, Bo Zhang, Ruijun XuICCV 2021 · 362 citations
- iDARTS: Differentiable Architecture Search with Stochastic Implicit GradientsMiao Zhang, Steven W. Su, Shirui Pan, Xiaojun Chang et al.ICML 2021 · 81 citations
- Geometry-Aware Gradient Algorithms for Neural Architecture SearchLiam Li, Mikhail Khodak, Nina Balcan, Ameet TalwalkarICLR 2021 · 73 citations
- Adapting Neural Architectures Between DomainsYanxi Li, Zhaohui Yang, Yunhe Wang, Chang XuNeurIPS 2020 · 34 citations
- Binarized Neural Architecture SearchHanlin Chen, Li'an Zhuo, Baochang Zhang, Xiawu Zheng et al.AAAI 2020 · 27 citations
Related papers
- Fast and Practical Neural Architecture SearchJiequan Cui, Pengguang Chen, Ruiyu Li, Shu Liu et al.ICCV 2019 · 69 citations
- Learning Latent Architectural Distribution in Differentiable Neural Architecture Search via Variational Information MaximizationYaoming Wang, Yuchen Liu, Wenrui Dai, Chenglin Li et al.ICCV 2021 · 9 citations
- Rapid Neural Architecture Search by Learning to Generate Graphs from DatasetsHayeon Lee, Eunyoung Hyung, Sung Ju HwangICLR 2021 · 57 citations
- DrNAS: Dirichlet Neural Architecture SearchXiangning Chen, Ruochen Wang, Minhao Cheng, Xiaocheng Tang et al.ICLR 2021 · 7 citations
- Densely Connected Search Space for More Flexible Neural Architecture SearchJiemin Fang, Yuzhu Sun, Qian Zhang, Yuan Li et al.CVPR 2020
