ISTA-NAS: Efficient and Consistent Neural Architecture Search by Sparse Coding
Yibo Yang, Hongyang Li, Shan You, Fei Wang, Chen Qian, Zhouchen Lin
摘要
Neural architecture search (NAS) aims to produce the optimal sparse solution from a high-dimensional space spanned by all candidate connections. Current gradientbased NAS methods commonly ignore the constraint of sparsity in the search phase, but project the optimized solution onto a sparse one by post-processing. As a result, the dense super-net for search is inefficient to train and has a gap with the projected architecture for evaluation. In this paper, we formulate neural architecture search as a sparse coding problem. We perform the differentiable search on a compressed lower-dimensional space that has the same validation loss as the original sparse solution space, and recover an architecture by solving the sparse coding problem. The differentiable search and architecture recovery are optimized in an alternate manner. By doing so, our network for search at each update satisfies the sparsity constraint and is efficient to train. In order to also eliminate the depth and width gap between the network in search and the target-net in evaluation, we further propose a method to search and evaluate in one stage under the target-net settings. When training finishes, architecture variables are absorbed into network weights. Thus we get the searched architecture and optimized parameters in a single run. In experiments, our two-stage method on CIFAR-10 requires only 0.05 GPU-day for search. Our one-stage method produces state-of-the-art performances on both CIFAR-10 and ImageNet at the cost of only evaluation time 1 .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper27
- Agree to Disagree: Adaptive Ensemble Knowledge Distillation in Gradient SpaceShangchen Du, Shan You, Xiaojie Li, Jianlong Wu 等NeurIPS 2020 · 被引用 144 次
- NAS-OoD: Neural Architecture Search for Out-of-Distribution GeneralizationHaoyue Bai, Fengwei Zhou, Lanqing Hong, Nanyang Ye 等ICCV 2021 · 被引用 46 次
- Efficient Equivariant NetworkLingshen He, Yuxuan Chen, Zhengyang Shen, Yiming Dong 等NeurIPS 2021 · 被引用 46 次
- Locally Free Weight Sharing for Network Width SearchXiu Su, Shan You, Tao Huang, Fei Wang 等ICLR 2021 · 被引用 45 次
- K-shot NAS: Learnable Weight-Sharing for NAS with K-shot SupernetsXiu Su, Shan You, Mingkai Zheng, Fei Wang 等ICML 2021 · 被引用 38 次
它引用的顶会 Paper9
- Once-for-All: Train One Network and Specialize it for Efficient DeploymentHan Cai, Chuang Gan, Tianzhe Wang, Zhekai Zhang 等ICLR 2020 · 被引用 1,522 次
- Progressive Differentiable Architecture Search: Bridging the Depth Gap Between Search and EvaluationXin Chen, Lingxi Xie, Jun Wu, Qi TianICCV 2019 · 被引用 725 次
- NAS evaluation is frustratingly hardAntoine Yang, Pedro M. Esperança, Fabio Maria CarlucciICLR 2020 · 被引用 180 次
- AtomNAS: Fine-Grained End-to-End Neural Architecture SearchJieru Mei, Yingwei Li, Xiaochen Lian, Xiaojie Jin 等ICLR 2020 · 被引用 110 次
- Efficient Neural Architecture Search via Proximal IterationsQuanming Yao, Ju Xu, Wei-Wei Tu, Zhanxing ZhuAAAI 2020 · 被引用 108 次
相关 Paper
- Towards Improving the Consistency, Efficiency, and Flexibility of Differentiable Neural Architecture SearchYibo Yang, Shan You, Hongyang Li, Fei Wang 等CVPR 2021
- Unchain the Search Space with Hierarchical Differentiable Architecture SearchGuanting Liu, Yujie Zhong, Sheng Guo, Matthew R. Scott 等AAAI 2021 · 被引用 3 次
- UNAS: Differentiable Architecture Search Meets Reinforcement LearningArash Vahdat, Arun Mallya, Ming-Yu Liu, Jan KautzCVPR 2020
- AutoSpace: Neural Architecture Search with Less Human InterferenceDaquan Zhou, Xiaojie Jin, Xiaochen Lian, Linjie Yang 等ICCV 2021 · 被引用 11 次
- You only search once: on lightweight differentiable architecture search for resource-constrained embedded platformsXiangzhong Luo, Di Liu, Hao Kong, Shuo Huai 等DAC 2022 · 被引用 13 次
