K-shot NAS: Learnable Weight-Sharing for NAS with K-shot Supernets
Xiu Su, Shan You, Mingkai Zheng, Fei Wang, Chen Qian, Changshui Zhang, Chang Xu
摘要
In one-shot weight sharing for NAS, the weights of each operation (at each layer) are supposed to be identical for all architectures (paths) in the supernet. However, this rules out the possibility of adjusting operation weights to cater for different paths, which limits the reliability of the evaluation results. In this paper, instead of counting on a single supernet, we introduce -shot supernets and take their weights for each operation as a dictionary. The operation weight for each path is represented as a convex combination of items in a dictionary with a simplex code. This enables a matrix approximation of the stand-alone weight matrix with a higher rank (). A simplex-net is introduced to produce architecture-customized code for each path. As a result, all paths can adaptively learn how to share weights in the -shot supernets and acquire corresponding weights for better evaluation. -shot supernets and simplex-net can be iteratively trained, and we further extend the search to the channel dimension. Extensive experiments on benchmark datasets validate that K-shot NAS significantly improves the evaluation accuracy of paths and thus brings in impressive performance improvements.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper17
- Weakly Supervised Contrastive LearningMingkai Zheng, Fei Wang, Shan You, Chen Qian 等ICCV 2021 · 被引用 153 次
- On Redundancy and Diversity in Cell-based Neural Architecture SearchXingchen Wan, Binxin Ru, Pedro M. Esperança, Zhenguo LiICLR 2022 · 被引用 27 次
- GreedyNASv2: Greedier Search with a Greedy Path FilterTao Huang, Shan You, Fei Wang, Chen Qian 等CVPR 2022 · 被引用 16 次
- NAS-LID: Efficient Neural Architecture Search with Local Intrinsic DimensionXin He, Jiangchao Yao, Yuxin Wang, Zhenheng Tang 等AAAI 2023 · 被引用 16 次
- Do Not Train It: A Linear Neural Architecture Search of Graph Neural NetworksPeng Xu, Lin Zhang, Xuanzhou Liu, Jiaqi Sun 等ICML 2023 · 被引用 14 次
它引用的顶会 Paper9
- NAS-Bench-201: Extending the Scope of Reproducible Neural Architecture SearchXuanyi Dong, Yi YangICLR 2020 · 被引用 825 次
- PC-DARTS: Partial Channel Connections for Memory-Efficient Architecture SearchYuhui Xu, Lingxi Xie, Xiaopeng Zhang, Xin Chen 等ICLR 2020 · 被引用 691 次
- Agree to Disagree: Adaptive Ensemble Knowledge Distillation in Gradient SpaceShangchen Du, Shan You, Xiaojie Li, Jianlong Wu 等NeurIPS 2020 · 被引用 144 次
- DARTS-: Robustly Stepping out of Performance Collapse Without IndicatorsXiangxiang Chu, Xiaoxing Wang, Bo Zhang, Shun Lu 等ICLR 2021 · 被引用 72 次
- ISTA-NAS: Efficient and Consistent Neural Architecture Search by Sparse CodingYibo Yang, Hongyang Li, Shan You, Fei Wang 等NeurIPS 2020 · 被引用 66 次
相关 Paper
- Few-Shot Neural Architecture SearchYiyang Zhao, Linnan Wang, Yuandong Tian, Rodrigo Fonseca 等ICML 2021 · 被引用 100 次
- Distribution Consistent Neural Architecture SearchJunyi Pan, Chong Sun, Yizhou Zhou, Ying Zhang 等CVPR 2022 · 被引用 9 次
- SUMNAS: Supernet with Unbiased Meta-Features for Neural Architecture SearchHyeonmin Ha, Ji-Hoon Kim, Semin Park, Byung-Gon ChunICLR 2022 · 被引用 5 次
- PA&DA: Jointly Sampling PAth and DAta for Consistent NASShun Lu, Yu Hu, Longxing Yang, Zihao Sun 等CVPR 2023
- ShiftNAS: Improving One-shot NAS via Probability ShiftMingyang Zhang, Xinyi Yu, Haodong Zhao, Linlin OuICCV 2023 · 被引用 9 次
