Overcoming Multi-Model Forgetting in One-Shot NAS With Diversity Maximization
Miao Zhang, Huiqi Li, Shirui Pan, Xiaojun Chang, Steven W. Su
Abstract
One-Shot Neural Architecture Search (NAS) significantly improves the computational efficiency through weight sharing. However, this approach also introduces multi-model forgetting during the supernet training (architecture search phase), where the performance of previous architectures degrades when sequentially training new architectures with partially-shared weights. To overcome such catastrophic forgetting, the state-of-the-art method assumes that the shared weights are optimal when jointly optimizing a posterior probability. However, this strict assumption is not necessarily held for One-Shot NAS in practice. In this paper, we formulate the supernet training in the One-Shot NAS as a constrained optimization problem of continual learning that the learning of current architecture should not degrade the performance of previous architectures. We propose a Novelty Search based Architecture Selection (NSAS) loss function and demonstrate that the posterior probability could be calculated without the strict assumption when maximizing the diversity of the selected constraints. A greedy novelty search method is devised to find the most representative subset to regularize the supernet training. We apply our proposed approach to two One-Shot NAS baselines, random sampling NAS (RandomNAS) and gradient-based sampling NAS (GDAS). Extensive experiments demonstrate that our method enhances the predictive ability of the supernet in One-Shot NAS and achieves remarkable performance on CIFAR-10, CIFAR-100, and PTB with efficiency.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 0ca237cb-960c-4a6b-93ad-4da20cd27c8bCited by top-tier papers22
- Zen-NAS: A Zero-Shot NAS for High-Performance Image RecognitionMing Lin, Pichao Wang, Zhenhong Sun, Hesen Chen et al.ICCV 2021 · 164 citations
- BossNAS: Exploring Hybrid CNN-transformers with Block-wisely Self-supervised Neural Architecture SearchChanglin Li, Tao Tang, Guangrun Wang, Jiefeng Peng et al.ICCV 2021 · 123 citations
- Evaluating Efficient Performance Estimators of Neural ArchitecturesXuefei Ning, Changcheng Tang, Wenshuo Li, Zixuan Zhou et al.NeurIPS 2021 · 99 citations
- iDARTS: Differentiable Architecture Search with Stochastic Implicit GradientsMiao Zhang, Steven W. Su, Shirui Pan, Xiaojun Chang et al.ICML 2021 · 81 citations
- AdvRush: Searching for Adversarially Robust Neural ArchitecturesJisoo Mok, Byunggook Na, Hyeokjun Choe, Sungroh YoonICCV 2021 · 55 citations
Builds on4
- Evaluating The Search Phase of Neural Architecture SearchKaicheng Yu, Christian Sciuto, Martin Jaggi, Claudiu Musat et al.ICLR 2020 · 370 citations
- FairNAS: Rethinking Evaluation Fairness of Weight Sharing Neural Architecture SearchXiangxiang Chu, Bo Zhang, Ruijun XuICCV 2021 · 362 citations
- Multinomial Distribution Learning for Effective Neural Architecture SearchXiawu Zheng, Rongrong Ji, Lang Tang, Baochang Zhang et al.ICCV 2019 · 100 citations
- Improving One-Shot NAS by Suppressing the Posterior FadingXiang Li, Chen Lin, Chuming Li, Ming Sun et al.CVPR 2020
Related papers
- SUMNAS: Supernet with Unbiased Meta-Features for Neural Architecture SearchHyeonmin Ha, Ji-Hoon Kim, Semin Park, Byung-Gon ChunICLR 2022 · 5 citations
- Distribution Consistent Neural Architecture SearchJunyi Pan, Chong Sun, Yizhou Zhou, Ying Zhang et al.CVPR 2022 · 9 citations
- Posterior-Guided Neural Architecture SearchYizhou Zhou, Xiaoyan Sun, Chong Luo, Zheng-Jun Zha et al.AAAI 2020 · 8 citations
- PA&DA: Jointly Sampling PAth and DAta for Consistent NASShun Lu, Yu Hu, Longxing Yang, Zihao Sun et al.CVPR 2023
- Differentiable Neural Architecture Search in Equivalent Space with Exploration EnhancementMiao Zhang, Huiqi Li, Shirui Pan, Xiaojun Chang et al.NeurIPS 2020 · 39 citations
