PreNAS: Preferred One-Shot Learning Towards Efficient Neural Architecture Search
Haibin Wang, Ce Ge, Hesen Chen, Xiuyu Sun
Abstract
The wide application of pre-trained models is driving the trend of once-for-all training in one-shot neural architecture search (NAS). However, training within a huge sample space damages the performance of individual subnets and requires much computation to search for an optimal model. In this paper, we present PreNAS, a search-free NAS approach that accentuates target models in one-shot training. Specifically, the sample space is dramatically reduced in advance by a zero-cost selector, and weight-sharing one-shot training is performed on the preferred architectures to alleviate update conflicts. Extensive experiments have demonstrated that PreNAS consistently outperforms state-of-the-art one-shot NAS competitors for both Vision Transformer and convolutional architectures, and importantly, enables instant specialization with zero search cost. Our code is available at https://github.com/tinyvision/PreNAS.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers6
- ParZC: Parametric Zero-Cost Proxies for Efficient NASPeijie Dong, Lujun Li, Zhenheng Tang, Xiang Liu et al.AAAI 2025 · 14 citations
- Vision Transformer Neural Architecture Search for Out-of-Distribution Generalization: Benchmark and InsightsSy-Tuyen Ho, Tuan Van Vo, Somayeh Ebrahimkhani, Ngai-Man CheungNeurIPS 2024 · 5 citations
- NeuroFlux: Memory-Efficient CNN Training Using Adaptive Local LearningDhananjay Saikumar, Blesson VargheseEuroSys 2024 · 2 citations
- Searching Efficient Semantic Segmentation Architectures via Dynamic Path SelectionYuxi Liu, Min Liu, Shuai Jiang, Yi Tang et al.NeurIPS 2025
- TAS-LoRA: Transformer Architecture Search with Mixture-of-LoRA ExpertsJeimin Jeon, Hyunju Lee, Bumsub HamCVPR 2026
Builds on26
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Training data-efficient image transformers & distillation through attentionHugo Touvron, Matthieu Cord, Matthijs Douze, Francisco Massa et al.ICML 2021 · 8,974 citations
- RandAugment: Practical Automated Data Augmentation with a Reduced Search SpaceEkin Dogus Cubuk, Barret Zoph, Jonathon Shlens, Quoc LeNeurIPS 2020 · 4,453 citations
- Random Erasing Data AugmentationZhun Zhong, Liang Zheng, Guoliang Kang, Shaozi Li et al.AAAI 2020 · 4,134 citations
- Tokens-to-Token ViT: Training Vision Transformers from Scratch on ImageNetLi Yuan, Yunpeng Chen, Tao Wang, Weihao Yu et al.ICCV 2021 · 2,462 citations
Related papers
- ShiftNAS: Improving One-shot NAS via Probability ShiftMingyang Zhang, Xinyi Yu, Haodong Zhao, Linlin OuICCV 2023 · 9 citations
- SUMNAS: Supernet with Unbiased Meta-Features for Neural Architecture SearchHyeonmin Ha, Ji-Hoon Kim, Semin Park, Byung-Gon ChunICLR 2022 · 5 citations
- Distribution Consistent Neural Architecture SearchJunyi Pan, Chong Sun, Yizhou Zhou, Ying Zhang et al.CVPR 2022 · 9 citations
- Masked Distillation Advances Self-Supervised Transformer Architecture SearchCaixia Yan, Xiaojun Chang, Zhihui Li, Lina Yao et al.ICLR 2024 · 3 citations
- Searching by Generating: Flexible and Efficient One-Shot NAS With Architecture GeneratorSian-Yao Huang, Wei-Ta ChuCVPR 2021
