Does Unsupervised Architecture Representation Learning Help Neural Architecture Search?
Shen Yan, Yu Zheng, Wei Ao, Xiao Zeng, Mi Zhang
Abstract
Existing Neural Architecture Search (NAS) methods either encode neural architectures using discrete encodings that do not scale well, or adopt supervised learning-based methods to jointly learn architecture representations and optimize architecture search on such representations which incurs search bias. Despite the widespread use, architecture representations learned in NAS are still poorly understood. We observe that the structural properties of neural architectures are hard to preserve in the latent space if architecture representation learning and search are coupled, resulting in less effective search performance. In this work, we find empirically that pre-training architecture representations using only neural architectures without their accuracies as labels improves the downstream architecture search efficiency. To explain this finding, we visualize how unsupervised architecture representation learning better encourages neural architectures with similar connections and operators to cluster together. This helps map neural architectures with similar performance to the same regions in the latent space and makes the transition of architectures in the latent space relatively smooth, which considerably benefits diverse downstream search strategies. 34th Conference on Neural Information Processing Systems (NeurIPS 2020),
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 4d01fbd5-075f-4778-998f-0ae9df00e10bCited by top-tier papers25
- BANANAS: Bayesian Optimization with Neural Architectures for Neural Architecture SearchColin White, Willie Neiswanger, Yash SavaniAAAI 2021 · 401 citations
- Stronger NAS with Weaker PredictorsJunru Wu, Xiyang Dai, Dongdong Chen, Yinpeng Chen et al.NeurIPS 2021 · 60 citations
- TNASP: A Transformer-based NAS Predictor with a Self-evolution FrameworkShun Lu, Jixiang Li, Jianchao Tan, Sen Yang et al.NeurIPS 2021 · 51 citations
- NAS-Bench-x11 and the Power of Learning CurvesShen Yan, Colin White, Yash Savani, Frank HutterNeurIPS 2021 · 36 citations
- CATE: Computation-aware Neural Architecture Encoding with TransformersShen Yan, Kaiqiang Song, Fei Liu, Mi ZhangICML 2021 · 35 citations
Builds on12
- NAS-Bench-201: Extending the Scope of Reproducible Neural Architecture SearchXuanyi Dong, Yi YangICLR 2020 · 825 citations
- Understanding and Robustifying Differentiable Architecture SearchArber Zela, Thomas Elsken, Tonmoy Saikia, Yassine Marrakchi et al.ICLR 2020 · 408 citations
- BANANAS: Bayesian Optimization with Neural Architectures for Neural Architecture SearchColin White, Willie Neiswanger, Yash SavaniAAAI 2021 · 401 citations
- From Variational to Deterministic AutoencodersPartha Ghosh, Mehdi S. M. Sajjadi, Antonio Vergari, Michael J. Black et al.ICLR 2020 · 298 citations
- Graph Structure of Neural NetworksJiaxuan You, Jure Leskovec, Kaiming He, Saining XieICML 2020 · 168 citations
Related papers
- Encodings for Prediction-based Neural Architecture SearchYash Akhauri, Mohamed S. AbdelfattahICML 2024 · 8 citations
- A Semi-Supervised Assessor of Neural ArchitecturesYehui Tang, Yunhe Wang, Yixing Xu, Hanting Chen et al.CVPR 2020
- CARL: Causality-Guided Architecture Representation Learning for an Interpretable Performance PredictorHan Ji, Yuqi Feng, Jiahao Fan, Yanan SunICCV 2025 · 1 citation
- Understanding Architectures Learnt by Cell-based Neural Architecture SearchYao Shu, Wei Wang, Shaofeng CaiICLR 2020 · 92 citations
- Block-Wisely Supervised Neural Architecture Search With Knowledge DistillationChanglin Li, Jiefeng Peng, Liuchun Yuan, Guangrun Wang et al.CVPR 2020
