CATE: Computation-aware Neural Architecture Encoding with Transformers
Shen Yan, Kaiqiang Song, Fei Liu, Mi Zhang
Abstract
Recent works (White et al., 2020a; Yan et al., 2020) demonstrate the importance of architecture encodings in Neural Architecture Search (NAS). These encodings encode either structure or computation information of the neural architectures. Compared to structure-aware encodings, computation-aware encodings map architectures with similar accuracies to the same region, which improves the downstream architecture search performance (Zhang et al., 2019; White et al., 2020a). In this work, we introduce a Computation-Aware Transformer-based Encoding method called CATE. Different from existing computation-aware encodings based on fixed transformation (e.g. path encoding), CATE employs a pairwise pre-training scheme to learn computation-aware encodings using Transformers with cross-attention. Such learned encodings contain dense and contextualized computation information of neural architectures. We compare CATE with eleven encodings under three major encoding-dependent NAS subroutines in both small and large search spaces. Our experiments show that CATE is beneficial to the downstream search, especially in the large search space. Moreover, the outside search space experiment demonstrates its superior generalization ability beyond the search space on which it was trained. Our code is available at: https://github.com/MSU-MLSys-Lab/CATE .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 7cd54939-6560-49ce-b05c-6bc206c15790Cited by top-tier papers13
- NAS-Bench-x11 and the Power of Learning CurvesShen Yan, Colin White, Yash Savani, Frank HutterNeurIPS 2021 · 36 citations
- PINAT: A Permutation INvariance Augmented Transformer for NAS PredictorShun Lu, Yu Hu, Peihao Wang, Yan Han et al.AAAI 2023 · 31 citations
- Efficient Data Subset Selection to Generalize Training Across Models: Transductive and Inductive NetworksEeshaan Jain, Tushar Nandy, Gaurav Aggarwal, Ashish Tendulkar et al.NeurIPS 2023 · 30 citations
- DiffusionNAG: Predictor-guided Neural Architecture Generation with Diffusion ModelsSohyun An, Hayeon Lee, Jaehyeong Jo, Seanie Lee et al.ICLR 2024 · 21 citations
- TA-GATES: An Encoding Scheme for Neural Network ArchitecturesXuefei Ning, Zixuan Zhou, Junbo Zhao, Tianchen Zhao et al.NeurIPS 2022 · 19 citations
Builds on17
- NAS-Bench-201: Extending the Scope of Reproducible Neural Architecture SearchXuanyi Dong, Yi YangICLR 2020 · 825 citations
- Understanding and Robustifying Differentiable Architecture SearchArber Zela, Thomas Elsken, Tonmoy Saikia, Yassine Marrakchi et al.ICLR 2020 · 408 citations
- BANANAS: Bayesian Optimization with Neural Architectures for Neural Architecture SearchColin White, Willie Neiswanger, Yash SavaniAAAI 2021 · 401 citations
- Exploring Randomly Wired Neural Networks for Image RecognitionSaining Xie, Alexander Kirillov, Ross B. Girshick, Kaiming HeICCV 2019 · 384 citations
- Stabilizing Differentiable Architecture Search via Perturbation-based RegularizationXiangning Chen, Cho-Jui HsiehICML 2020 · 235 citations
Related papers
- Does Unsupervised Architecture Representation Learning Help Neural Architecture Search?Shen Yan, Yu Zheng, Wei Ao, Xiao Zeng et al.NeurIPS 2020 · 129 citations
- Encodings for Prediction-based Neural Architecture SearchYash Akhauri, Mohamed S. AbdelfattahICML 2024 · 8 citations
- TNASP: A Transformer-based NAS Predictor with a Self-evolution FrameworkShun Lu, Jixiang Li, Jianchao Tan, Sen Yang et al.NeurIPS 2021 · 51 citations
- Prior Knowledge Guided Neural Architecture GenerationJingrong Xie, Han Ji, Yanan SunICML 2025
- A Study on Encodings for Neural Architecture SearchColin White, Willie Neiswanger, Sam Nolen, Yash SavaniNeurIPS 2020 · 88 citations
