PINAT: A Permutation INvariance Augmented Transformer for NAS Predictor
Shun Lu, Yu Hu, Peihao Wang, Yan Han, Jianchao Tan, Jixiang Li, Sen Yang, Ji Liu
Abstract
Time-consuming performance evaluation is the bottleneck of traditional Neural Architecture Search (NAS) methods. Predictor-based NAS can speed up performance evaluation by directly predicting performance, rather than training a large number of sub-models and then validating their performance. Most predictor-based NAS approaches use a proxy dataset to train model-based predictors efficiently but suffer from performance degradation and generalization problems. We attribute these problems to the poor abilities of existing predictors to character the sub-models' structure, specifically the topology information extraction and the node feature representation of the input graph data. To address these problems, we propose a Transformer-like NAS predictor PINAT, consisting of a partial Permutation INvariance Augmentation module serving as both token embedding layer and attention head, as well as a Laplacian matrix to be the positional encoding. Our design produces more representative features of the encoded architecture and outperforms state-of-the-art NAS predictors on six search spaces: NAS-Bench-101, NAS-Bench-201, DARTS, ProxylessNAS, PPI, and ModelNet. The code is available at https://github.com/ShunLu91/PINAT.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext bde0e244-fa84-4b14-9613-6b8fbe14fb92Cited by top-tier papers10
- SWAP-NAS: Sample-Wise Activation Patterns for Ultra-fast NASYameng Peng, Andy Song, Haytham M. Fayek, Vic Ciesielski et al.ICLR 2024 · 22 citations
- ParZC: Parametric Zero-Cost Proxies for Efficient NASPeijie Dong, Lujun Li, Zhenheng Tang, Xiang Liu et al.AAAI 2025 · 14 citations
- Permutation Equivariance of Transformers and its ApplicationsHengyuan Xu, Liyao Xiang, Hangyu Ye, Dixi Yao et al.CVPR 2024 · 11 citations
- AutoGO: Automated Computation Graph Optimization for Neural Network EvolutionMohammad Salameh, Keith G. Mills, Negar Hassanpour, Fred X. Han et al.NeurIPS 2023 · 7 citations
- Building Optimal Neural Architectures Using Interpretable KnowledgeKeith G. Mills, Fred X. Han, Mohammad Salameh, Shengyao Lu et al.CVPR 2024 · 3 citations
Builds on19
- Random Erasing Data AugmentationZhun Zhong, Liang Zheng, Guoliang Kang, Shaozi Li et al.AAAI 2020 · 4,134 citations
- KPConv: Flexible and Deformable Convolution for Point CloudsHugues Thomas, Charles R. Qi, Jean-Emmanuel Deschaud, Beatriz Marcotegui et al.ICCV 2019 · 3,193 citations
- Do Vision Transformers See Like Convolutional Neural Networks?Maithra Raghu, Thomas Unterthiner, Simon Kornblith, Chiyuan Zhang et al.NeurIPS 2021 · 1,553 citations
- NAS-Bench-201: Extending the Scope of Reproducible Neural Architecture SearchXuanyi Dong, Yi YangICLR 2020 · 825 citations
- Progressive Differentiable Architecture Search: Bridging the Depth Gap Between Search and EvaluationXin Chen, Lingxi Xie, Jun Wu, Qi TianICCV 2019 · 725 citations
Related papers
- TNASP: A Transformer-based NAS Predictor with a Self-evolution FrameworkShun Lu, Jixiang Li, Jianchao Tan, Sen Yang et al.NeurIPS 2021 · 51 citations
- ReNAS: Relativistic Evaluation of Neural Architecture SearchYixing Xu, Yunhe Wang, Kai Han, Yehui Tang et al.CVPR 2021
- BRP-NAS: Prediction-based NAS using GCNsLukasz Dudziak, Thomas Chau, Mohamed S. Abdelfattah, Royson Lee et al.NeurIPS 2020 · 233 citations
- Surprisingly Strong Performance Prediction with Neural Graph FeaturesGabriela Kadlecová, Jovita Lukasik, Martin Pilát, Petra Vidnerová et al.ICML 2024 · 12 citations
- Bridge the Gap Between Architecture Spaces via A Cross-Domain PredictorYuqiao Liu, Yehui Tang, Zeqiong Lv, Yunhe Wang et al.NeurIPS 2022 · 14 citations
