PINAT: A Permutation INvariance Augmented Transformer for NAS Predictor
Shun Lu, Yu Hu, Peihao Wang, Yan Han, Jianchao Tan, Jixiang Li, Sen Yang, Ji Liu
摘要
Time-consuming performance evaluation is the bottleneck of traditional Neural Architecture Search (NAS) methods. Predictor-based NAS can speed up performance evaluation by directly predicting performance, rather than training a large number of sub-models and then validating their performance. Most predictor-based NAS approaches use a proxy dataset to train model-based predictors efficiently but suffer from performance degradation and generalization problems. We attribute these problems to the poor abilities of existing predictors to character the sub-models' structure, specifically the topology information extraction and the node feature representation of the input graph data. To address these problems, we propose a Transformer-like NAS predictor PINAT, consisting of a partial Permutation INvariance Augmentation module serving as both token embedding layer and attention head, as well as a Laplacian matrix to be the positional encoding. Our design produces more representative features of the encoded architecture and outperforms state-of-the-art NAS predictors on six search spaces: NAS-Bench-101, NAS-Bench-201, DARTS, ProxylessNAS, PPI, and ModelNet. The code is available at https://github.com/ShunLu91/PINAT.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper10
- SWAP-NAS: Sample-Wise Activation Patterns for Ultra-fast NASYameng Peng, Andy Song, Haytham M. Fayek, Vic Ciesielski 等ICLR 2024 · 被引用 22 次
- ParZC: Parametric Zero-Cost Proxies for Efficient NASPeijie Dong, Lujun Li, Zhenheng Tang, Xiang Liu 等AAAI 2025 · 被引用 14 次
- Permutation Equivariance of Transformers and its ApplicationsHengyuan Xu, Liyao Xiang, Hangyu Ye, Dixi Yao 等CVPR 2024 · 被引用 11 次
- AutoGO: Automated Computation Graph Optimization for Neural Network EvolutionMohammad Salameh, Keith G. Mills, Negar Hassanpour, Fred X. Han 等NeurIPS 2023 · 被引用 7 次
- Building Optimal Neural Architectures Using Interpretable KnowledgeKeith G. Mills, Fred X. Han, Mohammad Salameh, Shengyao Lu 等CVPR 2024 · 被引用 3 次
它引用的顶会 Paper19
- Random Erasing Data AugmentationZhun Zhong, Liang Zheng, Guoliang Kang, Shaozi Li 等AAAI 2020 · 被引用 4,134 次
- KPConv: Flexible and Deformable Convolution for Point CloudsHugues Thomas, Charles R. Qi, Jean-Emmanuel Deschaud, Beatriz Marcotegui 等ICCV 2019 · 被引用 3,193 次
- Do Vision Transformers See Like Convolutional Neural Networks?Maithra Raghu, Thomas Unterthiner, Simon Kornblith, Chiyuan Zhang 等NeurIPS 2021 · 被引用 1,553 次
- NAS-Bench-201: Extending the Scope of Reproducible Neural Architecture SearchXuanyi Dong, Yi YangICLR 2020 · 被引用 825 次
- Progressive Differentiable Architecture Search: Bridging the Depth Gap Between Search and EvaluationXin Chen, Lingxi Xie, Jun Wu, Qi TianICCV 2019 · 被引用 725 次
相关 Paper
- TNASP: A Transformer-based NAS Predictor with a Self-evolution FrameworkShun Lu, Jixiang Li, Jianchao Tan, Sen Yang 等NeurIPS 2021 · 被引用 51 次
- ReNAS: Relativistic Evaluation of Neural Architecture SearchYixing Xu, Yunhe Wang, Kai Han, Yehui Tang 等CVPR 2021
- BRP-NAS: Prediction-based NAS using GCNsLukasz Dudziak, Thomas Chau, Mohamed S. Abdelfattah, Royson Lee 等NeurIPS 2020 · 被引用 233 次
- Surprisingly Strong Performance Prediction with Neural Graph FeaturesGabriela Kadlecová, Jovita Lukasik, Martin Pilát, Petra Vidnerová 等ICML 2024 · 被引用 12 次
- Bridge the Gap Between Architecture Spaces via A Cross-Domain PredictorYuqiao Liu, Yehui Tang, Zeqiong Lv, Yunhe Wang 等NeurIPS 2022 · 被引用 14 次
