TA-GATES: An Encoding Scheme for Neural Network Architectures
Xuefei Ning, Zixuan Zhou, Junbo Zhao, Tianchen Zhao, Yiping Deng, Changcheng Tang, Shuang Liang, Huazhong Yang, Yu Wang
Abstract
Neural architecture search tries to shift the manual design of neural network (NN) architectures to algorithmic design. In these cases, the NN architecture itself can be viewed as data and needs to be modeled. A better modeling could help explore novel architectures automatically and open the black box of automated architecture design. To this end, this work proposes a new encoding scheme for neural architectures, the Training-Analogous Graph-based ArchiTecture Encoding Scheme (TA-GATES). TA-GATES encodes an NN architecture in a way that is analogous to its training. Extensive experiments demonstrate that the flexibility and discriminative power of TA-GATES lead to better modeling of NN architectures. We expect our methodology of explicitly modeling the NN training process to benefit broader automated deep learning systems. The code is available at https: //github.com/walkerning/aw_nas .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 6a5c8188-e3a4-4679-bc28-c02a65f9c0a5Cited by top-tier papers7
- Dynamic Ensemble of Low-Fidelity Experts: Mitigating NAS "Cold-Start"Junbo Zhao, Xuefei Ning, Enshu Liu, Binxin Ru et al.AAAI 2023 · 5 citations
- Learning to Flow from Generative Pretext Tasks for Neural Architecture EncodingSunwoo Kim, Hyunjin Hwang, Kijung ShinNeurIPS 2025 · 2 citations
- CARL: Causality-Guided Architecture Representation Learning for an Interpretable Performance PredictorHan Ji, Yuqi Feng, Jiahao Fan, Yanan SunICCV 2025 · 1 citation
- NN-Former: Rethinking Graph Structure in Neural Architecture RepresentationRuihan Xu, Haokui Zhang, Yaowei Wang, Wei Zeng et al.CVPR 2025
- CloserToMe: A Unified Framework for Accurate and Transferable Latency Prediction Across Heterogeneous DevicesCheng Tang, Guochong Sui, Wenqi Lou, Zihan Wang et al.AAAI 2026
Builds on21
- RandAugment: Practical Automated Data Augmentation with a Reduced Search SpaceEkin Dogus Cubuk, Barret Zoph, Jonathon Shlens, Quoc LeNeurIPS 2020 · 4,453 citations
- Pruning neural networks without any data by iteratively conserving synaptic flowHidenori Tanaka, Daniel Kunin, Daniel L. K. Yamins, Surya GanguliNeurIPS 2020 · 884 citations
- NAS-Bench-201: Extending the Scope of Reproducible Neural Architecture SearchXuanyi Dong, Yi YangICLR 2020 · 825 citations
- Picking Winning Tickets Before Training by Preserving Gradient FlowChaoqi Wang, Guodong Zhang, Roger B. GrosseICLR 2020 · 743 citations
- PC-DARTS: Partial Channel Connections for Memory-Efficient Architecture SearchYuhui Xu, Lingxi Xie, Xiaopeng Zhang, Xin Chen et al.ICLR 2020 · 691 citations
Related papers
- A Semi-Supervised Assessor of Neural ArchitecturesYehui Tang, Yunhe Wang, Yixing Xu, Hanting Chen et al.CVPR 2020
- Generative Teaching Networks: Accelerating Neural Architecture Search by Learning to Generate Synthetic Training DataFelipe Petroski Such, Aditya Rawal, Joel Lehman, Kenneth O. Stanley et al.ICML 2020 · 180 citations
- Do Not Train It: A Linear Neural Architecture Search of Graph Neural NetworksPeng Xu, Lin Zhang, Xuanzhou Liu, Jiaqi Sun et al.ICML 2023 · 14 citations
- Encodings for Prediction-based Neural Architecture SearchYash Akhauri, Mohamed S. AbdelfattahICML 2024 · 8 citations
- Deep and Flexible Graph Neural Architecture SearchWentao Zhang, Zheyu Lin, Yu Shen, Yang Li et al.ICML 2022 · 5 citations
