Neural architecture search as program transformation exploration
Jack Turner, Elliot J. Crowley, Michael F. P. O'Boyle
Abstract
Improving the performance of deep neural networks (DNNs) is important to both the compiler and neural architecture search (NAS) communities. Compilers apply program transformations in order to exploit hardware parallelism and memory hierarchy. However, legality concerns mean they fail to exploit the natural robustness of neural networks. In contrast, NAS techniques mutate networks by operations such as the grouping or bottlenecking of convolutions, exploiting the resilience of DNNs. In this work, we express such neural architecture operations as program transformations whose legality depends on a notion of representational capacity. This allows them to be combined with existing transformations into a unified optimization framework. This unification allows us to express existing NAS operations as combinations of simpler transformations. Crucially, it allows us to generate and explore new tensor convolutions. We prototyped the combined framework in TVM and were able to find optimizations across different DNNs, that significantly reduce inference time -over 3× in the majority of cases. Furthermore, our scheme dramatically reduces NAS search time. Code is available at this https url.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext dc59abc3-0faf-42e9-a027-026819d4d1d6Cited by top-tier papers7
- GradSign: Model Performance Inference with Theoretical InsightsZhihao Zhang, Zhihao JiaICLR 2022 · 37 citations
- Towards Theoretically Inspired Neural Initialization OptimizationYibo Yang, Hong Wang, Haobo Yuan, Zhouchen LinNeurIPS 2022 · 15 citations
- MAGIS: Memory Optimization via Coordinated Graph Transformation and Scheduling for DNNRenze Chen, Zijian Ding, Size Zheng, Chengrui Zhang et al.ASPLOS 2024 · 14 citations
- Neural architecture search using property guided synthesisCharles Jin, Phitchaya Mangpo Phothilimthana, Sudip RoyOOPSLA 2022 · 8 citations
- UNICO: Unified Hardware Software Co-Optimization for Robust Neural Network AccelerationBahador Rashidi, Chao Gao, Shan Lu, Zhisheng Wang et al.MICRO 2023 · 6 citations
Builds on7
- NAS-Bench-201: Extending the Scope of Reproducible Neural Architecture SearchXuanyi Dong, Yi YangICLR 2020 · 825 citations
- Ansor: Generating High-Performance Tensor Programs for Deep LearningLianmin Zheng, Chengfan Jia, Minmin Sun, Zhao Wu et al.OSDI 2020 · 551 citations
- Neural Architecture Search without TrainingJoe Mellor, Jack Turner, Amos Storkey, Elliot J. CrowleyICML 2021 · 477 citations
- Evaluating The Search Phase of Neural Architecture SearchKaicheng Yu, Christian Sciuto, Martin Jaggi, Claudiu Musat et al.ICLR 2020 · 370 citations
- NAS-Bench-1Shot1: Benchmarking and Dissecting One-shot Neural Architecture SearchArber Zela, Julien Siems, Frank HutterICLR 2020 · 156 citations
Related papers
- Syno: Structured Synthesis for Neural OperatorsYongqi Zhuo, Zhengyuan Su, Chenggang Zhao, Mingyu GaoASPLOS 2025
- EINNET: Optimizing Tensor Programs with Derivation-Based TransformationsLiyan Zheng, Haojie Wang, Jidong Zhai, Muyan Hu et al.OSDI 2023 · 19 citations
- Felix: Optimizing Tensor Programs with Gradient DescentYifan Zhao, Hashim Sharif, Vikram S. Adve, Sasa MisailovicASPLOS 2024 · 13 citations
- PET: Optimizing Tensor Programs with Partially Equivalent Transformations and Automated CorrectionsHaojie Wang, Jidong Zhai, Mingyu Gao, Zixuan Ma et al.OSDI 2021 · 77 citations
- Explainable-DSE: An Agile and Explainable Exploration of Efficient HW/SW Codesigns of Deep Learning Accelerators Using Bottleneck AnalysisShail Dave, Tony Nowatzki, Aviral ShrivastavaASPLOS 2023 · 7 citations
