Building Optimal Neural Architectures Using Interpretable Knowledge
Keith G. Mills, Fred X. Han, Mohammad Salameh, Shengyao Lu, Chunhua Zhou, Jiao He, Fengyu Sun, Di Niu
Abstract
Neural Architecture Search is a costly practice. The fact that a search space can span a vast number of design choices with each architecture evaluation taking nontrivial overhead makes it hard for an algorithm to sufficiently explore candidate networks. In this paper, we propose Auto-Build, a scheme which learns to align the latent embeddings of operations and architecture modules with the ground-truth performance of the architectures they appear in. By doing so, AutoBuild is capable of assigning interpretable importance scores to architecture modules, such as individual operation features and larger macro operation sequences such that high-performance neural networks can be constructed without any need for search. Through experiments performed on state-of-the-art image classification, segmentation, and Stable Diffusion models, we show that by mining a relatively small set of evaluated architectures, AutoBuild can learn to build high-quality architectures directly or help to reduce search space to focus on relevant areas, finding better architectures that outperform both the original labeled ones and ones found by search baselines. Code available at https://github.com/Ascend-Research/AutoBuild
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 59fa9229-0e30-4be4-a673-ccf5f87e4d38Cited by top-tier papers4
- CARL: Causality-Guided Architecture Representation Learning for an Interpretable Performance PredictorHan Ji, Yuqi Feng, Jiahao Fan, Yanan SunICCV 2025 · 1 citation
- Progressive Neural Architecture GenerationCaiyang Yu, Chen Huang, Yun Liu, Chenwei Tang et al.CVPR 2026
- Prior Knowledge Guided Neural Architecture GenerationJingrong Xie, Han Ji, Yanan SunICML 2025
- Qua2SeDiMo: Quantifiable Quantization Sensitivity of Diffusion ModelsKeith G. Mills, Mohammad Salameh, Ruichen Chen, Negar Hassanpour et al.AAAI 2025
Builds on19
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Searching for MobileNetV3Andrew Howard, Ruoming Pang, Hartwig Adam, Quoc V. Le et al.ICCV 2019 · 9,163 citations
- SDXL: Improving Latent Diffusion Models for High-Resolution Image SynthesisDustin Podell, Zion English, Kyle Lacey, Andreas Blattmann et al.ICLR 2024 · 4,569 citations
- How Attentive are Graph Attention Networks?Shaked Brody, Uri Alon, Eran YahavICLR 2022 · 1,717 citations
- Once-for-All: Train One Network and Specialize it for Efficient DeploymentHan Cai, Chuang Gan, Tianzhe Wang, Zhekai Zhang et al.ICLR 2020 · 1,522 citations
Related papers
- Not All Diffusion Model Activations Have Been Evaluated as Discriminative FeaturesBenyuan Meng, Qianqian Xu, Zitai Wang, Xiaochun Cao et al.NeurIPS 2024 · 26 citations
- A Semi-Supervised Assessor of Neural ArchitecturesYehui Tang, Yunhe Wang, Yixing Xu, Hanting Chen et al.CVPR 2020
- AutoSpace: Neural Architecture Search with Less Human InterferenceDaquan Zhou, Xiaojie Jin, Xiaochen Lian, Linjie Yang et al.ICCV 2021 · 11 citations
- SGAS: Sequential Greedy Architecture SearchGuohao Li, Guocheng Qian, Itzel C. Delgadillo, Matthias Müller et al.CVPR 2020
- Neural Graph Embedding for Neural Architecture SearchWei Li, Shaogang Gong, Xiatian ZhuAAAI 2020 · 31 citations
