AIO-P: Expanding Neural Performance Predictors beyond Image Classification
Keith G. Mills, Di Niu, Mohammad Salameh, Weichen Qiu, Fred X. Han, Puyuan Liu, Jialin Zhang, Wei Lu, Shangling Jui
摘要
Evaluating neural network performance is critical to deep neural network design but a costly procedure. Neural predictors provide an efficient solution by treating architectures as samples and learning to estimate their performance on a given task. However, existing predictors are task-dependent, predominantly estimating neural network performance on image classification benchmarks. They are also search-space dependent; each predictor is designed to make predictions for a specific architecture search space with predefined topologies and set of operations. In this paper, we propose a novel All-in-One Predictor (AIO-P), which aims to pretrain neural predictors on architecture examples from multiple, separate computer vision (CV) task domains and multiple architecture spaces, and then transfer to unseen downstream CV tasks or neural architectures. We describe our proposed techniques for general graph representation, efficient predictor pretraining and knowledge infusion techniques, as well as methods to transfer to downstream tasks/spaces. Extensive experimental results show that AIO-P can achieve Mean Absolute Error (MAE) and Spearman’s Rank Correlation (SRCC) below 1p% and above 0.5, respectively, on a breadth of target downstream CV tasks with or without fine-tuning, outperforming a number of baselines. Moreover, AIO-P can directly transfer to new architectures not seen during training, accurately rank them and serve as an effective performance estimator when paired with an algorithm designed to preserve performance while reducing FLOPs.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- AutoGO: Automated Computation Graph Optimization for Neural Network EvolutionMohammad Salameh, Keith G. Mills, Negar Hassanpour, Fred X. Han 等NeurIPS 2023 · 被引用 7 次
- Building Optimal Neural Architectures Using Interpretable KnowledgeKeith G. Mills, Fred X. Han, Mohammad Salameh, Shengyao Lu 等CVPR 2024 · 被引用 3 次
它引用的顶会 Paper11
- Once-for-All: Train One Network and Specialize it for Efficient DeploymentHan Cai, Chuang Gan, Tianzhe Wang, Zhekai Zhang 等ICLR 2020 · 被引用 1,522 次
- Pruning neural networks without any data by iteratively conserving synaptic flowHidenori Tanaka, Daniel Kunin, Daniel L. K. Yamins, Surya GanguliNeurIPS 2020 · 被引用 884 次
- NAS-Bench-201: Extending the Scope of Reproducible Neural Architecture SearchXuanyi Dong, Yi YangICLR 2020 · 被引用 825 次
- BANANAS: Bayesian Optimization with Neural Architectures for Neural Architecture SearchColin White, Willie Neiswanger, Yash SavaniAAAI 2021 · 被引用 401 次
- Semi-Supervised Neural Architecture SearchRenqian Luo, Xu Tan, Rui Wang, Tao Qin 等NeurIPS 2020 · 被引用 106 次
相关 Paper
- Bridge the Gap Between Architecture Spaces via A Cross-Domain PredictorYuqiao Liu, Yehui Tang, Zeqiong Lv, Yunhe Wang 等NeurIPS 2022 · 被引用 14 次
- GENNAPE: Towards Generalized Neural Architecture Performance EstimatorsKeith G. Mills, Fred X. Han, Jialin Zhang, Fabian Chudak 等AAAI 2023 · 被引用 9 次
- A Semi-Supervised Assessor of Neural ArchitecturesYehui Tang, Yunhe Wang, Yixing Xu, Hanting Chen 等CVPR 2020
- ReNAS: Relativistic Evaluation of Neural Architecture SearchYixing Xu, Yunhe Wang, Kai Han, Yehui Tang 等CVPR 2021
- Homogeneous Architecture Augmentation for Neural PredictorYuqiao Liu, Yehui Tang, Yanan SunICCV 2021 · 被引用 33 次
