Neural Architecture Search Using Deep Neural Networks and Monte Carlo Tree Search
Linnan Wang, Yiyang Zhao, Yuu Jinnai, Yuandong Tian, Rodrigo Fonseca
Abstract
Neural Architecture Search (NAS) has shown great success in automating the design of neural networks, but the prohibitive amount of computations behind current NAS methods requires further investigations in improving the sample efficiency and the network evaluation cost to get better results in a shorter time. In this paper, we present a novel scalable Monte Carlo Tree Search (MCTS) based NAS agent, named AlphaX, to tackle these two aspects. AlphaX improves the search efficiency by adaptively balancing the exploration and exploitation at the state level, and by a Meta-Deep Neural Network (DNN) to predict network accuracies for biasing the search toward a promising region. To amortize the network evaluation cost, AlphaX accelerates MCTS rollouts with a distributed design and reduces the number of epochs in evaluating a network by transfer learning, which is guided with the tree structure in MCTS. In 12 GPU days and 1000 samples, AlphaX found an architecture that reaches 97.84% top-1 accuracy on CIFAR-10, and 75.5% top-1 accuracy on Ima-geNet, exceeding SOTA NAS methods in both the accuracy and sampling efficiency. Particularly, we also evaluate Al-phaX on NASBench-101, a large scale NAS dataset; AlphaX is 3x and 2.8x more sample efficient than Random Search and Regularized Evolution in finding the global optimum. Finally, we show the searched architecture improves a variety of vision applications from Neural Style Transfer, to Image Captioning and Object Detection.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext bd5c6fb2-4d23-4115-a1e4-b4825eba8a7dCited by top-tier papers4
- Rethink DARTS Search Space and Renovate a New BenchmarkJiuling Zhang, Zhiming DingICML 2023 · 3 citations
- A New Paradigm in Tuning Learned Indexes: A Reinforcement Learning Enhanced ApproachTaiyi Wang, Liang Liang, Guang Yang, Thomas Heinis et al.SIGMOD 2025 · 2 citations
- Prioritized Architecture Sampling With Monto-Carlo Tree SearchXiu Su, Tao Huang, Yanxi Li, Shan You et al.CVPR 2021
- FBNetV3: Joint Architecture-Recipe Search Using Predictor PretrainingXiaoliang Dai, Alvin Wan, Peizhao Zhang, Bichen Wu et al.CVPR 2021
Related papers
- Combinatorial Neural BanditsTaehyun Hwang, Kyuwook Chai, Min-hwan OhICML 2023 · 7 citations
- Rapid Neural Architecture Search by Learning to Generate Graphs from DatasetsHayeon Lee, Eunyoung Hyung, Sung Ju HwangICLR 2021 · 57 citations
- Towards Fast Adaptation of Neural Architectures with Meta LearningDongze Lian, Yin Zheng, Yintao Xu, Yanxiong Lu et al.ICLR 2020 · 95 citations
- Multinomial Distribution Learning for Effective Neural Architecture SearchXiawu Zheng, Rongrong Ji, Lang Tang, Baochang Zhang et al.ICCV 2019 · 100 citations
- Fast and Practical Neural Architecture SearchJiequan Cui, Pengguang Chen, Ruiyu Li, Shu Liu et al.ICCV 2019 · 69 citations
