RZ-NAS: Enhancing LLM-guided Neural Architecture Search via Reflective Zero-Cost Strategy
Zipeng Ji, Guanghui Zhu, Chunfeng Yuan, Yihua Huang
Abstract
LLM-to-NAS is a promising field at the intersection of Large Language Models (LLMs) and Neural Architecture Search (NAS), as recent research has explored the potential of architecture generation leveraging LLMs on multiple search spaces. However, the existing LLM-to-NAS methods face the challenges of limited search spaces, time-cost search efficiency, and uncompetitive performance across standard NAS benchmarks and multiple downstream tasks. In this work, we propose Reflective Zero-cost NAS (RZ-NAS) that can search NAS architectures with humanoid reflections and training-free metrics to elicit the power of LLMs. We rethink LLMs' roles in NAS in current work and design a structured, promptbased to comprehensively understand the search task and architectures from both text and code levels. By integrating LLM reflection modules, we use LLM-generated feedback to provide linguistic guidance within architecture optimization. RZ-NAS enables effective search within both micro and macro search spaces without extensive time cost, achieving SOTA performance across multiple downstream tasks. vital and useful tool. Recently, the combination of Large Language Models (LLMs) with NAS represents a cuttingedge development in automated machine learning, seeking to alleviate the difficulties of manual designs and explore novel architectures on diverse NAS tasks. However, most of the existing work remains in the exploratory phase, where LLMs generate neural architectures directly through textual prompts (Zhao et al., 2023; Yu et al., 2023; Wei et al., 2023) . This approach suffers from two key drawbacks: (1) Reproducibility: Stochastic LLM responses hinder consistent results. ( 2 ) Interpretability: Text-based prompts lack clarity on the design rationale, making optimization and trust difficult. Moreover, LLM-to-NAS methods that focus on generating architecture code (Lehman et al., 2024) can only support tiny search spaces and simplified networks, thus exhibiting poorer performance compared to established NAS algorithms on standard NAS benchmarks (Chen et al., 2023) . Furthermore, current LLM-to-NAS algorithms rely on iterative or evolutionary methods, facing high computational costs, since each architecture requires full training for evaluation. To address this, we aim to design a novel LLMto-NAS algorithm that (1) enhances LLMs' understanding of NAS architectures from both text-level and code-level, (2) addresses the time-cost issue in existing LLM-to-NAS methods, and (3) achieves better performance and scalability for broader search spaces and standard benchmarks.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext bb99de4a-71be-4134-8a83-66e6eaf71173Cited by top-tier papers4
- pTNAS: Progressive Neural Architecture Search for Tabular DataNaili Xing, Shaofeng Cai, Lingze Zeng, Jiaqi Zhu et al.ICML 2026 · 4 citations
- Knowledgeable Language Models as Black-Box Optimizers for Personalized MedicineMichael S. Yao, Osbert Bastani, Alma Andersson, Tommaso Biancalani et al.ICLR 2026
- SAGE-NAS: Synergizing LLM-Based Semantic Agent with Graph-Based Evaluator for Neural Architecture SearchKaiqi Lin, Jianping LuoICML 2026
- Structured Progressive Knowledge Activation for LLM-Driven Neural Architecture SearchZhen Liu, Yuhan Liu, Jinjun Wang, Wei Song et al.ICML 2026
Builds on17
- Reflexion: language agents with verbal reinforcement learningNoah Shinn, Federico Cassano, Ashwin Gopinath, Karthik Narasimhan et al.NeurIPS 2023 · 5,828 citations
- Pruning neural networks without any data by iteratively conserving synaptic flowHidenori Tanaka, Daniel Kunin, Daniel L. K. Yamins, Surya GanguliNeurIPS 2020 · 884 citations
- Picking Winning Tickets Before Training by Preserving Gradient FlowChaoqi Wang, Guodong Zhang, Roger B. GrosseICLR 2020 · 743 citations
- PC-DARTS: Partial Channel Connections for Memory-Efficient Architecture SearchYuhui Xu, Lingxi Xie, Xiaopeng Zhang, Xin Chen et al.ICLR 2020 · 691 citations
- Rethinking the Role of Demonstrations: What Makes In-Context Learning Work?Sewon Min, Xinxi Lyu, Ari Holtzman, Mikel Artetxe et al.EMNLP 2022 · 634 citations
Related papers
- Revolutionizing Training-Free NAS: Towards Efficient Automatic Proxy Discovery via Large Language ModelsHaidong Kang, Lihong Lin, Hanling WangNeurIPS 2025 · 3 citations
- LM-Searcher: Cross-domain Neural Architecture Search with LLMs via Unified Numerical EncodingYuxuan Hu, Jihao Liu, Ke Wang, Jinliang Zheng et al.EMNLP 2025
- L-SWAG: Layer-Sample Wise Activation with Gradients Information for Zero-Shot NAS on Vision TransformersSofia Casarin, Sergio Escalera, Oswald LanzCVPR 2025
- NADER: Neural Architecture Design via Multi-Agent CollaborationZekang Yang, Wang Zeng, Sheng Jin, Chen Qian et al.CVPR 2025
- Per-Architecture Training-Free Metric Optimization for Neural Architecture SearchMingzhuo Lin, Jianping LuoNeurIPS 2025 · 3 citations
