Curriculum reinforcement learning for quantum architecture search under hardware errors
Yash J. Patel, Akash Kundu, Mateusz Ostaszewski, Xavier Bonet-Monroig, Vedran Dunjko, Onur Danaci
Abstract
The key challenge in the noisy intermediate-scale quantum era is finding useful circuits compatible with current device limitations. Variational quantum algorithms (VQAs) offer a potential solution by fixing the circuit architecture and optimizing individual gate parameters in an external loop. However, parameter optimization can become intractable, and the overall performance of the algorithm depends heavily on the initially chosen circuit architecture. Several quantum architecture search (QAS) algorithms have been developed to design useful circuit architectures automatically. In the case of parameter optimization alone, noise effects have been observed to dramatically influence the performance of the optimizer and final outcomes, which is a key line of study. However, the effects of noise on the architecture search, which could be just as critical, are poorly understood. This work addresses this gap by introducing a curriculum-based reinforcement learning QAS (CRLQAS) algorithm designed to tackle challenges in realistic VQA deployment. The algorithm incorporates (i) a 3D architecture encoding and restrictions on environment dynamics to explore the search space of possible circuits efficiently, (ii) an episode halting scheme to steer the agent to find shorter circuits, and (iii) a novel variant of simultaneous perturbation stochastic approximation as an optimizer for faster convergence. To facilitate studies, we developed an optimized simulator for our algorithm, significantly improving computational efficiency in simulating noisy quantum circuits by employing the Pauli-transfer matrix formalism in the Pauli-Liouville basis. Numerical experiments focusing on quantum chemistry tasks demonstrate that CRLQAS outperforms existing QAS algorithms across several metrics in both noiseless and noisy environments.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext cfc6651f-dd11-414c-b1e7-8bedf7fc3dc0Cited by top-tier papers2
- TensorRL-QAS: Reinforcement learning with tensor networks for improved quantum architecture searchAkash Kundu, Stefano ManginiNeurIPS 2025 · 9 citations
- Layerwise Federated Learning for Heterogeneous Quantum Clients using QuorusJason Han, Nicholas S. DiBrita, Daniel Leeds, Jianqiang Li et al.ICLR 2026 · 6 citations
Builds on3
- QuantumNAS: Noise-Adaptive Search for Robust Quantum CircuitsHanrui Wang, Yongshan Ding, Jiaqi Gu, Yujun Lin et al.HPCA 2022 · 199 citations
- QuantumDARTS: Differentiable Quantum Architecture Search for Variational Quantum AlgorithmsWenjie Wu, Ge Yan, Xudong Lu, Kaisen Pan et al.ICML 2023 · 42 citations
- Qubit Routing Using Graph Neural Network Aided Monte Carlo Tree SearchAnimesh Sinha, Utkarsh Azad, Harjinder SinghAAAI 2022 · 31 citations
Related papers
- Training-Free Quantum Architecture SearchZhimin He, Maijie Deng, Shenggen Zheng, Lvzhou Li et al.AAAI 2024 · 39 citations
- Reinforcement learning for optimization of variational quantum circuit architecturesMateusz Ostaszewski, Lea M. Trenkwalder, Wojciech Masarczyk, Eleanor Scerri et al.NeurIPS 2021 · 204 citations
- Q-MAML: Quantum Model-Agnostic Meta-Learning for Variational Quantum AlgorithmsJunyong Lee, Jeihee Cho, Shiho KimAAAI 2025 · 9 citations
- Quarl: A Learning-Based Quantum Circuit OptimizerZikun Li, Jinjun Peng, Yixuan Mei, Sina Lin et al.OOPSLA 2024 · 21 citations
- Quantum Policy Gradient Algorithm with Optimized Action DecodingNico Meyer, Daniel D. Scherer, Axel Plinge, Christopher Mutschler et al.ICML 2023 · 31 citations
