Parametrized Quantum Policies for Reinforcement Learning
Sofiène Jerbi, Casper Gyurik, Simon C. Marshall, Hans J. Briegel, Vedran Dunjko
Abstract
With the advent of real-world quantum computing, the idea that parametrized quantum computations can be used as hypothesis families in a quantum-classical machine learning system is gaining increasing traction. Such hybrid systems have already shown the potential to tackle real-world tasks in supervised and generative learning, and recent works have established their provable advantages in special artificial tasks. Yet, in the case of reinforcement learning, which is arguably most challenging and where learning boosts would be extremely valuable, no proposal has been successful in solving even standard benchmarking tasks, nor in showing a theoretical learning advantage over classical algorithms. In this work, we achieve both. We propose a hybrid quantum-classical reinforcement learning model using very few qubits, which we show can be effectively trained to solve several standard benchmarking environments. Moreover, we demonstrate, and formally prove, the ability of parametrized quantum circuits to solve certain learning tasks that are intractable for classical models, including current state-of-art deep neural networks, under the widely-believed classical hardness of the discrete logarithm problem.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers7
- Non-asymptotic Approximation Error Bounds of Parameterized Quantum CircuitsZhan Yu, Qiuhao Chen, Yuling Jiao, Yinan Li et al.NeurIPS 2024 · 35 citations
- Quantum Policy Gradient Algorithm with Optimized Action DecodingNico Meyer, Daniel D. Scherer, Axel Plinge, Christopher Mutschler et al.ICML 2023 · 31 citations
- Near-optimal Quantum algorithms for multivariate mean estimationArjan Cornelissen, Yassine Hamoudi, Sofiène JerbiSTOC 2022 · 16 citations
- Offline Quantum Reinforcement Learning in a Conservative MannerZhihao Cheng, Kaining Zhang, Li Shen, Dacheng TaoAAAI 2023 · 7 citations
- eQMARL: Entangled Quantum Multi-Agent Reinforcement Learning for Distributed Cooperation over Quantum ChannelsAlexander C. DeRieux, Walid SaadICLR 2025
Related papers
- Learning to Optimize Variational Quantum Circuits to Solve Combinatorial ProblemsSami Khairy, Ruslan Shaydulin, Lukasz Cincio, Yuri Alexeev et al.AAAI 2020 · 155 citations
- VSQL: Variational Shadow Quantum Learning for ClassificationGuangxi Li, Zhixin Song, Xin WangAAAI 2021 · 55 citations
- Recurrent Quantum Neural NetworksJohannes BauschNeurIPS 2020 · 223 citations
- On the Relation between Trainability and Dequantization of Variational Quantum Learning ModelsElies Gil-Fuster, Casper Gyurik, Adrián Pérez-Salinas, Vedran DunjkoICLR 2025
- Concentration of Data Encoding in Parameterized Quantum CircuitsGuangxi Li, Ruilin Ye, Xuanqiang Zhao, Xin WangNeurIPS 2022 · 42 citations
