Benchmarking Quantum Reinforcement Learning
Nico Meyer, Christian Ufrecht, George Yammine, Georgios D. Kontes, Christopher Mutschler, Daniel D. Scherer
Abstract
Benchmarking and establishing proper statistical validation metrics for reinforcement learning (RL) remain ongoing challenges, where no consensus has been established yet. The emergence of quantum computing and its potential applications in quantum reinforcement learning (QRL) further complicate benchmarking efforts. To enable valid performance comparisons and to streamline current research in this area, we propose a novel benchmarking methodology, which is based on a statistical estimator for sample complexity and a definition of statistical outperformance. Furthermore, considering QRL, our methodology casts doubt on some previous claims regarding its superiority. We conducted experiments on a novel benchmarking environment with flexible levels of complexity. While we still identify possible advantages, our findings are more nuanced overall. We discuss the potential limitations of these results and explore their implications for empirical research on quantum advantage in QRL.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Builds on9
- Implementation Matters in Deep RL: A Case Study on PPO and TRPOLogan Engstrom, Andrew Ilyas, Shibani Santurkar, Dimitris Tsipras et al.ICLR 2020 · 305 citations
- Parametrized Quantum Policies for Reinforcement LearningSofiène Jerbi, Casper Gyurik, Simon C. Marshall, Hans J. Briegel et al.NeurIPS 2021 · 201 citations
- Towards a Standardised Performance Evaluation Protocol for Cooperative MARLRihab Gorsane, Omayma Mahjoub, Ruan de Kock, Roland Dubb et al.NeurIPS 2022 · 79 citations
- Evaluating the Performance of Reinforcement Learning AlgorithmsScott M. Jordan, Yash Chandak, Daniel Cohen, Mengxue Zhang et al.ICML 2020 · 59 citations
- What Matters for On-Policy Deep Actor-Critic Methods? A Large-Scale StudyMarcin Andrychowicz, Anton Raichuk, Piotr Stanczyk, Manu Orsini et al.ICLR 2021 · 52 citations
Related papers
- Offline Quantum Reinforcement Learning in a Conservative MannerZhihao Cheng, Kaining Zhang, Li Shen, Dacheng TaoAAAI 2023 · 7 citations
- Optimal Goal-Reaching Reinforcement Learning via Quasimetric LearningTongzhou Wang, Antonio Torralba, Phillip Isola, Amy ZhangICML 2023 · 88 citations
- Breaking the Computational Barrier: Provably Efficient Actor–Critic for Low-Rank MDPsRuiquan Huang, Donghao Li, Yingbin LIANG, Jing YangICML 2026
- SupermarQ: A Scalable Quantum Benchmark SuiteTeague Tomesh, Pranav Gokhale, Victory Omole, Gokul Subramanian Ravi et al.HPCA 2022 · 132 citations
- Quantum Policy Gradient Algorithm with Optimized Action DecodingNico Meyer, Daniel D. Scherer, Axel Plinge, Christopher Mutschler et al.ICML 2023 · 31 citations
