Benchmarking Quantum Reinforcement Learning
Nico Meyer, Christian Ufrecht, George Yammine, Georgios D. Kontes, Christopher Mutschler, Daniel D. Scherer
摘要
Benchmarking and establishing proper statistical validation metrics for reinforcement learning (RL) remain ongoing challenges, where no consensus has been established yet. The emergence of quantum computing and its potential applications in quantum reinforcement learning (QRL) further complicate benchmarking efforts. To enable valid performance comparisons and to streamline current research in this area, we propose a novel benchmarking methodology, which is based on a statistical estimator for sample complexity and a definition of statistical outperformance. Furthermore, considering QRL, our methodology casts doubt on some previous claims regarding its superiority. We conducted experiments on a novel benchmarking environment with flexible levels of complexity. While we still identify possible advantages, our findings are more nuanced overall. We discuss the potential limitations of these results and explore their implications for empirical research on quantum advantage in QRL.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper9
- Implementation Matters in Deep RL: A Case Study on PPO and TRPOLogan Engstrom, Andrew Ilyas, Shibani Santurkar, Dimitris Tsipras 等ICLR 2020 · 被引用 305 次
- Parametrized Quantum Policies for Reinforcement LearningSofiène Jerbi, Casper Gyurik, Simon C. Marshall, Hans J. Briegel 等NeurIPS 2021 · 被引用 201 次
- Towards a Standardised Performance Evaluation Protocol for Cooperative MARLRihab Gorsane, Omayma Mahjoub, Ruan de Kock, Roland Dubb 等NeurIPS 2022 · 被引用 79 次
- Evaluating the Performance of Reinforcement Learning AlgorithmsScott M. Jordan, Yash Chandak, Daniel Cohen, Mengxue Zhang 等ICML 2020 · 被引用 59 次
- What Matters for On-Policy Deep Actor-Critic Methods? A Large-Scale StudyMarcin Andrychowicz, Anton Raichuk, Piotr Stanczyk, Manu Orsini 等ICLR 2021 · 被引用 52 次
相关 Paper
- Offline Quantum Reinforcement Learning in a Conservative MannerZhihao Cheng, Kaining Zhang, Li Shen, Dacheng TaoAAAI 2023 · 被引用 7 次
- Optimal Goal-Reaching Reinforcement Learning via Quasimetric LearningTongzhou Wang, Antonio Torralba, Phillip Isola, Amy ZhangICML 2023 · 被引用 88 次
- Breaking the Computational Barrier: Provably Efficient Actor–Critic for Low-Rank MDPsRuiquan Huang, Donghao Li, Yingbin LIANG, Jing YangICML 2026
- SupermarQ: A Scalable Quantum Benchmark SuiteTeague Tomesh, Pranav Gokhale, Victory Omole, Gokul Subramanian Ravi 等HPCA 2022 · 被引用 132 次
- Quantum Policy Gradient Algorithm with Optimized Action DecodingNico Meyer, Daniel D. Scherer, Axel Plinge, Christopher Mutschler 等ICML 2023 · 被引用 31 次
