Lune

ICML2026Top-tier venue

Quantum Robust Inner Minimization for Reinforcement Learning with Quadratic Speed-Up in Query Complexity

Hyun Kyu Lee, Joongheon Kim, Sung Whan Yoon

2026Year

Abstract

Robust reinforcement learning (RRL) aims to tackle unexpected environmental changes by optimizing policies against the worst case. However, RRL remains impractical due to the cost of the Max-Min optimization, where it suffers from the exhaustive query complexity for finding the worst-case (dubbed 'Min') within the environmental uncertainty set U\mathcal{U}, i.e., O(∣U∣)\mathcal{O}(|\mathcal{U}|). By viewing this via a lens of quantum perspective, we raise a pivotal question: If we can query from the environment with quantum superpositions, is it possible to accelerate the Max-Min optimization of RRL? Our answer is 'Yes'. Our method, called quantum robust inner minimization (QRIM), encodes the uncertainty set with quantum superposition and amplifies low-return cases, thus enabling RL for solving the robust (i.e., worst-case) Bellman equation. Importantly, QRIM achieves a quadratic speed-up in query complexity without altering the outer RL pipeline, i.e., O(∣U∣)\mathcal{O}(\sqrt{|\mathcal{U}|}). Validated through classical simulations to real quantum hardware execution, QRIM learns more robust policies with quadratically reduced queries than classical RL.

Ask about this paper

Your agent reads all of it.

Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.

Questions to start from

Your agent calls

Luneget_paper_fulltext

Ask in Lune

Free to start. No credit card required.

lune papers fulltext c4aeed5a-ff5c-4e38-aecf-470a67f54029

Builds on8

Related papers

Dusk over the sea between two cliffs drawn in fine vertical lines