Lune

ICML2026顶会

Quantum Robust Inner Minimization for Reinforcement Learning with Quadratic Speed-Up in Query Complexity

Hyun Kyu Lee, Joongheon Kim, Sung Whan Yoon

出版方
2026年份

摘要

Robust reinforcement learning (RRL) aims to tackle unexpected environmental changes by optimizing policies against the worst case. However, RRL remains impractical due to the cost of the Max-Min optimization, where it suffers from the exhaustive query complexity for finding the worst-case (dubbed 'Min') within the environmental uncertainty set U\mathcal{U}, i.e., O(∣U∣)\mathcal{O}(|\mathcal{U}|). By viewing this via a lens of quantum perspective, we raise a pivotal question: If we can query from the environment with quantum superpositions, is it possible to accelerate the Max-Min optimization of RRL? Our answer is 'Yes'. Our method, called quantum robust inner minimization (QRIM), encodes the uncertainty set with quantum superposition and amplifies low-return cases, thus enabling RL for solving the robust (i.e., worst-case) Bellman equation. Importantly, QRIM achieves a quadratic speed-up in query complexity without altering the outer RL pipeline, i.e., O(∣U∣)\mathcal{O}(\sqrt{|\mathcal{U}|}). Validated through classical simulations to real quantum hardware execution, QRIM learns more robust policies with quadratically reduced queries than classical RL.

问问这篇 Paper

智能体会读完全文。

Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。

可以从这些问题问起

智能体调用

Luneget_paper_fulltext

在 Lune 里问

免费开始,无需绑卡

lune papers fulltext c4aeed5a-ff5c-4e38-aecf-470a67f54029

它引用的顶会 Paper8

相关 Paper

黄昏的海面,两侧是细线勾勒的悬崖