Lune

NeurIPS2025顶会

Non-rectangular Robust MDPs with Normed Uncertainty Sets

Navdeep Kumar, Adarsh Gupta, Maxence Mohamed Elfatihi, Giorgia Ramponi, Kfir Y. Levy, Shie Mannor

2025年份
3被引次数
1顶会引用

摘要

Robust policy evaluation for non-rectangular uncertainty set is generally NP-hard, even in approximation. Consequently, existing approaches suffer from either exponential iteration complexity or significant accuracy gaps. Interestingly, we identify a powerful class of L p -bounded uncertainty sets that avoid these complexity barriers due to their structural simplicity. We further show that this class can be decomposed into infinitely many sa -rectangular L p -bounded sets and leverage its structural properties to derive a novel dual formulation for L p robust Markov Decision Processes (MDPs). This formulation reveals key insights into the adversary’s strategy and leads to the first polynomial-time robust policy evaluation algorithm for L 1 -normed non-rectangular robust MDPs.

问问这篇 Paper

智能体会读完全文。

Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。

可以从这些问题问起

智能体调用

Luneget_paper_fulltext

在 Lune 里问

免费开始,无需绑卡

引用它的顶会 Paper1

问问它们各自怎么用它

它引用的顶会 Paper9

相关 Paper

黄昏的海面,两侧是细线勾勒的悬崖