Lune

NeurIPS2024顶会

State-free Reinforcement Learning

Mingyu Chen, Aldo Pacchiano, Xuezhou Zhang

2024年份

摘要

In this work, we study the state-free RL problem, where the algorithm does not have the states information before interacting with the environment. Specifically, denote the reachable state set by SΠ:={s∣max⁡π∈ΠqP,π(s)>0}{S}^\Pi := \{ s|\max_{\pi\in \Pi}q^{P, \pi}(s)>0 \}, we design an algorithm which requires no information on the state space SS while having a regret that is completely independent of S{S} and only depend on SΠ{S}^\Pi. We view this as a concrete first step towards parameter-free RL, with the goal of designing RL algorithms that require no hyper-parameter tuning.

问问这篇 Paper

智能体会读完全文。

Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。

可以从这些问题问起

智能体调用

Luneget_paper_fulltext

在 Lune 里问

免费开始,无需绑卡

它引用的顶会 Paper16

相关 Paper

黄昏的海面,两侧是细线勾勒的悬崖