Lune

NeurIPS2024Top-tier venue

State-free Reinforcement Learning

Mingyu Chen, Aldo Pacchiano, Xuezhou Zhang

2024Year

Abstract

In this work, we study the state-free RL problem, where the algorithm does not have the states information before interacting with the environment. Specifically, denote the reachable state set by SΠ:={s∣max⁡π∈ΠqP,π(s)>0}{S}^\Pi := \{ s|\max_{\pi\in \Pi}q^{P, \pi}(s)>0 \}, we design an algorithm which requires no information on the state space SS while having a regret that is completely independent of S{S} and only depend on SΠ{S}^\Pi. We view this as a concrete first step towards parameter-free RL, with the goal of designing RL algorithms that require no hyper-parameter tuning.

Ask about this paper

Your agent reads all of it.

Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.

Questions to start from

Your agent calls

Luneget_paper_fulltext

Ask in Lune

Free to start. No credit card required.

lune papers fulltext 3ef93f53-a1a9-481d-b50d-04f161901096

Builds on16

Related papers

Dusk over the sea between two cliffs drawn in fine vertical lines