Intelligent Switching for Reset-Free RL
Darshan Patil, Janarthanan Rajendran, Glen Berseth, Sarath Chandar
Abstract
In the real world, the strong episode resetting mechanisms that are needed to train agents in simulation are unavailable. The resetting assumption limits the potential of reinforcement learning in the real world, as providing resets to an agent usually requires the creation of additional handcrafted mechanisms or human interventions. Recent work aims to train agents (forward) with learned resets by constructing a second (backward) agent that returns the forward agent to the initial state. We find that the termination and timing of the transitions between these two agents are crucial for algorithm success. With this in mind, we create a new algorithm, Reset Free RL with Intelligently Switching Controller (RISC) which intelligently switches between the two agents based on the agent's confidence in achieving its current goal. Our new method achieves state-of-the-art performance on several challenging environments for reset-free RL.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 29a40e70-ada0-43bf-8eb7-da52700ec3e6Cited by top-tier papers1
Ask how each one uses itBuilds on9
- Deep Reinforcement Learning at the Edge of the Statistical PrecipiceRishabh Agarwal, Max Schwarzer, Pablo Samuel Castro, Aaron C. Courville et al.NeurIPS 2021 · 1,067 citations
- The Ingredients of Real World Robotic Reinforcement LearningHenry Zhu, Justin Yu, Abhishek Gupta, Dhruv Shah et al.ICLR 2020 · 202 citations
- Reset-Free Lifelong Learning with Skill-Space PlanningKevin Lu, Aditya Grover, Pieter Abbeel, Igor MordatchICLR 2021 · 42 citations
- Autonomous Reinforcement Learning via Subgoal CurriculaArchit Sharma, Abhishek Gupta, Sergey Levine, Karol Hausman et al.NeurIPS 2021 · 41 citations
- Autonomous Reinforcement Learning: Formalism and BenchmarkingArchit Sharma, Kelvin Xu, Nikhil Sardana, Abhishek Gupta et al.ICLR 2022 · 39 citations
Related papers
- Demonstration-free Autonomous Reinforcement Learning via Implicit and Bidirectional CurriculumJigang Kim, Daesol Cho, H. Jin KimICML 2023 · 4 citations
- Provable Reset-free Reinforcement Learning by No-Regret ReductionHoai-An Nguyen, Ching-An ChengICML 2023 · 3 citations
- Continual Learning of Control Primitives : Skill Discovery via Reset-GamesKelvin Xu, Siddharth Verma, Chelsea Finn, Sergey LevineNeurIPS 2020 · 37 citations
- A State-Distribution Matching Approach to Non-Episodic Reinforcement LearningArchit Sharma, Rehaan Ahmad, Chelsea FinnICML 2022 · 23 citations
- Learning Robust Policy against Disturbance in Transition Dynamics via State-Conservative Policy OptimizationYufei Kuang, Miao Lu, Jie Wang, Qi Zhou et al.AAAI 2022 · 29 citations
