Robust and Scalable Autonomous Reinforcement Learning in Irreversible Environments
Sang-Hyun Lee
摘要
Reinforcement learning (RL) typically assumes repetitive resets to provide an agent with diverse and unbiased experiences. These resets require significant human intervention and result in poor training efficiency in real-world settings. Autonomous RL (ARL) addresses this challenge by jointly training forward and reset policies. While recent ARL algorithms have shown promise in reducing human intervention, they assume narrow support over the distributions of initial or goal states and rely on task-specific knowledge to identify irreversible states. In this paper, we propose a robust and scalable ARL algorithm, called RSA, that enables an agent to handle diverse initial and goal states and to avoid irreversible states without task-specific knowledge. RSA generates a curriculum by identifying informative states based on the learning progress of an agent. We hypothesize that informative states are neither overly difficult nor trivially easy for the agent being trained. To detect and avoid irreversible states without task-specific knowledge, RSA encodes the behaviors exhibited in those states rather than the states themselves. Experimental results demonstrate that RSA outperforms existing ARL algorithms with fewer manual resets in both reversible and irreversible environments.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper7
- The Ingredients of Real World Robotic Reinforcement LearningHenry Zhu, Justin Yu, Abhishek Gupta, Dhruv Shah 等ICLR 2020 · 被引用 202 次
- Autonomous Reinforcement Learning via Subgoal CurriculaArchit Sharma, Abhishek Gupta, Sergey Levine, Karol Hausman 等NeurIPS 2021 · 被引用 41 次
- Autonomous Reinforcement Learning: Formalism and BenchmarkingArchit Sharma, Kelvin Xu, Nikhil Sardana, Abhishek Gupta 等ICLR 2022 · 被引用 39 次
- Continual Learning of Control Primitives : Skill Discovery via Reset-GamesKelvin Xu, Siddharth Verma, Chelsea Finn, Sergey LevineNeurIPS 2020 · 被引用 37 次
- When to Ask for Help: Proactive Interventions in Autonomous Reinforcement LearningAnnie Xie, Fahim Tajwar, Archit Sharma, Chelsea FinnNeurIPS 2022 · 被引用 30 次
相关 Paper
- Demonstration-free Autonomous Reinforcement Learning via Implicit and Bidirectional CurriculumJigang Kim, Daesol Cho, H. Jin KimICML 2023 · 被引用 4 次
- Automated curriculum generation through setter-solver interactionsSébastien Racanière, Andrew K. Lampinen, Adam Santoro, David P. Reichert 等ICLR 2020 · 被引用 41 次
- Intelligent Switching for Reset-Free RLDarshan Patil, Janarthanan Rajendran, Glen Berseth, Sarath ChandarICLR 2024 · 被引用 1 次
- Safe Reinforcement Learning via Curriculum InductionMatteo Turchetta, Andrey Kolobov, Shital Shah, Andreas Krause 等NeurIPS 2020 · 被引用 109 次
- Self-Paced Deep Reinforcement LearningPascal Klink, Carlo D'Eramo, Jan Peters, Joni PajarinenNeurIPS 2020 · 被引用 83 次
