Weakly-Supervised Reinforcement Learning for Controllable Behavior
Lisa Lee, Ben Eysenbach, Ruslan Salakhutdinov, Shixiang Shane Gu, Chelsea Finn
摘要
Reinforcement learning (RL) is a powerful framework for learning to take actions to solve tasks. However, in many settings, an agent must winnow down the inconceivably large space of all possible tasks to the single task that it is currently being asked to solve. Can we instead constrain the space of tasks to those that are semantically meaningful? In this work, we introduce a framework for using weak supervision to automatically disentangle this semantically meaningful subspace of tasks from the enormous space of nonsensical "chaff" tasks. We show that this learned subspace enables efficient exploration and provides a representation that captures distance between states. On a variety of challenging, vision-based continuous control problems, our approach leads to substantial performance gains, particularly as the complexity of the environment grows.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper9
- Maximum Entropy Gain Exploration for Long Horizon Multi-goal Reinforcement LearningSilviu Pitis, Harris Chan, Stephen Zhao, Bradly C. Stadie 等ICML 2020 · 被引用 145 次
- Learning Domain Invariant Representations in Goal-conditioned Block MDPsBeining Han, Chongyi Zheng, Harris Chan, Keiran Paster 等NeurIPS 2021 · 被引用 20 次
- Policy Learning Using Weak SupervisionJingkang Wang, Hongyi Guo, Zhaowei Zhu, Yang LiuNeurIPS 2021 · 被引用 16 次
- Distributional Reward Estimation for Effective Multi-agent Deep Reinforcement LearningJifeng Hu, Yanchao Sun, Hechang Chen, Sili Huang 等NeurIPS 2022 · 被引用 13 次
- Learning State Representations via Retracing in Reinforcement LearningChangmin Yu, Dong Li, Jianye Hao, Jun Wang 等ICLR 2022 · 被引用 9 次
它引用的顶会 Paper6
- Model Based Reinforcement Learning for AtariLukasz Kaiser, Mohammad Babaeizadeh, Piotr Milos, Blazej Osinski 等ICLR 2020 · 被引用 969 次
- Improving Sample Efficiency in Model-Free Reinforcement Learning from ImagesDenis Yarats, Amy Zhang, Ilya Kostrikov, Brandon Amos 等AAAI 2021 · 被引用 506 次
- Stochastic Latent Actor-Critic: Deep Reinforcement Learning with a Latent Variable ModelAlex X. Lee, Anusha Nagabandi, Pieter Abbeel, Sergey LevineNeurIPS 2020 · 被引用 437 次
- Skew-Fit: State-Covering Self-Supervised Reinforcement LearningVitchyr Pong, Murtaza Dalal, Steven Lin, Ashvin Nair 等ICML 2020 · 被引用 303 次
- Weakly Supervised Disentanglement with GuaranteesRui Shu, Yining Chen, Abhishek Kumar, Stefano Ermon 等ICLR 2020 · 被引用 148 次
相关 Paper
- Semantic Exploration from Language Abstractions and Pretrained RepresentationsAllison C. Tam, Neil C. Rabinowitz, Andrew K. Lampinen, Nicholas A. Roy 等NeurIPS 2022 · 被引用 85 次
- Reinforcement Learning with Prototypical RepresentationsDenis Yarats, Rob Fergus, Alessandro Lazaric, Lerrel PintoICML 2021 · 被引用 262 次
- Reinforcement Learning with Neural Radiance FieldsDanny Driess, Ingmar Schubert, Pete Florence, Yunzhu Li 等NeurIPS 2022 · 被引用 72 次
- Possibility Before Utility: Learning And Using Hierarchical AffordancesRobby Costales, Shariq Iqbal, Fei ShaICLR 2022 · 被引用 5 次
- Denoised MDPs: Learning World Models Better Than the World ItselfTongzhou Wang, Simon S. Du, Antonio Torralba, Phillip Isola 等ICML 2022 · 被引用 63 次
