Learning Robot Skills with Temporal Variational Inference
Tanmay Shankar, Abhinav Gupta
摘要
In this paper, we address the discovery of robotic options from demonstrations in an unsupervised manner. Specifically, we present a framework to jointly learn low-level control policies and higher-level policies of how to use them from demonstrations of a robot performing various tasks. By representing options as continuous latent variables, we frame the problem of learning these options as latent variable inference. We then present a temporal formulation of variational inference based on a temporal factorization of trajectory likelihoods,that allows us to infer options in an unsupervised manner. We demonstrate the ability of our framework to learn such options across three robotic demonstration datasets.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper27
- Accelerating Robotic Reinforcement Learning via Parameterized Action PrimitivesMurtaza Dalal, Deepak Pathak, Ruslan SalakhutdinovNeurIPS 2021 · 被引用 121 次
- Reinforcement Learning with Action ChunkingQiyang Li, Zhiyuan Zhou, Sergey LevineNeurIPS 2025 · 被引用 114 次
- Learning Multimodal Behaviors from Scratch with Diffusion Policy GradientSteven Li, Rickmer Krohn, Tao Chen, Anurag Ajay 等NeurIPS 2024 · 被引用 61 次
- RODE: Learning Roles to Decompose Multi-Agent TasksTonghan Wang, Tarun Gupta, Anuj Mahajan, Bei Peng 等ICLR 2021 · 被引用 60 次
- TRAIL: Near-Optimal Imitation Learning with Suboptimal DataMengjiao Yang, Sergey Levine, Ofir NachumICLR 2022 · 被引用 54 次
它引用的顶会 Paper1
相关 Paper
- Bayesian Nonparametrics for Offline Skill DiscoveryValentin Villecroze, Harry J. Braviner, Panteha Naderian, Chris J. Maddison 等ICML 2022 · 被引用 9 次
- Hierarchical Reinforcement Learning by Discovering Intrinsic OptionsJesse Zhang, Haonan Yu, Wei XuICLR 2021 · 被引用 97 次
- Data-efficient Hindsight Off-policy Option LearningMarkus Wulfmeier, Dushyant Rao, Roland Hafner, Thomas Lampe 等ICML 2021 · 被引用 48 次
- A Unified Algorithm Framework for Unsupervised Discovery of Skills based on Determinantal Point ProcessJiayu Chen, Vaneet Aggarwal, Tian LanNeurIPS 2023 · 被引用 8 次
- Adversarial Option-Aware Hierarchical Imitation LearningMingxuan Jing, Wenbing Huang, Fuchun Sun, Xiaojian Ma 等ICML 2021 · 被引用 28 次
