Learning Compound Tasks without Task-specific Knowledge via Imitation and Self-supervised Learning
Sang-Hyun Lee, Seung-Woo Seo
摘要
Most real-world tasks are compound tasks that consist of multiple simpler sub-tasks. The main challenge of learning compound tasks is that we have no explicit supervision to learn the hierarchical structure of compound tasks. To address this challenge, previous imitation learning methods exploit task-specific knowledge, e.g., labeling demonstrations manually or specifying termination conditions for each sub-task. However, the need for task-specific knowledge makes it difficult to scale imitation learning to real-world tasks. In this paper, we propose an imitation learning method that can learn compound tasks without task-specific knowledge. The key idea behind our method is to leverage a self-supervised learning framework to learn the hierarchical structure of compound tasks. Our work also proposes a taskagnostic regularization technique to prevent unstable switching between sub-tasks, which has been a common degenerate case in previous works. We evaluate our method against several baselines on compound tasks. The results show that our method achieves state-of-the-art performance on compound tasks, outperforming prior imitation learning methods.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper7
- Adversarial Option-Aware Hierarchical Imitation LearningMingxuan Jing, Wenbing Huang, Fuchun Sun, Xiaojian Ma 等ICML 2021 · 被引用 28 次
- Learning Options via CompressionYiding Jiang, Evan Zheran Liu, Benjamin Eysenbach, J. Zico Kolter 等NeurIPS 2022 · 被引用 26 次
- Learning Task Decomposition with Ordered Memory Policy NetworkYuchen Lu, Yikang Shen, Siyuan Zhou, Aaron C. Courville 等ICLR 2021 · 被引用 17 次
- Inverse Contextual Bandits: Learning How Behavior Evolves over TimeAlihan Hüyük, Daniel Jarrett, Mihaela van der SchaarICML 2022 · 被引用 14 次
- Unsupervised Skill Discovery for Learning Shared Structures across Changing EnvironmentsSang-Hyun Lee, Seung-Woo SeoICML 2023 · 被引用 6 次
相关 Paper
- One-shot Imitation in a Non-Stationary Environment via Multi-Modal SkillSangwoo Shin, Daehee Lee, Minjong Yoo, Woo Kyung Kim 等ICML 2023 · 被引用 12 次
- Self-Supervised Reinforcement Learning that Transfers using Random FeaturesBoyuan Chen, Chuning Zhu, Pulkit Agrawal, Kaiqing Zhang 等NeurIPS 2023 · 被引用 16 次
- Ess-InfoGAIL: Semi-supervised Imitation Learning from Imbalanced DemonstrationsHuiqiao Fu, Kaiqiang Tang, Yuanyang Lu, Yiming Qi 等NeurIPS 2023 · 被引用 15 次
- Watch, Try, Learn: Meta-Learning from Demonstrations and RewardsAllan Zhou, Eric Jang, Daniel Kappler, Alexander Herzog 等ICLR 2020 · 被引用 53 次
- Hierarchical Reinforcement Learning by Discovering Intrinsic OptionsJesse Zhang, Haonan Yu, Wei XuICLR 2021 · 被引用 97 次
