f-GAIL: Learning f-Divergence for Generative Adversarial Imitation Learning
Xin Zhang, Yanhua Li, Ziming Zhang, Zhi-Li Zhang
Abstract
Imitation learning (IL) aims to learn a policy from expert demonstrations that minimizes the discrepancy between the learner and expert behaviors. Various imitation learning algorithms have been proposed with different pre-determined divergences to quantify the discrepancy. This naturally gives rise to the following question: Given a set of expert demonstrations, which divergence can recover the expert policy more accurately with higher data efficiency? In this work, we propose -GAIL, a new generative adversarial imitation learning (GAIL) model, that automatically learns a discrepancy measure from the -divergence family as well as a policy capable of producing expert-like behaviors. Compared with IL baselines with various predefined divergence measures, -GAIL learns better policies with higher data efficiency in six physics-based control tasks.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 58d97a2a-9ab6-4fc7-bd88-29068e59264cCited by top-tier papers4
- Planning for Sample Efficient Imitation LearningZhao-Heng Yin, Weirui Ye, Qifeng Chen, Yang GaoNeurIPS 2022 · 32 citations
- DiffAIL: Diffusion Adversarial Imitation LearningBingzheng Wang, Guoqiang Wu, Teng Pang, Yan Zhang et al.AAAI 2024 · 24 citations
- Inverse Contextual Bandits: Learning How Behavior Evolves over TimeAlihan Hüyük, Daniel Jarrett, Mihaela van der SchaarICML 2022 · 14 citations
- FedSkill: Privacy Preserved Interpretable Skill Learning via ImitationYushan Jiang, Wenchao Yu, Dongjin Song, Lu Wang et al.KDD 2023 · 5 citations
Builds on1
Related papers
- Learning to Weight Imperfect DemonstrationsYunke Wang, Chang Xu, Bo Du, Honglak LeeICML 2021 · 57 citations
- Diffusion-Reward Adversarial Imitation LearningChun-Mao Lai, Hsiang-Chun Wang, Ping-Chun Hsieh, Yu-Chiang Frank Wang et al.NeurIPS 2024 · 28 citations
- C-GAIL: Stabilizing Generative Adversarial Imitation Learning with Control TheoryTianjiao Luo, Tim Pearce, Huayu Chen, Jianfei Chen et al.NeurIPS 2024 · 13 citations
- Variational Adversarial Kernel Learned Imitation LearningFan Yang, Alina Vereshchaka, Yufan Zhou, Changyou Chen et al.AAAI 2020 · 9 citations
- Exploring Gradient Explosion in Generative Adversarial Imitation Learning: A Probabilistic PerspectiveWanying Wang, Yichen Zhu, Yirui Zhou, Chaomin Shen et al.AAAI 2024 · 13 citations
