Learning Soft Constraints From Constrained Expert Demonstrations
Ashish Gaurav, Kasra Rezaee, Guiliang Liu, Pascal Poupart
摘要
Inverse reinforcement learning (IRL) methods assume that the expert data is generated by an agent optimizing some reward function. However, in many settings, the agent may optimize a reward function subject to some constraints, where the constraints induce behaviors that may be otherwise difficult to express with just a reward function. We consider the setting where the reward function is given, and the constraints are unknown, and propose a method that is able to recover these constraints satisfactorily from the expert data. While previous work has focused on recovering hard constraints, our method can recover cumulative soft constraints that the agent satisfies on average per episode. In IRL fashion, our method solves this problem by adjusting the constraint function iteratively through a constrained optimization procedure, until the agent behavior matches the expert behavior. We demonstrate our approach on synthetic environments, robotics environments and real world highway driving scenarios.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper8
- Multi-Modal Inverse Constrained Reinforcement Learning from a Mixture of DemonstrationsGuanren Qiao, Guiliang Liu, Pascal Poupart, Zhiqiang XuNeurIPS 2023 · 被引用 28 次
- Uncertainty-aware Constraint Inference in Inverse Constrained Reinforcement LearningSheng Xu, Guiliang LiuICLR 2024 · 被引用 12 次
- Robust Inverse Constrained Reinforcement Learning under Model MisspecificationSheng Xu, Guiliang LiuICML 2024 · 被引用 7 次
- Confidence Aware Inverse Constrained Reinforcement LearningSriram Ganapathi Subramanian, Guiliang Liu, Mohammed Elmahgiubi, Kasra Rezaee 等ICML 2024 · 被引用 5 次
- Learning Constraints from Offline Demonstrations via Superior Distribution Correction EstimationGuorui Quan, Zhiqiang Xu, Guiliang LiuICML 2024 · 被引用 4 次
它引用的顶会 Paper4
- Maximum Likelihood Constraint Inference for Inverse Reinforcement LearningDexter R. R. Scobee, S. Shankar SastryICLR 2020 · 被引用 74 次
- DC3: A learning method for optimization with hard constraintsPriya L. Donti, David Rolnick, J. Zico KolterICLR 2021 · 被引用 64 次
- Deep Inverse Q-learning with ConstraintsGabriel Kalweit, Maria Hügle, Moritz Werling, Joschka BoedeckerNeurIPS 2020 · 被引用 39 次
- Inverse Constrained Reinforcement LearningShehryar Malik, Usman Anwar, Alireza Aghasi, Ali AhmedICML 2021 · 被引用 14 次
相关 Paper
- Benchmarking Constraint Inference in Inverse Reinforcement LearningGuiliang Liu, Yudong Luo, Ashish Gaurav, Kasra Rezaee 等ICLR 2023 · 被引用 2 次
- Provably Efficient Exploration in Inverse Constrained Reinforcement LearningBo Yue, Jian Li, Guiliang LiuICML 2025
- Learning Shared Safety Constraints from Multi-task DemonstrationsKonwoo Kim, Gokul Swamy, Zuxin Liu, Ding Zhao 等NeurIPS 2023 · 被引用 31 次
- Understanding Constraint Inference in Safety-Critical Inverse Reinforcement LearningBo Yue, Shufan Wang, Ashish Gaurav, Jian Li 等ICLR 2025
- Inverse Reinforcement Learning in a Continuous State Space with Formal GuaranteesGregory Dexter, Kevin Bello, Jean HonorioNeurIPS 2021 · 被引用 9 次
