Inverse Constrained Reinforcement Learning
Shehryar Malik, Usman Anwar, Alireza Aghasi, Ali Ahmed
摘要
In real world settings, numerous constraints are present which are hard to specify mathematically. However, for the real world deployment of reinforcement learning (RL), it is critical that RL agents are aware of these constraints, so that they can act safely. In this work, we consider the problem of learning constraints from demonstrations of a constraint-abiding agent's behavior. We experimentally validate our approach and show that our framework can successfully learn the most likely constraints that the agent respects. We further show that these learned constraints are transferable to new agents that may have different morphologies and/or reward functions. Previous works in this regard have either mainly been restricted to tabular (discrete) settings, specific types of constraints or assume the environment's transition dynamics. In contrast, our framework is able to learn arbitrary Markovian constraints in high-dimensions in a completely modelfree setting. The code is available at: https: //github.com/shehryar-malik/icrl.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper27
- Sustainable Online Reinforcement Learning for Auto-biddingZhiyu Mou, Yusen Huo, Rongquan Bai, Mingzhou Xie 等NeurIPS 2022 · 被引用 53 次
- Distributed Inverse Constrained Reinforcement Learning for Multi-agent SystemsShicheng Liu, Minghui ZhuNeurIPS 2022 · 被引用 41 次
- Learning Multi-agent Behaviors from Distributed and Streaming DemonstrationsShicheng Liu, Minghui ZhuNeurIPS 2023 · 被引用 34 次
- Multi-Modal Inverse Constrained Reinforcement Learning from a Mixture of DemonstrationsGuanren Qiao, Guiliang Liu, Pascal Poupart, Zhiqiang XuNeurIPS 2023 · 被引用 28 次
- Meta Inverse Constrained Reinforcement Learning: Convergence Guarantee and Generalization AnalysisShicheng Liu, Minghui ZhuICLR 2024 · 被引用 26 次
它引用的顶会 Paper2
- Off-Dynamics Reinforcement Learning: Training for Transfer with Domain ClassifiersBenjamin Eysenbach, Shreyas Chaudhari, Swapnil Asawa, Sergey Levine 等ICLR 2021 · 被引用 120 次
- Maximum Likelihood Constraint Inference for Inverse Reinforcement LearningDexter R. R. Scobee, S. Shankar SastryICLR 2020 · 被引用 74 次
相关 Paper
- Learning Shared Safety Constraints from Multi-task DemonstrationsKonwoo Kim, Gokul Swamy, Zuxin Liu, Ding Zhao 等NeurIPS 2023 · 被引用 31 次
- Robust Inverse Constrained Reinforcement Learning under Model MisspecificationSheng Xu, Guiliang LiuICML 2024 · 被引用 7 次
- Benchmarking Constraint Inference in Inverse Reinforcement LearningGuiliang Liu, Yudong Luo, Ashish Gaurav, Kasra Rezaee 等ICLR 2023 · 被引用 2 次
- From Text to Trajectory: Exploring Complex Constraint Representation and Decomposition in Safe Reinforcement LearningPusen Dong, Tianchen Zhu, Yue Qiu, Haoyi Zhou 等NeurIPS 2024 · 被引用 2 次
- Constraints Penalized Q-learning for Safe Offline Reinforcement LearningHaoran Xu, Xianyuan Zhan, Xiangyu ZhuAAAI 2022 · 被引用 127 次
