Scalable Kernel Inverse Optimization
Youyuan Long, Tolga Ok, Pedro Zattoni Scroccaro, Peyman Mohajerin Esfahani
摘要
Inverse Optimization (IO) is a framework for learning the unknown objective function of an expert decision-maker from a past dataset. In this paper, we extend the hypothesis class of IO objective functions to a reproducing kernel Hilbert space (RKHS), thereby enhancing feature representation to an infinite-dimensional space. We demonstrate that a variant of the representer theorem holds for a specific training loss, allowing the reformulation of the problem as a finite-dimensional convex optimization program. To address scalability issues commonly associated with kernel methods, we propose the Sequential Selection Optimization (SSO) algorithm to efficiently train the proposed Kernel Inverse Optimization (KIO) model. Finally, we validate the generalization capabilities of the proposed KIO model and the effectiveness of the SSO algorithm through learning-from-demonstration tasks on the MuJoCo benchmark.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper2
相关 Paper
- Meta-Learning Hypothesis Spaces for Sequential Decision-makingParnian Kassraie, Jonas Rothfuss, Andreas KrauseICML 2022 · 被引用 6 次
- Efficient Exploration of Reward Functions in Inverse Reinforcement Learning via Bayesian OptimizationSreejith Balakrishnan, Quoc Phong Nguyen, Bryan Kian Hsiang Low, Harold SohNeurIPS 2020 · 被引用 33 次
- On Computation and Generalization of Generative Adversarial Imitation LearningMinshuo Chen, Yizhou Wang, Tianyi Liu, Zhuoran Yang 等ICLR 2020 · 被引用 42 次
- Is Inverse Reinforcement Learning Harder than Standard Reinforcement Learning? A Theoretical PerspectiveLei Zhao, Mengdi Wang, Yu BaiICML 2024 · 被引用 3 次
- Koopman Kernel RegressionPetar Bevanda, Max Beier, Armin Lederer, Stefan Sosnowski 等NeurIPS 2023 · 被引用 36 次
