Lune

AAAI2020Top-tier venue

IPO: Interior-Point Policy Optimization under Constraints

Yongshuai Liu, Jiaxin Ding, Xin Liu

2020Year
231Citations
41Top-tier citations

Abstract

In this paper, we study reinforcement learning (RL) algorithms to solve real-world decision problems with the objective of maximizing the long-term reward as well as satisfying cumulative constraints. We propose a novel first-order policy optimization method, Interior-point Policy Optimization (IPO), which augments the objective with logarithmic barrier functions, inspired by the interior-point method. Our proposed method is easy to implement with performance guarantees and can handle general types of cumulative multi-constraint settings. We conduct extensive evaluations to compare our approach with state-of-the-art baselines. Our algorithm outperforms the baseline algorithms, in terms of reward maximization and constraint satisfaction.

Ask about this paper

Your agent reads all of it.

Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.

Questions to start from

Your agent calls

Luneget_paper_fulltext

Ask in Lune

Free to start. No credit card required.

lune papers fulltext 30f359f1-7178-4cd8-87f5-4cf1338c408a

Cited by top-tier papers41

Ask how each one uses it

Related papers

Dusk over the sea between two cliffs drawn in fine vertical lines