Prediction, Consistency, Curvature: Representation Learning for Locally-Linear Control
Nir Levine, Yinlam Chow, Rui Shu, Ang Li, Mohammad Ghavamzadeh, Hung Bui
Abstract
Many real-world sequential decision-making problems can be formulated as optimal control with high-dimensional observations and unknown dynamics. A promising approach is to embed the high-dimensional observations into a lower-dimensional latent representation space, estimate the latent dynamics model, then utilize this model for control in the latent space. An important open question is how to learn a representation that is amenable to existing control algorithms? In this paper, we focus on learning representations for locally-linear control algorithms, such as iterative LQR (iLQR). By formulating and analyzing the representation learning problem from an optimal control perspective, we establish three underlying principles that the learned representation should comprise: 1) accurate prediction in the observation space, 2) consistency between latent and observation space dynamics, and 3) low curvature in the latent space transitions. These principles naturally correspond to a loss function that consists of three terms: prediction, consistency, and curvature (PCC). Crucially, to make PCC tractable, we derive an amortized variational bound for the PCC loss function. Extensive experiments on benchmark domains demonstrate that the new variational-PCC learning algorithm benefits from significantly more stable and reproducible training, and leads to superior control performance. Further ablation studies give support to the importance of all three PCC components for learning a good latent space for control.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext bd0fe9a0-d907-4a7c-9027-5345cdaad1caCited by top-tier papers15
- Representation Matters: Offline Pretraining for Sequential Decision MakingMengjiao Yang, Ofir NachumICML 2021 · 126 citations
- Temporal Predictive Coding For Model-Based Planning In Latent SpaceTung D. Nguyen, Rui Shu, Tuan Pham, Hung Bui et al.ICML 2021 · 65 citations
- The Differentiable Cross-Entropy MethodBrandon Amos, Denis YaratsICML 2020 · 60 citations
- Unsupervised Learning of Lagrangian Dynamics from Images for Prediction and ControlYaofeng Desmond Zhong, Naomi Ehrich LeonardNeurIPS 2020 · 49 citations
- Value Gradient weighted Model-Based Reinforcement LearningClaas Voelcker, Victor Liao, Animesh Garg, Amir-massoud FarahmandICLR 2022 · 37 citations
Builds on1
Related papers
- Control-Aware Representations for Model-based Reinforcement LearningBrandon Cui, Yinlam Chow, Mohammad GhavamzadehICLR 2021 · 8 citations
- Predictive Coding for Locally-Linear ControlRui Shu, Tung Nguyen, Yinlam Chow, Tuan Pham et al.ICML 2020 · 28 citations
- iLQR-VAE : control-based learning of input-driven dynamics with applications to neural dataMarine Schimel, Ta-Chu Kao, Kristopher T. Jensen, Guillaume HennequinICLR 2022 · 40 citations
- Rich-Observation Reinforcement Learning with Continuous Latent DynamicsYuda Song, Lili Wu, Dylan J. Foster, Akshay KrishnamurthyICML 2024 · 2 citations
- Learning the Linear Quadratic Regulator from Nonlinear ObservationsZakaria Mhammedi, Dylan J. Foster, Max Simchowitz, Dipendra Misra et al.NeurIPS 2020 · 33 citations
