Revisiting Implicit Differentiation for Learning Problems in Optimal Control
Ming Xu, Timothy L. Molloy, Stephen Gould
摘要
This paper proposes a new method for differentiating through optimal trajectories arising from non-convex, constrained discrete-time optimal control (COC) problems using the implicit function theorem (IFT). Previous works solve a differential Karush-Kuhn-Tucker (KKT) system for the trajectory derivative, and achieve this efficiently by solving an auxiliary Linear Quadratic Regulator (LQR) problem. In contrast, we directly evaluate the matrix equations which arise from applying variable elimination on the Lagrange multiplier terms in the (differential) KKT system. By appropriately accounting for the structure of the terms within the resulting equations, we show that the trajectory derivatives scale linearly with the number of timesteps. Furthermore, our approach allows for easy parallelization, significantly improved scalability with model size, direct computation of vector-Jacobian products and improved numerical stability compared to prior works. As an additional contribution, we unify prior works, addressing claims that computing trajectory derivatives using IFT scales quadratically with the number of timesteps. We evaluate our method on a both synthetic benchmark and four challenging, learning from demonstration benchmarks including a 6-DoF maneuvering quadrotor and 6-DoF rocket powered landing. 37th Conference on Neural Information Processing Systems (NeurIPS 2023).
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- Accelerating Optimization via Differentiable Stopping TimeZhonglin Xie, Yiman Fong, Haoran Yuan, Zaiwen WenNeurIPS 2025 · 被引用 1 次
- DiffOP: Reinforcement Learning of Optimization-Based Control Policies via Implicit Policy GradientsYuexin Bian, Jie Feng, Yuanyuan ShiAAAI 2026 · 被引用 1 次
- DiLQR: Differentiable Iterative Linear Quadratic Regulator via Implicit DifferentiationShuyuan Wang, Philip D. Loewen, Michael G. Forbes, R. Bhushan Gopaluni 等ICML 2025
- DiffTORI: Differentiable Trajectory Optimization for Deep Reinforcement and Imitation LearningWeikang Wan, Ziyu Wang, Yufei Wang, Zackory Erickson 等NeurIPS 2024
- BADControl: Backdoor Attacks Against Control SystemsLuis Burbano, Hampei Sasahara, Ruoyu Song, Z. Berkay Celik 等USENIX Security 2026
它引用的顶会 Paper7
- Differentiation of Blackbox Combinatorial SolversMarin Vlastelica Pogancic, Anselm Paulus, Vít Musil, Georg Martius 等ICLR 2020 · 被引用 341 次
- Fast Differentiable Sorting and RankingMathieu Blondel, Olivier Teboul, Quentin Berthet, Josip DjolongaICML 2020 · 被引用 285 次
- Pontryagin Differentiable Programming: An End-to-End Learning and Control FrameworkWanxin Jin, Zhaoran Wang, Zhuoran Yang, Shaoshuai MouNeurIPS 2020 · 被引用 133 次
- Theseus: A Library for Differentiable Nonlinear OptimizationLuis Pineda, Taosha Fan, Maurizio Monge, Shobha Venkataraman 等NeurIPS 2022 · 被引用 124 次
- Safe Pontryagin Differentiable ProgrammingWanxin Jin, Shaoshuai Mou, George J. PappasNeurIPS 2021 · 被引用 65 次
相关 Paper
- Alternating Differentiation for Optimization LayersHaixiang Sun, Ye Shi, Jingya Wang, Hoang Duong Tuan 等ICLR 2023 · 被引用 3 次
- Infinite-Horizon Differentiable Model Predictive ControlSebastian East, Marco Gallieri, Jonathan Masci, Jan Koutník 等ICLR 2020 · 被引用 39 次
- A Penalty Approach For Differentiation Through Black-box Quadratic Programming SolversYuxuan Linghu, Zhiyuan Liu, Qi DengICML 2026 · 被引用 1 次
- Zeroth-Order Optimization with Trajectory-Informed Derivative EstimationYao Shu, Zhongxiang Dai, Weicong Sng, Arun Verma 等ICLR 2023
- Learning Differential Equations that are Easy to SolveJacob Kelly, Jesse Bettencourt, Matthew J. Johnson, David DuvenaudNeurIPS 2020 · 被引用 134 次
