Differentiable Model Predictive Control on the GPU
Emre Adabag, Marcus Greiff, John Subosits, Thomas Jonathan Lew
Abstract
Differentiable model predictive control (MPC) offers a powerful framework for combining learning and control. However, its adoption has been limited by the inherently sequential nature of traditional optimization algorithms, which are challenging to parallelize on modern computing hardware like GPUs. In this work, we tackle this bottleneck by introducing a GPU-accelerated differentiable optimization tool for MPC. This solver leverages sequential quadratic programming and a custom preconditioned conjugate gradient (PCG) routine with tridiagonal preconditioning to exploit the problem's structure and enable efficient parallelization. We demonstrate substantial speedups over CPU- and GPU-based baselines, significantly improving upon state-of-the-art training times on benchmark reinforcement learning and imitation learning tasks. Finally, we showcase the method on the challenging task of reinforcement learning for driving at the limits of handling, where it enables robust drifting of a Toyota Supra through water puddles.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext e5c632a5-f9fa-4cc7-bb5a-df477d01ad47Builds on5
- Theseus: A Library for Differentiable Nonlinear OptimizationLuis Pineda, Taosha Fan, Maurizio Monge, Shobha Venkataraman et al.NeurIPS 2022 · 124 citations
- DOC: Differentiable Optimal Control for Retargeting Motions onto Legged RobotsRuben Grandia, Farbod Farshidian, Espen Knoop, Christian Schumacher et al.SIGGRAPH 2023 · 20 citations
- Leveraging augmented-Lagrangian techniques for differentiating over infeasible quadratic programs in machine learningAntoine Bambade, Fabian Schramm, Adrien B. Taylor, Justin CarpentierICLR 2024 · 9 citations
- CRONOS: Enhancing Deep Learning with Scalable GPU Accelerated Convex Neural NetworksMiria Feng, Zachary Frangella, Mert PilanciNeurIPS 2024 · 6 citations
- DiffTORI: Differentiable Trajectory Optimization for Deep Reinforcement and Imitation LearningWeikang Wan, Ziyu Wang, Yufei Wang, Zackory Erickson et al.NeurIPS 2024
Related papers
- Infinite-Horizon Differentiable Model Predictive ControlSebastian East, Marco Gallieri, Jonathan Masci, Jan Koutník et al.ICLR 2020 · 39 citations
- Parallel Q-Learning: Scaling Off-policy Reinforcement Learning under Massively Parallel SimulationZechu Li, Tao Chen, Zhang-Wei Hong, Anurag Ajay et al.ICML 2023 · 27 citations
- GMI-DRL: Empowering Multi-GPU DRL with Adaptive-Grained ParallelismYuke Wang, Boyuan Feng, Zheng Wang, Guyue Huang et al.USENIX ATC 2025 · 3 citations
- DiLQR: Differentiable Iterative Linear Quadratic Regulator via Implicit DifferentiationShuyuan Wang, Philip D. Loewen, Michael G. Forbes, R. Bhushan Gopaluni et al.ICML 2025
- Accelerating Quadratic Optimization with Reinforcement LearningJeffrey Ichnowski, Paras Jain, Bartolomeo Stellato, Goran Banjac et al.NeurIPS 2021 · 62 citations
