Physics-Informed Neural Network Policy Iteration: Algorithms, Convergence, and Verification
Yiming Meng, Ruikun Zhou, Amartya Mukherjee, Maxwell Fitzsimmons, Christopher Song, Jun Liu
Abstract
Solving nonlinear optimal control problems is a challenging task, particularly for high-dimensional problems. We propose algorithms for model-based policy iterations to solve nonlinear optimal control problems with convergence guarantees. The main component of our approach is an iterative procedure that utilizes neural approximations to solve linear partial differential equations (PDEs), ensuring convergence. We present two variants of the algorithms. The first variant formulates the optimization problem as a linear least square problem, drawing inspiration from extreme learning machine (ELM) for solving PDEs. This variant efficiently handles low-dimensional problems with high accuracy. The second variant is based on a physics-informed neural network (PINN) for solving PDEs and has the potential to address high-dimensional problems. We demonstrate that both algorithms outperform traditional approaches, such as Galerkin methods, by a significant margin. We provide a theoretical analysis of both algorithms in terms of convergence of neural approximations towards the true optimal solutions in a general setting. Furthermore, we employ formal verification techniques to demonstrate the verifiable stability of the resulting controllers.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 2c86c18e-7660-4fcb-a892-b2c0818343d1Cited by top-tier papers4
- Physics-Informed Approach for Exploratory Hamilton-Jacobi-Bellman Equations via Policy IterationsYeongjong Kim, Namkyeong Cho, Minseok Kim, Yeoneung KimAAAI 2026 · 4 citations
- Continuous-Time Value Iteration for Multi-Agent Reinforcement LearningXuefeng Wang, Lei Zhang, Henglin Pu, Ahmed Hussain Qureshi et al.ICLR 2026 · 3 citations
- A Physics-preserved Transfer Learning Method for Differential EquationsHaoran Yang, Chuan-Xian RenNeurIPS 2025 · 1 citation
- Safe Continuous-time Multi-Agent Reinforcement Learning via Epigraph FormXuefeng Wang, Lei Zhang, Henglin Pu, Husheng Li et al.ICLR 2026 · 1 citation
Builds on3
- Training language models to follow instructions with human feedbackLong Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida et al.NeurIPS 2022 · 24,707 citations
- Neural Lyapunov Control of Unknown Nonlinear Systems with Stability GuaranteesRuikun Zhou, Thanin Quartz, Hans De Sterck, Jun LiuNeurIPS 2022 · 109 citations
- Continuous-time Model-based Reinforcement LearningÇagatay Yildiz, Markus Heinonen, Harri LähdesmäkiICML 2021 · 72 citations
Related papers
- Learning a Neural Solver for Parametric PDEs to Enhance Physics-Informed MethodsLise Le Boudec, Emmanuel de Bézenac, Louis Serrano, Ramon Daniel Regueiro-Espino et al.ICLR 2025
- Physics-informed Neural Networks for Functional Differential Equations: Cylindrical Approximation and Its Convergence GuaranteesTaiki Miyagawa, Takeru YokotaNeurIPS 2024 · 8 citations
- Is Physics Informed Loss Always Suitable for Training Physics Informed Neural Network?Chuwei Wang, Shanda Li, Di He, Liwei WangNeurIPS 2022 · 36 citations
- Bi-level Physics-Informed Neural Networks for PDE Constrained Optimization using Broyden's HypergradientsZhongkai Hao, Chengyang Ying, Hang Su, Jun Zhu et al.ICLR 2023 · 2 citations
- How does PDE order affect the convergence of PINNs?Changhoon Song, Yesom Park, Myungjoo KangNeurIPS 2024 · 17 citations
