Physics-Informed Neural Network Policy Iteration: Algorithms, Convergence, and Verification
Yiming Meng, Ruikun Zhou, Amartya Mukherjee, Maxwell Fitzsimmons, Christopher Song, Jun Liu
摘要
Solving nonlinear optimal control problems is a challenging task, particularly for high-dimensional problems. We propose algorithms for model-based policy iterations to solve nonlinear optimal control problems with convergence guarantees. The main component of our approach is an iterative procedure that utilizes neural approximations to solve linear partial differential equations (PDEs), ensuring convergence. We present two variants of the algorithms. The first variant formulates the optimization problem as a linear least square problem, drawing inspiration from extreme learning machine (ELM) for solving PDEs. This variant efficiently handles low-dimensional problems with high accuracy. The second variant is based on a physics-informed neural network (PINN) for solving PDEs and has the potential to address high-dimensional problems. We demonstrate that both algorithms outperform traditional approaches, such as Galerkin methods, by a significant margin. We provide a theoretical analysis of both algorithms in terms of convergence of neural approximations towards the true optimal solutions in a general setting. Furthermore, we employ formal verification techniques to demonstrate the verifiable stability of the resulting controllers.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- Physics-Informed Approach for Exploratory Hamilton-Jacobi-Bellman Equations via Policy IterationsYeongjong Kim, Namkyeong Cho, Minseok Kim, Yeoneung KimAAAI 2026 · 被引用 4 次
- Continuous-Time Value Iteration for Multi-Agent Reinforcement LearningXuefeng Wang, Lei Zhang, Henglin Pu, Ahmed Hussain Qureshi 等ICLR 2026 · 被引用 3 次
- A Physics-preserved Transfer Learning Method for Differential EquationsHaoran Yang, Chuan-Xian RenNeurIPS 2025 · 被引用 1 次
- Safe Continuous-time Multi-Agent Reinforcement Learning via Epigraph FormXuefeng Wang, Lei Zhang, Henglin Pu, Husheng Li 等ICLR 2026 · 被引用 1 次
它引用的顶会 Paper3
- Training language models to follow instructions with human feedbackLong Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida 等NeurIPS 2022 · 被引用 24,707 次
- Neural Lyapunov Control of Unknown Nonlinear Systems with Stability GuaranteesRuikun Zhou, Thanin Quartz, Hans De Sterck, Jun LiuNeurIPS 2022 · 被引用 109 次
- Continuous-time Model-based Reinforcement LearningÇagatay Yildiz, Markus Heinonen, Harri LähdesmäkiICML 2021 · 被引用 72 次
相关 Paper
- Learning a Neural Solver for Parametric PDEs to Enhance Physics-Informed MethodsLise Le Boudec, Emmanuel de Bézenac, Louis Serrano, Ramon Daniel Regueiro-Espino 等ICLR 2025
- Physics-informed Neural Networks for Functional Differential Equations: Cylindrical Approximation and Its Convergence GuaranteesTaiki Miyagawa, Takeru YokotaNeurIPS 2024 · 被引用 8 次
- Is Physics Informed Loss Always Suitable for Training Physics Informed Neural Network?Chuwei Wang, Shanda Li, Di He, Liwei WangNeurIPS 2022 · 被引用 36 次
- Bi-level Physics-Informed Neural Networks for PDE Constrained Optimization using Broyden's HypergradientsZhongkai Hao, Chengyang Ying, Hang Su, Jun Zhu 等ICLR 2023 · 被引用 2 次
- How does PDE order affect the convergence of PINNs?Changhoon Song, Yesom Park, Myungjoo KangNeurIPS 2024 · 被引用 17 次
