Physics-Informed Approach for Exploratory Hamilton-Jacobi-Bellman Equations via Policy Iterations
Yeongjong Kim, Namkyeong Cho, Minseok Kim, Yeoneung Kim
摘要
We propose a mesh-free policy iteration framework based on physics-informed neural networks (PINNs) for solving entropy-regularized stochastic control problems. The method iteratively alternates between soft policy evaluation and improvement using automatic differentiation and neural approximation, without relying on spatial discretization. We present a detailed L 2 error analysis that decomposes the total approximation error into three sources: iteration error, policy network error, and PDE residual error. The proposed algorithm is validated with a range of challenging control tasks, including high-dimensional linear-quadratic regulation in 5D and 10D, as well as nonlinear systems such as pendulum and cartpole problems. Numerical results confirm the scalability, accuracy, and robustness of our approach across both linear and nonlinear benchmarks.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper1
相关 Paper
- Continuous-Time Value Iteration for Multi-Agent Reinforcement LearningXuefeng Wang, Lei Zhang, Henglin Pu, Ahmed Hussain Qureshi 等ICLR 2026 · 被引用 3 次
- TINNs: Time-Induced Neural Networks for Solving Time-Dependent PDEsChen-Yang Dai, Che-Chia Chang, Te-Sheng Lin, Ming-Chih Lai 等ICML 2026 · 被引用 2 次
- MUSA-PINN: Multi-scale Weak-form Physics-Informed Neural Networks for Fluid Flow in Complex GeometriesWeizheng Zhang, Xunjie Xie, Hao Pan, Xiaowei Duan 等ICML 2026
- Generic bounds on the approximation error for physics-informed (and) operator learningTim De Ryck, Siddhartha MishraNeurIPS 2022 · 被引用 93 次
- Physics-informed Neural Networks for Functional Differential Equations: Cylindrical Approximation and Its Convergence GuaranteesTaiki Miyagawa, Takeru YokotaNeurIPS 2024 · 被引用 8 次
