Physics-Informed Approach for Exploratory Hamilton-Jacobi-Bellman Equations via Policy Iterations
Yeongjong Kim, Namkyeong Cho, Minseok Kim, Yeoneung Kim
Abstract
We propose a mesh-free policy iteration framework based on physics-informed neural networks (PINNs) for solving entropy-regularized stochastic control problems. The method iteratively alternates between soft policy evaluation and improvement using automatic differentiation and neural approximation, without relying on spatial discretization. We present a detailed L 2 error analysis that decomposes the total approximation error into three sources: iteration error, policy network error, and PDE residual error. The proposed algorithm is validated with a range of challenging control tasks, including high-dimensional linear-quadratic regulation in 5D and 10D, as well as nonlinear systems such as pendulum and cartpole problems. Numerical results confirm the scalability, accuracy, and robustness of our approach across both linear and nonlinear benchmarks.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext cae31a79-8c3a-4cd0-86fa-147f0c35ff5eBuilds on1
Related papers
- Continuous-Time Value Iteration for Multi-Agent Reinforcement LearningXuefeng Wang, Lei Zhang, Henglin Pu, Ahmed Hussain Qureshi et al.ICLR 2026 · 3 citations
- TINNs: Time-Induced Neural Networks for Solving Time-Dependent PDEsChen-Yang Dai, Che-Chia Chang, Te-Sheng Lin, Ming-Chih Lai et al.ICML 2026 · 2 citations
- MUSA-PINN: Multi-scale Weak-form Physics-Informed Neural Networks for Fluid Flow in Complex GeometriesWeizheng Zhang, Xunjie Xie, Hao Pan, Xiaowei Duan et al.ICML 2026
- Generic bounds on the approximation error for physics-informed (and) operator learningTim De Ryck, Siddhartha MishraNeurIPS 2022 · 93 citations
- Physics-informed Neural Networks for Functional Differential Equations: Cylindrical Approximation and Its Convergence GuaranteesTaiki Miyagawa, Takeru YokotaNeurIPS 2024 · 8 citations
