KAM Theory Meets Statistical Learning Theory: Hamiltonian Neural Networks with Non-zero Training Loss
Yuhan Chen, Takashi Matsubara, Takaharu Yaguchi
摘要
Many physical phenomena are described by Hamiltonian mechanics using an energy function (the Hamiltonian). Recently, the Hamiltonian neural network, which approximates the Hamiltonian as a neural network, and its extensions have attracted much attention. This is a very powerful method, but its use in theoretical studies remains limited. In this study, by combining the statistical learning theory and Kolmogorov-Arnold-Moser (KAM) theory, we provide a theoretical analysis of the behavior of Hamiltonian neural networks when the learning error is not completely zero. A Hamiltonian neural network with non-zero errors can be considered as a perturbation from the true dynamics, and the perturbation theory of the Hamilton equation is widely known as the KAM theory. To apply the KAM theory, we provide a generalization error bound for Hamiltonian neural networks by deriving an estimate of the covering number of the gradient of the multilayer perceptron, which is the key ingredient of the model. This error bound gives an L ∞ bound on the Hamiltonian that is required in the application of the KAM theory.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper8
- Generalizing Convolutional Neural Networks for Equivariance to Lie Groups on Arbitrary Continuous DataMarc Finzi, Samuel Stanton, Pavel Izmailov, Andrew Gordon WilsonICML 2020 · 被引用 372 次
- Symplectic ODE-Net: Learning Hamiltonian Dynamics with ControlYaofeng Desmond Zhong, Biswadip Dey, Amit ChakrabortyICLR 2020 · 被引用 319 次
- Symplectic Recurrent Neural NetworksZhengdao Chen, Jianyu Zhang, Martín Arjovsky, Léon BottouICLR 2020 · 被引用 261 次
- Hamiltonian Generative NetworksPeter Toth, Danilo J. Rezende, Andrew Jaegle, Sébastien Racanière 等ICLR 2020 · 被引用 242 次
- Simplifying Hamiltonian and Lagrangian Neural Networks via Explicit ConstraintsMarc Finzi, Ke Alexander Wang, Andrew Gordon WilsonNeurIPS 2020 · 被引用 168 次
相关 Paper
- When are dynamical systems learned from time series data statistically accurate?Jeongjin Park, Nicole Yang, Nisha ChandramoorthyNeurIPS 2024 · 被引用 17 次
- Generalization Bounds and Model Complexity for Kolmogorov-Arnold NetworksXianyang Zhang, Huijuan ZhouICLR 2025
- Hamiltonian Neural PDE Solvers through Functional ApproximationAnthony Y. Zhou, Amir Barati FarimaniNeurIPS 2025 · 被引用 1 次
- UEPI: Universal Energy-Behavior-Preserving Integrators for Energy Conservative/Dissipative Differential EquationsElena Celledoni, Brynjulf Owren, Chong Shen, Baige Xu 等NeurIPS 2025 · 被引用 3 次
- Structure Preserving Neural Networks: A Case Study in the Entropy Closure of the Boltzmann EquationSteffen Schotthöfer, Tianbai Xiao, Martin Frank, Cory D. HauckICML 2022 · 被引用 14 次
