RLEKF: An Optimizer for Deep Potential with Ab Initio Accuracy
Siyu Hu, Wentao Zhang, Qiuchen Sha, Feng Pan, Lin-Wang Wang, Weile Jia, Guangming Tan, Tong Zhao
摘要
It is imperative to accelerate the training of neural network force field such as Deep Potential, which usually requires thousands of images based on first-principles calculation and a couple of days to generate an accurate potential energy surface. To this end, we propose a novel optimizer named reorganized layer extended Kalman filtering (RLEKF), an optimized version of global extended Kalman filtering (GEKF) with a strategy of splitting big and gathering small layers to overcome the O(N^2) computational cost of GEKF. This strategy provides an approximation of the dense weights error covariance matrix with a sparse diagonal block matrix for GEKF. We implement both RLEKF and the baseline Adam in our alphaDynamics package and numerical experiments are performed on 13 unbiased datasets. Overall, RLEKF converges faster with slightly better accuracy. For example, a test on a typical system, bulk copper, shows that RLEKF converges faster by both the number of training epochs (x11.67) and wall-clock time (x1.19). Besides, we theoretically prove that the updates of weights converge and thus are against the gradient exploding problem. Experimental results verify that RLEKF is not sensitive to the initialization of weights. The RLEKF sheds light on other AI-for-science applications where training a large neural network (with tons of thousands parameters) is a bottleneck.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper2
相关 Paper
- Training one DeePMD Model in Minutes: a Step towards Online LearningSiyu Hu, Tong Zhao, Qiuchen Sha, Enji Li 等PPoPP 2024 · 被引用 3 次
- A Layer-Wise Natural Gradient Optimizer for Training Deep Neural NetworksXiaolei Liu, Shaoshuai Li, Kaixin Gao, Binfeng WangNeurIPS 2024 · 被引用 2 次
- High-order differentiable autoencoder for nonlinear model reductionSiyuan Shen, Yin Yang, Tianjia Shao, He Wang 等SIGGRAPH 2021 · 被引用 42 次
- KOALA++: Efficient Kalman-Based Optimization with Gradient-Covariance ProductsZixuan Xia, Aram Davtyan, Paolo FavaroNeurIPS 2025 · 被引用 2 次
- KOALA: A Kalman Optimization Algorithm with Loss AdaptivityAram Davtyan, Sepehr Sameni, Llukman Cerkezi, Givi Meishvili 等AAAI 2022 · 被引用 5 次
