Extending the limit of molecular dynamics with ab initio accuracy to 10 billion atoms
Zhuoqiang Guo, Denghui Lu, Yujin Yan, Siyu Hu, Rongrong Liu, Guangming Tan, Ninghui Sun, Wanrun Jiang, Lijun Liu, Yixiao Chen, Linfeng Zhang, Mohan Chen
Abstract
High-performance computing, together with a neural network model trained from data generated with first-principles methods, has greatly boosted applications of ab initio molecular dynamics in terms of spatial and temporal scales on modern supercomputers. Previous state-of-the-art can achieve 1 -- 2 nanoseconds molecular dynamics simulation per day for 100-million atoms on the entire Summit supercomputer. In this paper, we have significantly reduced the memory footprint and computational time by a comprehensive approach with both algorithmic and system innovations. The neural network model is compressed by model tabulation, kernel fusion, and redundancy removal. Then optimizations such as acceleration of customized kernel, tabulation of activation function, MPI+OpenMP parallelization are implemented on GPU and ARM architectures. Testing results of the copper system show that the optimized code can scale up to the entire machine of both Fugaku and Summit, and the corresponding system size can be extended by a factor of 134 to an unprecedented 17 billion atoms. The strong scaling of a 13.5-million atom copper system shows that the time-to-solution can be 7 times faster, reaching 11.2 nanoseconds per day. This work opens the door for unprecedentedly large-scale molecular dynamics simulations based on ab initio accuracy and can be potentially utilized in studying more realistic applications such as mechanical properties of metals, semiconductor devices, batteries, etc. The optimization techniques detailed in this paper also provide insight for relevant high-performance computing applications.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 6e55dd82-52d2-4cd0-a9f0-1d7ff3449c71Cited by top-tier papers4
- DistMLIP: A Distributed Inference Platform for Machine Learning Interatomic PotentialsKevin Han, Bowen Deng, Amir Barati Farimani, Gerbrand CederICLR 2026 · 10 citations
- Scaling Molecular Dynamics with ab initio Accuracy to 149 Nanoseconds per DayJianxiong Li, Boyang Li, Zhuoqiang Guo, Mingzhen Li et al.SC 2024 · 9 citations
- RLEKF: An Optimizer for Deep Potential with Ab Initio AccuracySiyu Hu, Wentao Zhang, Qiuchen Sha, Feng Pan et al.AAAI 2023 · 5 citations
- Deep Learning-Enabled Supercritical Flame Simulation at Detailed Chemistry and Real-Fluid Accuracy Towards Trillion-Cell ScaleZhuoqiang Guo, Runze Mao, Lijun Liu, Guangming Tan et al.SC 2025 · 1 citation
Builds on7
- ZeRO: memory optimizations toward training trillion parameter modelsSamyam Rajbhandari, Jeff Rasley, Olatunji Ruwase, Yuxiong HeSC 2020 · 852 citations
- Chimera: efficiently training large-scale neural networks with bidirectional pipelinesShigang Li, Torsten HoeflerSC 2021 · 124 citations
- A Mean Field Analysis Of Deep ResNet And Beyond: Towards Provably Optimization Via Overparameterization From DepthYiping Lu, Chao Ma, Yulong Lu, Jianfeng Lu et al.ICML 2020 · 85 citations
- Accelerating sparse DNN models without hardware-support via tile-wise sparsityCong Guo, Bo Yang Hsueh, Jingwen Leng, Yuxian Qiu et al.SC 2020 · 65 citations
- GEMS: GPU-enabled memory-aware model-parallelism system for distributed DNN trainingArpan Jain, Ammar Ahmad Awan, Asmaa M. Aljuhani, Jahanzeb Maqbool Hashmi et al.SC 2020 · 48 citations
Related papers
- Enabling large-scale correlated electronic structure calculations: scaling the RI-MP2 method on summitGiuseppe M. J. Barca, Jorge L. Galvez Vallejo, David L. Poole, Melisa Alkan et al.SC 2021 · 18 citations
- Training one DeePMD Model in Minutes: a Step towards Online LearningSiyu Hu, Tong Zhao, Qiuchen Sha, Enji Li et al.PPoPP 2024 · 3 citations
- Scaling Correlated Fragment Molecular Orbital Calculations on SummitGiuseppe M. J. Barca, Calum Snowdon, Jorge L. Galvez Vallejo, Fazeleh S. Kazemian et al.SC 2022 · 25 citations
- MISA-AKMC : Achieve Kinetic Monte Carlo Simulation of 20 Quadrillion Atoms on GPU ClustersShunde Li, Zhijie Pan, Ningming Nie, Jue Wang et al.SC 2025 · 1 citation
- TensorKMC: kinetic Monte Carlo simulation of 50 trillion atoms driven by deep learning on a new generation of Sunway supercomputerHonghui Shang, Xin Chen, Xingyu Gao, Rongfen Lin et al.SC 2021 · 16 citations
