SC2020Top-tier venue
Scaling the hartree-fock matrix build on summit
Giuseppe M. J. Barca, David L. Poole, Jorge L. Galvez Vallejo, Melisa Alkan, Colleen Bertoni, Alistair P. Rendell, Mark S. Gordon
Abstract
Usage of Graphics Processing Units (GPU) has become strategic for simulating the chemistry of large molecular systems, with the majority of top supercomputers utilizing GPUs as their main source of computational horsepower. In this paper, a new fragmentation-based Hartree-Fock matrix build algorithm designed for scaling on many-GPU architectures is presented. The new algorithm uses a novel dynamic load balancing scheme based on a binned shell-pair container to distribute batches of significant shell quartets with the same code path to different GPUs. This maximizes computational throughput and load balancing, and eliminates GPU thread divergence due to integral screening. Additionally, the code uses a novel Fock digestion algorithm to contract electron repulsion integrals into the Fock matrix, which exploits all forms of permutational symmetry and eliminates thread synchronization requirements. The implementation demonstrates excellent scalability on the Summit computer, achieving good strong scaling performance up to 4096 nodes, and linear weak scaling up to 612 nodes.
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Your agent calls
Lunesearch_papers
Free to start. No credit card required.
Terminal
Install the CLIlune papers get 57f2558a-353c-429b-a73a-156c245ff6b6Related papers
- Scaling Correlated Fragment Molecular Orbital Calculations on SummitGiuseppe M. J. Barca, Calum Snowdon, Jorge L. Galvez Vallejo, Fazeleh S. Kazemian et al.SC 2022 · 25 citations
- Enabling large-scale correlated electronic structure calculations: scaling the RI-MP2 method on summitGiuseppe M. J. Barca, Jorge L. Galvez Vallejo, David L. Poole, Melisa Alkan et al.SC 2021 · 18 citations
- A Scalable Hybrid Total FETI Method for Massively Parallel FEM SimulationsKehao Lin, Chunbao Zhou, Yan Zeng, Ningming Nie et al.PPoPP 2023 · 2 citations
- Extending the limit of molecular dynamics with ab initio accuracy to 10 billion atomsZhuoqiang Guo, Denghui Lu, Yujin Yan, Siyu Hu et al.PPoPP 2022 · 50 citations
- Scalable All-pairs Shortest Paths for Huge Graphs on Multi-GPU ClustersPiyush Sao, Hao Lu, Ramakrishnan Kannan, Vijay Thakkar et al.HPDC 2021 · 6 citations
