JanusPipe: Efficient Pipeline Parallel Training for Machine Learning Interatomic Potentials
Hongyu Wang, Weijian Liu, Hongtao Xu, Yan Wang, Mingzhen Li, Weile Jia, Guangming Tan
Abstract
Discovering atom-level phenomena requires molecular dynamics (MD) simulations with ab initio accuracy. Machine learning interatomic potentials (MLIPs) enable stable, high-accuracy MD simulations, and their models exhibit scaling-law trends similar to large language models. However, the lack of scalable and efficient distributed training systems for conservative MLIPs makes them difficult to scale. This is because conservative MLIPs inherently follow a double-backward execution pattern, which involves computing gradients during the forward pass. This pattern creates a mismatch with existing distributed training systems, especially for pipeline parallelism. Therefore, we present JanusPipe, an efficient 3D-parallel (PP/DP/GP) training system tailored for conservative MLIPs. It integrates SymFold to enable memory-efficient pipeline parallelism for conservative MLIPs, and WaveK to reduce pipeline bubbles by balancing the four-phase compute time. Experimental results on 32 GPUs show that JanusPipe improves throughput by and on average over 1F1B and Hanayo, respectively.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext f380d1d0-82b5-4d95-abd1-c2dbc3f41252Builds on15
- PairNorm: Tackling Oversmoothing in GNNsLingxiao Zhao, Leman AkogluICLR 2020 · 590 citations
- Efficient large-scale language model training on GPU clusters using megatron-LMDeepak Narayanan, Mohammad Shoeybi, Jared Casper, Patrick LeGresley et al.SC 2021 · 576 citations
- EquiformerV2: Improved Equivariant Transformer for Scaling to Higher-Degree RepresentationsYi-Lun Liao, Brandon M. Wood, Abhishek Das, Tess E. SmidtICLR 2024 · 311 citations
- UMA: A Family of Universal Models for AtomsBrandon M. Wood, Misko Dzamba, Xiang Fu, Meng Gao et al.NeurIPS 2025 · 282 citations
- Chimera: efficiently training large-scale neural networks with bidirectional pipelinesShigang Li, Torsten HoeflerSC 2021 · 124 citations
Related papers
- DistMLIP: A Distributed Inference Platform for Machine Learning Interatomic PotentialsKevin Han, Bowen Deng, Amir Barati Farimani, Gerbrand CederICLR 2026 · 10 citations
- Physics-Informed Weakly Supervised Learning For Interatomic PotentialsMakoto Takamoto, Viktor Zaverkin, Mathias NiepertICML 2025
- FlashTP: Fused, Sparsity-Aware Tensor Product for Machine Learning Interatomic PotentialsSeung Yul Lee, Hojoon Kim, Yutack Park, Dawoon Jeong et al.ICML 2025
- Smooth Dynamic Cutoffs for Machine Learning Interatomic PotentialsKevin Han, Haolin Cong, Bowen Deng, Amir Barati FarimaniICML 2026 · 1 citation
- A recipe for scalable attention-based ML potentials: unlocking long-range accuracy with all-to-all node attentionEric Qu, Brandon Wood, Aditi Krishnapriyan, Zachary UlissiICML 2026 · 14 citations
