A recipe for scalable attention-based ML potentials: unlocking long-range accuracy with all-to-all node attention
Eric Qu, Brandon Wood, Aditi Krishnapriyan, Zachary Ulissi
Abstract
Machine-learning interatomic potentials (MLIPs) have advanced rapidly, with many top models relying on strong physics-based inductive biases. However, as models scale to larger systems like biomolecules and electrolytes, they struggle to accurately capture long-range (LR) interactions, leading current approaches to rely on explicit physics-based terms or components. In this work, we propose AllScAIP, a straightforward, attention-based, and energy-conserving MLIP model that scales to O(100 million) training samples. It addresses the long-range challenge using an all-to-all node attention component that is data-driven. Extensive ablations reveal that in low-data/small-model regimes, inductive biases improve sample efficiency. However, as data and model size scale, these benefits diminish or even reverse, while all-to-all attention remains critical for capturing LR interactions. Our model achieves state-of-the-art energy/force accuracy on molecular systems, as well as a number of physics-based evaluations (OMol25), while being competitive on materials (OMat24) and catalysts (OC20). Furthermore, it enables stable, long-timescale MD simulations that accurately recover experimental observables, including density and heat of vaporization predictions.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext bfe1d416-4364-4f8c-bf32-70e8c07d47c1Builds on8
- Directional Message Passing for Molecular GraphsJohannes Klicpera, Janek Groß, Stephan GünnemannICLR 2020 · 1,079 citations
- GemNet: Universal Directional Graph Neural Networks for MoleculesJohannes Gasteiger, Florian Becker, Stephan GünnemannNeurIPS 2021 · 665 citations
- EquiformerV2: Improved Equivariant Transformer for Scaling to Higher-Degree RepresentationsYi-Lun Liao, Brandon M. Wood, Abhishek Das, Tess E. SmidtICLR 2024 · 311 citations
- Reducing SO(3) Convolutions to SO(2) for Efficient Equivariant GNNsSaro Passaro, C. Lawrence ZitnickICML 2023 · 157 citations
- Spherical Channels for Modeling Atomic InteractionsLarry Zitnick, Abhishek Das, Adeesh Kolluru, Janice Lan et al.NeurIPS 2022 · 84 citations
Related papers
- The Importance of Being Scalable: Improving the Speed and Accuracy of Neural Network Interatomic Potentials Across Chemical DomainsEric Qu, Aditi S. KrishnapriyanNeurIPS 2024 · 63 citations
- Smooth Dynamic Cutoffs for Machine Learning Interatomic PotentialsKevin Han, Haolin Cong, Bowen Deng, Amir Barati FarimaniICML 2026 · 1 citation
- MatRIS: Toward Reliable and Efficient Pretrained Machine Learning Interatomic PotentialsYuanchang Zhou, Siyu Hu, Xiangyu Zhang, Hongyu Wang et al.ICLR 2026 · 9 citations
- Physics-Informed Weakly Supervised Learning For Interatomic PotentialsMakoto Takamoto, Viktor Zaverkin, Mathias NiepertICML 2025
- DistMLIP: A Distributed Inference Platform for Machine Learning Interatomic PotentialsKevin Han, Bowen Deng, Amir Barati Farimani, Gerbrand CederICLR 2026 · 10 citations
