A recipe for scalable attention-based ML potentials: unlocking long-range accuracy with all-to-all node attention
Eric Qu, Brandon Wood, Aditi Krishnapriyan, Zachary Ulissi
摘要
Machine-learning interatomic potentials (MLIPs) have advanced rapidly, with many top models relying on strong physics-based inductive biases. However, as models scale to larger systems like biomolecules and electrolytes, they struggle to accurately capture long-range (LR) interactions, leading current approaches to rely on explicit physics-based terms or components. In this work, we propose AllScAIP, a straightforward, attention-based, and energy-conserving MLIP model that scales to O(100 million) training samples. It addresses the long-range challenge using an all-to-all node attention component that is data-driven. Extensive ablations reveal that in low-data/small-model regimes, inductive biases improve sample efficiency. However, as data and model size scale, these benefits diminish or even reverse, while all-to-all attention remains critical for capturing LR interactions. Our model achieves state-of-the-art energy/force accuracy on molecular systems, as well as a number of physics-based evaluations (OMol25), while being competitive on materials (OMat24) and catalysts (OC20). Furthermore, it enables stable, long-timescale MD simulations that accurately recover experimental observables, including density and heat of vaporization predictions.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper8
- Directional Message Passing for Molecular GraphsJohannes Klicpera, Janek Groß, Stephan GünnemannICLR 2020 · 被引用 1,079 次
- GemNet: Universal Directional Graph Neural Networks for MoleculesJohannes Gasteiger, Florian Becker, Stephan GünnemannNeurIPS 2021 · 被引用 665 次
- EquiformerV2: Improved Equivariant Transformer for Scaling to Higher-Degree RepresentationsYi-Lun Liao, Brandon M. Wood, Abhishek Das, Tess E. SmidtICLR 2024 · 被引用 311 次
- Reducing SO(3) Convolutions to SO(2) for Efficient Equivariant GNNsSaro Passaro, C. Lawrence ZitnickICML 2023 · 被引用 157 次
- Spherical Channels for Modeling Atomic InteractionsLarry Zitnick, Abhishek Das, Adeesh Kolluru, Janice Lan 等NeurIPS 2022 · 被引用 84 次
相关 Paper
- The Importance of Being Scalable: Improving the Speed and Accuracy of Neural Network Interatomic Potentials Across Chemical DomainsEric Qu, Aditi S. KrishnapriyanNeurIPS 2024 · 被引用 63 次
- Smooth Dynamic Cutoffs for Machine Learning Interatomic PotentialsKevin Han, Haolin Cong, Bowen Deng, Amir Barati FarimaniICML 2026 · 被引用 1 次
- MatRIS: Toward Reliable and Efficient Pretrained Machine Learning Interatomic PotentialsYuanchang Zhou, Siyu Hu, Xiangyu Zhang, Hongyu Wang 等ICLR 2026 · 被引用 9 次
- Physics-Informed Weakly Supervised Learning For Interatomic PotentialsMakoto Takamoto, Viktor Zaverkin, Mathias NiepertICML 2025
- DistMLIP: A Distributed Inference Platform for Machine Learning Interatomic PotentialsKevin Han, Bowen Deng, Amir Barati Farimani, Gerbrand CederICLR 2026 · 被引用 10 次
