The Importance of Being Scalable: Improving the Speed and Accuracy of Neural Network Interatomic Potentials Across Chemical Domains
Eric Qu, Aditi S. Krishnapriyan
摘要
Scaling has been critical in improving model performance and generalization in machine learning. It involves how a model's performance changes with increases in model size or input data, as well as how efficiently computational resources are utilized to support this growth. Despite successes in other areas, the study of scaling in Neural Network Interatomic Potentials (NNIPs) remains limited. NNIPs act as surrogate models for ab initio quantum mechanical calculations. The dominant paradigm here is to incorporate many physical domain constraints into the model, such as rotational equivariance. We contend that these complex constraints inhibit the scaling ability of NNIPs, and are likely to lead to performance plateaus in the long run. In this work, we take an alternative approach and start by systematically studying NNIP scaling strategies. Our findings indicate that scaling the model through attention mechanisms is efficient and improves model expressivity. These insights motivate us to develop an NNIP architecture designed for scalability: the Efficiently Scaled Attention Interatomic Potential (EScAIP). EScAIP leverages a multi-head self-attention formulation within graph neural networks, applying attention at the neighbor-level representations. Implemented with highly-optimized attention GPU kernels, EScAIP achieves substantial gains in efficiency--at least 10x faster inference, 5x less memory usage--compared to existing NNIPs. EScAIP also achieves state-of-the-art performance on a wide range of datasets including catalysts (OC20 and OC22), molecules (SPICE), and materials (MPTrj). We emphasize that our approach should be thought of as a philosophy rather than a specific model, representing a proof-of-concept for developing general-purpose NNIPs that achieve better expressivity through scaling, and continue to scale efficiently with increased computational resources and training data.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper17
- E2Former: An Efficient and Equivariant Transformer with Linear-Scaling Tensor ProductsYunyang Li, Lin Huang, Zhihao Ding, Xinran Wei 等NeurIPS 2025 · 被引用 17 次
- Probing Equivariance and Symmetry Breaking in Convolutional NetworksSharvaree Vadgama, Mohammad Mohaiminul Islam, Domas Buracas, Christian Shewmake 等NeurIPS 2025 · 被引用 15 次
- A recipe for scalable attention-based ML potentials: unlocking long-range accuracy with all-to-all node attentionEric Qu, Brandon Wood, Aditi Krishnapriyan, Zachary UlissiICML 2026 · 被引用 14 次
- MatRIS: Toward Reliable and Efficient Pretrained Machine Learning Interatomic PotentialsYuanchang Zhou, Siyu Hu, Xiangyu Zhang, Hongyu Wang 等ICLR 2026 · 被引用 9 次
- Scaling Laws and Symmetry, Evidence from Neural Force FieldsNhat Khang Ngo, Siamak RavanbakhshICLR 2026 · 被引用 7 次
它引用的顶会 Paper19
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- Directional Message Passing for Molecular GraphsJohannes Klicpera, Janek Groß, Stephan GünnemannICLR 2020 · 被引用 1,079 次
- Scaling Vision TransformersXiaohua Zhai, Alexander Kolesnikov, Neil Houlsby, Lucas BeyerCVPR 2022 · 被引用 767 次
- GemNet: Universal Directional Graph Neural Networks for MoleculesJohannes Gasteiger, Florian Becker, Stephan GünnemannNeurIPS 2021 · 被引用 665 次
- EquiformerV2: Improved Equivariant Transformer for Scaling to Higher-Degree RepresentationsYi-Lun Liao, Brandon M. Wood, Abhishek Das, Tess E. SmidtICLR 2024 · 被引用 311 次
相关 Paper
- DistMLIP: A Distributed Inference Platform for Machine Learning Interatomic PotentialsKevin Han, Bowen Deng, Amir Barati Farimani, Gerbrand CederICLR 2026 · 被引用 10 次
- Smooth Dynamic Cutoffs for Machine Learning Interatomic PotentialsKevin Han, Haolin Cong, Bowen Deng, Amir Barati FarimaniICML 2026 · 被引用 1 次
- Injecting Domain Knowledge from Empirical Interatomic Potentials to Neural Networks for Predicting Material PropertiesZeren Shui, Daniel S. Karls, Mingjian Wen, Ilia A. Nikiforov 等NeurIPS 2022 · 被引用 10 次
- E2Former-V2: On-the-Fly Equivariant Attention with Linear Activation MemoryLin Huang, Chengxiang Huang, Ziang Wang, Yiyue Du 等ICML 2026 · 被引用 3 次
- Equivariant Atomic and Lattice Modeling Using Geometric Deep Learning for Crystal Structure OptimizationZiduo Yang, Yiming Zhao, Xian Wang, Wei Zhuo 等AAAI 2026 · 被引用 1 次
