Scaling Laws and Symmetry, Evidence from Neural Force Fields
Nhat Khang Ngo, Siamak Ravanbakhsh
Abstract
We present an empirical study in the geometric task of learning interatomic potentials, which shows equivariance matters even more at larger scales; we show a clear power-law scaling behaviour with respect to data, parameters and compute with ``architecture-dependent exponents''. In particular, we observe that equivariant architectures, which leverage task symmetry, scale better than non-equivariant models. Moreover, among equivariant architectures, higher-order representations translate to better scaling exponents. Our analysis also suggests that for compute-optimal training, the data and model sizes should scale in tandem regardless of the architecture. At a high level, these results suggest that, contrary to common belief, we should not leave it to the model to discover fundamental inductive biases such as symmetry, especially as we scale, because they change the inherent difficulty of the task and its scaling laws.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers1
Ask how each one uses itBuilds on32
- E(n) Equivariant Graph Neural NetworksVictor Garcia Satorras, Emiel Hoogeboom, Max WellingICML 2021 · 1,432 citations
- Directional Message Passing for Molecular GraphsJohannes Klicpera, Janek Groß, Stephan GünnemannICLR 2020 · 1,079 citations
- Scaling Vision TransformersXiaohua Zhai, Alexander Kolesnikov, Neil Houlsby, Lucas BeyerCVPR 2022 · 767 citations
- GemNet: Universal Directional Graph Neural Networks for MoleculesJohannes Gasteiger, Florian Becker, Stephan GünnemannNeurIPS 2021 · 665 citations
- Understanding over-squashing and bottlenecks on graphs via curvatureJake Topping, Francesco Di Giovanni, Benjamin Paul Chamberlain, Xiaowen Dong et al.ICLR 2022 · 628 citations
Related papers
- Higher-Rank Irreducible Cartesian Tensors for Equivariant Message PassingViktor Zaverkin, Francesco Alesiani, Takashi Maruyama, Federico Errica et al.NeurIPS 2024 · 19 citations
- UMA: A Family of Universal Models for AtomsBrandon M. Wood, Misko Dzamba, Xiang Fu, Meng Gao et al.NeurIPS 2025 · 282 citations
- Platonic Transformers: A Solid Choice For EquivarianceMohammad Mohaiminul Islam, Rishabh Anand, David Wessels, Friso de Kruiff et al.ICML 2026 · 7 citations
- EquiformerV2: Improved Equivariant Transformer for Scaling to Higher-Degree RepresentationsYi-Lun Liao, Brandon M. Wood, Abhishek Das, Tess E. SmidtICLR 2024 · 311 citations
- Smooth, exact rotational symmetrization for deep learning on point cloudsSergey Pozdnyakov, Michele CeriottiNeurIPS 2023 · 68 citations
