EquiformerV2: Improved Equivariant Transformer for Scaling to Higher-Degree Representations
Yi-Lun Liao, Brandon M. Wood, Abhishek Das, Tess E. Smidt
Abstract
Equivariant Transformers such as Equiformer have demonstrated the efficacy of applying Transformers to the domain of 3D atomistic systems. However, they are limited to small degrees of equivariant representations due to their computational complexity. In this paper, we investigate whether these architectures can scale well to higher degrees. Starting from Equiformer, we first replace SOp3q convolutions with eSCN convolutions to efficiently incorporate higher-degree tensors. Then, to better leverage the power of higher degrees, we propose three architectural improvements -attention re-normalization, separable S 2 activation and separable layer normalization. Putting this all together, we propose EquiformerV2, which outperforms previous state-of-the-art methods on large-scale OC20 dataset by up to 9% on forces, 4% on energies, offers better speed-accuracy trade-offs, and 2r eduction in DFT calculations needed for computing adsorption energies. Additionally, EquiformerV2 trained on only OC22 dataset outperforms GemNet-OC trained on both OC20 and OC22 datasets, achieving much better data efficiency. Finally, we compare EquiformerV2 with Equiformer on QM9 and OC20 S2EF-2M datasets to better understand the performance gain brought by higher degrees. INTRODUCTION In recent years, machine learning (ML) models have shown promising results in accelerating and scaling high-accuracy but compute-intensive quantum mechanical calculations by effectively accounting for key features of atomic systems, such as the discrete nature of atoms, and Euclidean and permutation symmetries (
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 2cbe5bc8-cb83-4ce3-aaff-331c3a713a6eCited by top-tier papers88
- UMA: A Family of Universal Models for AtomsBrandon M. Wood, Misko Dzamba, Xiang Fu, Meng Gao et al.NeurIPS 2025 · 282 citations
- FlowMM: Generating Materials with Riemannian Flow MatchingBenjamin Kurt Miller, Ricky T. Q. Chen, Anuroop Sriram, Brandon M. WoodICML 2024 · 101 citations
- FlowLLM: Flow Matching for Material Generation with Large Language Models as Base DistributionsAnuroop Sriram, Benjamin Kurt Miller, Ricky T. Q. Chen, Brandon M. WoodNeurIPS 2024 · 78 citations
- The Importance of Being Scalable: Improving the Speed and Accuracy of Neural Network Interatomic Potentials Across Chemical DomainsEric Qu, Aditi S. KrishnapriyanNeurIPS 2024 · 63 citations
- ET-Flow: Equivariant Flow-Matching for Molecular Conformer GenerationMajdi Hassan, Nikhil Shenoy, Jungyoon Lee, Hannes Stärk et al.NeurIPS 2024 · 50 citations
Builds on23
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah et al.NeurIPS 2020 · 64,255 citations
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Training data-efficient image transformers & distillation through attentionHugo Touvron, Matthieu Cord, Matthijs Douze, Francisco Massa et al.ICML 2021 · 8,974 citations
- Do Transformers Really Perform Badly for Graph Representation?Chengxuan Ying, Tianle Cai, Shengjie Luo, Shuxin Zheng et al.NeurIPS 2021 · 1,632 citations
- MACE: Higher Order Equivariant Message Passing Neural Networks for Fast and Accurate Force FieldsIlyes Batatia, Dávid Péter Kovács, Gregor N. C. Simm, Christoph Ortner et al.NeurIPS 2022 · 1,448 citations
Related papers
- Equiformer: Equivariant Graph Attention Transformer for 3D Atomistic GraphsYi-Lun Liao, Tess E. SmidtICLR 2023 · 65 citations
- Equivariant Transformers for Neural Network based Molecular PotentialsPhilipp Thölke, Gianni De FabritiisICLR 2022 · 277 citations
- E2Former-V2: On-the-Fly Equivariant Attention with Linear Activation MemoryLin Huang, Chengxiang Huang, Ziang Wang, Yiyue Du et al.ICML 2026 · 3 citations
- E2Former: An Efficient and Equivariant Transformer with Linear-Scaling Tensor ProductsYunyang Li, Lin Huang, Zhihao Ding, Xinran Wei et al.NeurIPS 2025 · 17 citations
- GotenNet: Rethinking Efficient 3D Equivariant Graph Neural NetworksSarp Aykent, Tian XiaICLR 2025
