Equiformer: Equivariant Graph Attention Transformer for 3D Atomistic Graphs
Yi-Lun Liao, Tess E. Smidt
摘要
Despite their widespread success in various domains, Transformer networks have yet to perform well across datasets in the domain of 3D atomistic graphs such as molecules even when 3D-related inductive biases like translational invariance and rotational equivariance are considered. In this paper, we demonstrate that Transformers can generalize well to 3D atomistic graphs and present Equiformer, a graph neural network leveraging the strength of Transformer architectures and incorporating SE(3)/E(3)-equivariant features based on irreducible representations (irreps). First, we propose a simple and effective architecture by only replacing original operations in Transformers with their equivariant counterparts and including tensor products. Using equivariant operations enables encoding equivariant information in channels of irreps features without complicating graph structures. With minimal modifications to Transformers, this architecture has already achieved strong empirical results. Second, we propose a novel attention mechanism called equivariant graph attention, which improves upon typical attention in Transformers through replacing dot product attention with multi-layer perceptron attention and including non-linear message passing. With these two innovations, Equiformer achieves competitive results to previous models on QM9, MD17 and OC20 datasets.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper137
- Simplifying and Empowering Transformers for Large-Graph RepresentationsQitian Wu, Wentao Zhao, Chenxiao Yang, Hengrui Zhang 等NeurIPS 2023 · 被引用 318 次
- EquiformerV2: Improved Equivariant Transformer for Scaling to Higher-Degree RepresentationsYi-Lun Liao, Brandon M. Wood, Abhishek Das, Tess E. SmidtICLR 2024 · 被引用 311 次
- Equivariant flow matchingLeon Klein, Andreas Krämer, Frank NoéNeurIPS 2023 · 被引用 169 次
- Reducing SO(3) Convolutions to SO(2) for Efficient Equivariant GNNsSaro Passaro, C. Lawrence ZitnickICML 2023 · 被引用 157 次
- Geometric Algebra TransformerJohann Brehmer, Pim de Haan, Sönke Behrends, Taco S. CohenNeurIPS 2023 · 被引用 81 次
它引用的顶会 Paper24
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah 等NeurIPS 2020 · 被引用 64,255 次
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- wav2vec 2.0: A Framework for Self-Supervised Learning of Speech RepresentationsAlexei Baevski, Yuhao Zhou, Abdelrahman Mohamed, Michael AuliNeurIPS 2020 · 被引用 9,451 次
- Training data-efficient image transformers & distillation through attentionHugo Touvron, Matthieu Cord, Matthijs Douze, Francisco Massa 等ICML 2021 · 被引用 8,974 次
- Open Graph Benchmark: Datasets for Machine Learning on GraphsWeihua Hu, Matthias Fey, Marinka Zitnik, Yuxiao Dong 等NeurIPS 2020 · 被引用 3,935 次
相关 Paper
- Geometric Transformer with Interatomic Positional EncodingYusong Wang, Shaoning Li, Tong Wang, Bin Shao 等NeurIPS 2023 · 被引用 25 次
- GotenNet: Rethinking Efficient 3D Equivariant Graph Neural NetworksSarp Aykent, Tian XiaICLR 2025
- GeoMFormer: A General Architecture for Geometric Molecular Representation LearningTianlang Chen, Shengjie Luo, Di He, Shuxin Zheng 等ICML 2024 · 被引用 9 次
- Generalist Equivariant Transformer Towards 3D Molecular Interaction LearningXiangzhe Kong, Wenbing Huang, Yang LiuICML 2024 · 被引用 29 次
- Periodic Graph Transformers for Crystal Material Property PredictionKeqiang Yan, Yi Liu, Yuchao Lin, Shuiwang JiNeurIPS 2022 · 被引用 167 次
