One Transformer Can Understand Both 2D & 3D Molecular Data
Shengjie Luo, Tianlang Chen, Yixian Xu, Shuxin Zheng, Tie-Yan Liu, Liwei Wang, Di He
Abstract
Unlike vision and language data which usually has a unique format, molecules can naturally be characterized using different chemical formulations. One can view a molecule as a 2D graph or define it as a collection of atoms located in a 3D space. For molecular representation learning, most previous works designed neural networks only for a particular data format, making the learned models likely to fail for other data formats. We believe a general-purpose neural network model for chemistry should be able to handle molecular tasks across data modalities. To achieve this goal, in this work, we develop a novel Transformer-based Molecular model called Transformer-M, which can take molecular data of 2D or 3D formats as input and generate meaningful semantic representations. Using the standard Transformer as the backbone architecture, Transformer-M develops two separated channels to encode 2D and 3D structural information and incorporate them with the atom features in the network modules. When the input data is in a particular format, the corresponding channel will be activated, and the other will be disabled. By training on 2D and 3D molecular data with properly designed supervised signals, Transformer-M automatically learns to leverage knowledge from different data modalities and correctly capture the representations. We conducted extensive experiments for Transformer-M. All empirical results show that Transformer-M can simultaneously achieve strong performance on 2D and 3D tasks, suggesting its broad applicability. The code and models will be made publicly available at https://github.com/lsj2408/Transformer-M .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 812efee2-3f2c-4156-8f07-740bf25391f3Cited by top-tier papers43
- Enabling Efficient Equivariant Operations in the Fourier Basis via Gaunt Tensor ProductsShengjie Luo, Tianlang Chen, Aditi S. KrishnapriyanICLR 2024 · 42 citations
- Towards Foundational Models for Molecular Learning on Large-Scale Multi-Task DatasetsDominique Beaini, Shenyang Huang, Joao Alex Cunha, Zhiyi Li et al.ICLR 2024 · 39 citations
- Learning Probabilistic Symmetrization for Architecture Agnostic EquivarianceJinwoo Kim, Dat Nguyen, Ayhan Suleymanzade, Hyeokjun An et al.NeurIPS 2023 · 32 citations
- Automated 3D Pre-Training for Molecular Property PredictionXu Wang, Huan Zhao, Wei-Wei Tu, Quanming YaoKDD 2023 · 28 citations
- Multimodal Molecular Pretraining via Modality BlendingQiying Yu, Yudi Zhang, Yuyan Ni, Shikun Feng et al.ICLR 2024 · 26 citations
Builds on32
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah et al.NeurIPS 2020 · 64,255 citations
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu et al.ICCV 2021 · 31,683 citations
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 24,064 citations
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
Related papers
- Unified 2D and 3D Pre-Training of Molecular RepresentationsJinhua Zhu, Yingce Xia, Lijun Wu, Shufang Xie et al.KDD 2022 · 53 citations
- Dual-view Molecular Pre-trainingJinhua Zhu, Yingce Xia, Lijun Wu, Shufang Xie et al.KDD 2023 · 47 citations
- Uni-Mol: A Universal 3D Molecular Representation Learning FrameworkGengmo Zhou, Zhifeng Gao, Qiankun Ding, Hang Zheng et al.ICLR 2023 · 254 citations
- GeoMFormer: A General Architecture for Geometric Molecular Representation LearningTianlang Chen, Shengjie Luo, Di He, Shuxin Zheng et al.ICML 2024 · 9 citations
- Learning Multi-view Molecular Representations with Structured and Unstructured KnowledgeYizhen Luo, Kai Yang, Massimo Hong, Xing Yi Liu et al.KDD 2024 · 9 citations
