One Transformer Can Understand Both 2D & 3D Molecular Data
Shengjie Luo, Tianlang Chen, Yixian Xu, Shuxin Zheng, Tie-Yan Liu, Liwei Wang, Di He
摘要
Unlike vision and language data which usually has a unique format, molecules can naturally be characterized using different chemical formulations. One can view a molecule as a 2D graph or define it as a collection of atoms located in a 3D space. For molecular representation learning, most previous works designed neural networks only for a particular data format, making the learned models likely to fail for other data formats. We believe a general-purpose neural network model for chemistry should be able to handle molecular tasks across data modalities. To achieve this goal, in this work, we develop a novel Transformer-based Molecular model called Transformer-M, which can take molecular data of 2D or 3D formats as input and generate meaningful semantic representations. Using the standard Transformer as the backbone architecture, Transformer-M develops two separated channels to encode 2D and 3D structural information and incorporate them with the atom features in the network modules. When the input data is in a particular format, the corresponding channel will be activated, and the other will be disabled. By training on 2D and 3D molecular data with properly designed supervised signals, Transformer-M automatically learns to leverage knowledge from different data modalities and correctly capture the representations. We conducted extensive experiments for Transformer-M. All empirical results show that Transformer-M can simultaneously achieve strong performance on 2D and 3D tasks, suggesting its broad applicability. The code and models will be made publicly available at https://github.com/lsj2408/Transformer-M .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper43
- Enabling Efficient Equivariant Operations in the Fourier Basis via Gaunt Tensor ProductsShengjie Luo, Tianlang Chen, Aditi S. KrishnapriyanICLR 2024 · 被引用 42 次
- Towards Foundational Models for Molecular Learning on Large-Scale Multi-Task DatasetsDominique Beaini, Shenyang Huang, Joao Alex Cunha, Zhiyi Li 等ICLR 2024 · 被引用 39 次
- Learning Probabilistic Symmetrization for Architecture Agnostic EquivarianceJinwoo Kim, Dat Nguyen, Ayhan Suleymanzade, Hyeokjun An 等NeurIPS 2023 · 被引用 32 次
- Automated 3D Pre-Training for Molecular Property PredictionXu Wang, Huan Zhao, Wei-Wei Tu, Quanming YaoKDD 2023 · 被引用 28 次
- Multimodal Molecular Pretraining via Modality BlendingQiying Yu, Yudi Zhang, Yuyan Ni, Shikun Feng 等ICLR 2024 · 被引用 26 次
它引用的顶会 Paper32
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah 等NeurIPS 2020 · 被引用 64,255 次
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu 等ICCV 2021 · 被引用 31,683 次
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 被引用 24,064 次
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
相关 Paper
- Unified 2D and 3D Pre-Training of Molecular RepresentationsJinhua Zhu, Yingce Xia, Lijun Wu, Shufang Xie 等KDD 2022 · 被引用 53 次
- Dual-view Molecular Pre-trainingJinhua Zhu, Yingce Xia, Lijun Wu, Shufang Xie 等KDD 2023 · 被引用 47 次
- Uni-Mol: A Universal 3D Molecular Representation Learning FrameworkGengmo Zhou, Zhifeng Gao, Qiankun Ding, Hang Zheng 等ICLR 2023 · 被引用 254 次
- GeoMFormer: A General Architecture for Geometric Molecular Representation LearningTianlang Chen, Shengjie Luo, Di He, Shuxin Zheng 等ICML 2024 · 被引用 9 次
- Learning Multi-view Molecular Representations with Structured and Unstructured KnowledgeYizhen Luo, Kai Yang, Massimo Hong, Xing Yi Liu 等KDD 2024 · 被引用 9 次
