Dual-view Molecular Pre-training
Jinhua Zhu, Yingce Xia, Lijun Wu, Shufang Xie, Wengang Zhou, Tao Qin, Houqiang Li, Tie-Yan Liu
Abstract
Molecular pre-training, which is about to learn an effective representation for molecules on large amount of data, has attracted substantial attention in cheminformatics and bioinformatics. A molecule can be viewed as either a graph (where atoms are connected by bonds) or a SMILES sequence (where depth-first-search is applied to the molecular graph with specific rules). The Transformer and graph neural networks (GNN) are two representative methods to deal with the sequential data and the graphic data, which can globally and locally model the molecules respectively and are supposed to be complementary. In this work, we propose to leverage both representations and design a new pre-training algorithm, dual-view molecule pre-training (briefly, DVMP), that can effectively combine the strengths of both types of molecule representations. DVMP has a Transformer branch and a GNN branch, and the two branches are pre-trained to maintain the semantic consistency of molecules. After pre-training, we can use either the Transformer branch (this one is recommended according to empirical results), the GNN branch, or both for downstream tasks. DVMP is tested on 11 molecular property prediction tasks and outperforms strong baselines. Furthermore, we test DVMP on three retrosynthesis tasks and it achieves state-of-the-art results. Our code is released at https://github.com/microsoft/DVMP.
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Your agent calls
Lunesearch_papers
Free to start. No credit card required.
Terminal
Install the CLIlune papers get 30746e8e-36b9-4583-bf1d-16314a768c03Cited by top-tier papers8
- Fragment-based Pretraining and Finetuning on Molecular GraphsKha-Dinh Luong, Ambuj K. SinghNeurIPS 2023 · 36 citations
- Learning Multi-view Molecular Representations with Structured and Unstructured KnowledgeYizhen Luo, Kai Yang, Massimo Hong, Xing Yi Liu et al.KDD 2024 · 9 citations
- Preference Optimization for Molecule Synthesis with Conditional Residual Energy-based ModelsSongtao Liu, Hanjun Dai, Yue Zhao, Peng LiuICML 2024 · 8 citations
- OPTFM: A Scalable Multi-View Graph Transformer for Hierarchical Pre-Training in Combinatorial OptimizationHao Yuan, Wenli Ouyang, Changwen Zhang, Congrui Li et al.NeurIPS 2025 · 3 citations
- Understanding Oversmoothing in Diffusion-Based GNNs From the Perspective of Operator Semigroup TheoryWeichen Zhao, Chenguang Wang, Xinyan Wang, Congying Han et al.KDD 2025 · 1 citation
Related papers
- Unified 2D and 3D Pre-Training of Molecular RepresentationsJinhua Zhu, Yingce Xia, Lijun Wu, Shufang Xie et al.KDD 2022 · 53 citations
- One Transformer Can Understand Both 2D & 3D Molecular DataShengjie Luo, Tianlang Chen, Yixian Xu, Shuxin Zheng et al.ICLR 2023 · 4 citations
- Mole-BERT: Rethinking Pre-training Graph Neural Networks for MoleculesJun Xia, Chengshuai Zhao, Bozhen Hu, Zhangyang Gao et al.ICLR 2023 · 119 citations
- MOL-Mamba: Enhancing Molecular Representation with Structural & Electronic InsightsJingjing Hu, Dan Guo, Zhan Si, Deguang Liu et al.AAAI 2025 · 9 citations
- Self-Supervised Graph Transformer on Large-Scale Molecular DataYu Rong, Yatao Bian, Tingyang Xu, Weiyang Xie et al.NeurIPS 2020 · 1,113 citations
