GTA: Graph Truncated Attention for Retrosynthesis
Seung-Woo Seo, You Young Song, June Yong Yang, Seohui Bae, Hankook Lee, Jinwoo Shin, Sung Ju Hwang, Eunho Yang
Abstract
Retrosynthesis is the task of predicting reactant molecules from a given product molecule and is, important in organic chemistry because the identification of a synthetic path is as demanding as the discovery of new chemical compounds. Recently, the retrosynthesis task has been solved automatically without human expertise using powerful deep learning models. Recent deep models are primarily based on seq2seq or graph neural networks depending on the function of molecular representation, sequence, or graph. Current state-of-theart models represent a molecule as a graph, but they require joint training with auxiliary prediction tasks, such as the most probable reaction template or reaction center prediction. Furthermore, they require additional labels by experienced chemists, thereby incurring additional cost. Herein, we propose a novel template-free model, i.e., Graph Truncated Attention (GTA), which leverages both sequence and graph representations by inserting graphical information into a seq2seq model. The proposed GTA model masks the self-attention layer using the adjacency matrix of product molecule in the encoder and applies a new loss using atom mapping acquired from an automated algorithm to the cross-attention layer in the decoder. Our model achieves new state-of-the-art records, i.e., exact match top-1 and top-10 accuracies of 51.1 % and 81.6 % on the USPTO-50k benchmark dataset, respectively, and 46.0 % and 70.0 % on the USPTO-full dataset, respectively, both without any reaction class information. The GTA model surpasses prior graph-based template-free models by 2 % and 7 % in terms of the top-1 and top-10 accuracies on the USPTO-50k dataset, respectively, and by over 6 % for both the top-1 and top-10 accuracies on the USPTO-full dataset.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 375230ba-2b79-470b-8837-a12e6980105dCited by top-tier papers9
- RetroBridge: Modeling Retrosynthesis with Markov BridgesIlia Igashov, Arne Schneuing, Marwin H. S. Segler, Michael M. Bronstein et al.ICLR 2024 · 34 citations
- RetroLens: A Human-AI Collaborative System for Multi-step Retrosynthetic Route PlanningChuhan Shi, Yicheng Hu, Shenan Wang, Shuai Ma et al.CHI 2023 · 21 citations
- Retrosynthesis Prediction with Local Template RetrievalShufang Xie, Rui Yan, Junliang Guo, Yingce Xia et al.AAAI 2023 · 20 citations
- Learning Chemical Rules of Retrosynthesis with Pre-trainingYinjie Jiang, Ying Wei, Fei Wu, Zhengxing Huang et al.AAAI 2023 · 12 citations
- GDiffRetro: Retrosynthesis Prediction with Dual Graph Enhanced Molecular Representation and Diffusion GenerationShengyin Sun, Wenhao Yu, Yuxiang Ren, Weitao Du et al.AAAI 2025 · 8 citations
Builds on1
Related papers
- Retroformer: Pushing the Limits of End-to-end Retrosynthesis TransformerYue Wan, Chang-Yu Hsieh, Ben Liao, Shengyu ZhangICML 2022 · 13 citations
- RetroXpert: Decompose Retrosynthesis Prediction Like A ChemistChaochao Yan, Qianggang Ding, Peilin Zhao, Shuangjia Zheng et al.NeurIPS 2020 · 151 citations
- Learning Graph Models for Retrosynthesis PredictionVignesh Ram Somnath, Charlotte Bunne, Connor W. Coley, Andreas Krause et al.NeurIPS 2021 · 137 citations
- Order Matters in Retrosynthesis: Structure-aware Generation via Reaction-Center-Guided Discrete Flow MatchingChenguang Wang, Zihan Zhou, LEI BAI, Tianshu YuICML 2026 · 1 citation
- Towards understanding retrosynthesis by energy-based modelsRuoxi Sun, Hanjun Dai, Li Li, Steven Kearnes et al.NeurIPS 2021 · 45 citations
