Order Matters in Retrosynthesis: Structure-aware Generation via Reaction-Center-Guided Discrete Flow Matching
Chenguang Wang, Zihan Zhou, LEI BAI, Tianshu Yu
Abstract
Template-free retrosynthesis methods treat the task as black-box sequence generation, limiting learning efficiency, while semi-template approaches rely on rigid reaction libraries that constrain generalization. We address this gap with a key insight: atom ordering in neural representations matters. Building on this insight, we propose a structure-aware template-free framework that encodes the two-stage nature of chemical reactions as a positional inductive bias. By placing reaction center atoms at the sequence head, our method transforms implicit chemical knowledge into explicit positional patterns that the model can readily capture. The proposed RetroDiT backbone, a graph transformer with rotary position embeddings, exploits this ordering to prioritize chemically critical regions. Combined with discrete flow matching, our approach decouples training from sampling and enables generation in 20--50 steps versus 500 for prior diffusion methods. Our method achieves state-of-the-art performance on both USPTO-50k (61.2% top-1) and the large-scale USPTO-Full (51.3% top-1) with predicted reaction centers. With oracle centers, performance reaches 71.1% and 63.4% respectively, surpassing foundation models trained on 10 billion reactions while using orders of magnitude less data. Ablation studies further reveal that structural priors outperform brute-force scaling: a 280K-parameter model with proper ordering matches a 65M-parameter model without it.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext ca3a3a7c-3c7b-4b08-9b18-1576d702606dBuilds on23
- Structured Denoising Diffusion Models in Discrete State-SpacesJacob Austin, Daniel D. Johnson, Jonathan Ho, Daniel Tarlow et al.NeurIPS 2021 · 2,256 citations
- Efficient Streaming Language Models with Attention SinksGuangxuan Xiao, Yuandong Tian, Beidi Chen, Song Han et al.ICLR 2024 · 1,714 citations
- Do Transformers Really Perform Badly for Graph Representation?Chengxuan Ying, Tianle Cai, Shengjie Luo, Shuxin Zheng et al.NeurIPS 2021 · 1,632 citations
- Score-Based Generative Modeling through Stochastic Differential EquationsYang Song, Jascha Sohl-Dickstein, Diederik P. Kingma, Abhishek Kumar et al.ICLR 2021 · 1,270 citations
- Simple and Effective Masked Diffusion Language ModelsSubham S. Sahoo, Marianne Arriola, Yair Schiff, Aaron Gokaslan et al.NeurIPS 2024 · 929 citations
Related papers
- GTA: Graph Truncated Attention for RetrosynthesisSeung-Woo Seo, You Young Song, June Yong Yang, Seohui Bae et al.AAAI 2021 · 73 citations
- Retroformer: Pushing the Limits of End-to-end Retrosynthesis TransformerYue Wan, Chang-Yu Hsieh, Ben Liao, Shengyu ZhangICML 2022 · 13 citations
- Learning Chemical Rules of Retrosynthesis with Pre-trainingYinjie Jiang, Ying Wei, Fei Wu, Zhengxing Huang et al.AAAI 2023 · 12 citations
- RetroXpert: Decompose Retrosynthesis Prediction Like A ChemistChaochao Yan, Qianggang Ding, Peilin Zhao, Shuangjia Zheng et al.NeurIPS 2020 · 151 citations
- A Graph to Graphs Framework for Retrosynthesis PredictionChence Shi, Minkai Xu, Hongyu Guo, Ming Zhang et al.ICML 2020 · 176 citations
