Advancing Retrosynthesis with Retrieval-Augmented Graph Generation
Anjie Qiao, Zhen Wang, Jiahua Rao, Yuedong Yang, Zhewei Wei
摘要
Diffusion-based molecular graph generative models have achieved significant success in template-free, single-step retrosynthesis prediction. However, these models typically generate reactants from scratch, often overlooking the fact that the scaffold of a product molecule typically remains unchanged during chemical reactions. To leverage this useful observation, we introduce a retrieval-augmented molecular graph generation framework. Our framework comprises three key components: a retrieval component that identifies similar molecules for the given product, an integration component that learns valuable clues from these molecules about which part of the product should remain unchanged, and a base generative model that is prompted by these clues to generate the corresponding reactants. We explore various design choices for critical and under-explored aspects of this framework and instantiate it as the Retrieval-Augmented RetroBridge (RARB). RARB demonstrates state-of-the-art performance on standard benchmarks, achieving a 14.8% relative improvement in top-1 accuracy over its base generative model, highlighting the effectiveness of retrieval augmentation. Additionally, RARB excels in handling out-of-distribution molecules, and its advantages remain significant even with smaller models or fewer denoising steps. These strengths make RARB highly valuable for real-world retrosynthesis applications, where extrapolation to novel molecules and high-throughput prediction are essential.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper22
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 被引用 13,211 次
- Scalable Diffusion Models with TransformersWilliam Peebles, Saining XieICCV 2023 · 被引用 5,568 次
- Video Diffusion ModelsJonathan Ho, Tim Salimans, Alexey A. Gritsenko, William Chan 等NeurIPS 2022 · 被引用 2,948 次
相关 Paper
- RetroBridge: Modeling Retrosynthesis with Markov BridgesIlia Igashov, Arne Schneuing, Marwin H. S. Segler, Michael M. Bronstein 等ICLR 2024 · 被引用 34 次
- Equivariant Denoisers Cannot Copy Graphs: Align Your Graph Diffusion ModelsNajwa Laabid, Severi Rissanen, Markus Heinonen, Arno Solin 等ICLR 2025
- Molecule Generation with Fragment Retrieval AugmentationSeul Lee, Karsten Kreis, Srimukh Prasad Veccham, Meng Liu 等NeurIPS 2024 · 被引用 36 次
- GDiffRetro: Retrosynthesis Prediction with Dual Graph Enhanced Molecular Representation and Diffusion GenerationShengyin Sun, Wenhao Yu, Yuxiang Ren, Weitao Du 等AAAI 2025 · 被引用 8 次
- RETRO SYNFLOW: Discrete Flow-Matching for Accurate and Diverse Single-Step RetrosynthesisRobin Yadav, Qi Yan, Guy Wolf, Joey Bose 等NeurIPS 2025 · 被引用 8 次
