Graph Denoising Diffusion for Inverse Protein Folding
Kai Yi, Bingxin Zhou, Yiqing Shen, Pietro Lió, Yuguang Wang
Abstract
Inverse protein folding is challenging due to its inherent one-to-many mapping characteristic, where numerous possible amino acid sequences can fold into a single, identical protein backbone. This task involves not only identifying viable sequences but also representing the sheer diversity of potential solutions. However, existing discriminative models, such as transformer-based auto-regressive models, struggle to encapsulate the diverse range of plausible solutions. In contrast, diffusion probabilistic models, as an emerging genre of generative approaches, offer the potential to generate a diverse set of sequence candidates for determined protein backbones. We propose a novel graph denoising diffusion model for inverse protein folding, where a given protein backbone guides the diffusion process on the corresponding amino acid residue types. The model infers the joint distribution of amino acids conditioned on the nodes' physiochemical properties and local environment. Moreover, we utilize amino acid replacement matrices for the diffusion forward process, encoding the biologically meaningful prior knowledge of amino acids from their spatial and sequential neighbors as well as themselves, which reduces the sampling space of the generative process. Our model achieves state-of-the-art performance over a set of popular baseline methods in sequence recovery and exhibits great potential in generating diverse protein sequences for a determined protein backbone structure. The code is available on https://github.com/ykiiiiii/GraDe_IF .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext cccaf931-c435-45e1-8813-65445a2b8b2fCited by top-tier papers25
- Generative Flows on Discrete State-Spaces: Enabling Multimodal Flows with Applications to Protein Co-DesignAndrew Campbell, Jason Yim, Regina Barzilay, Tom Rainforth et al.ICML 2024 · 283 citations
- Diffusion Language Models Are Versatile Protein LearnersXinyou Wang, Zaixiang Zheng, Fei Ye, Dongyu Xue et al.ICML 2024 · 113 citations
- Fast Solvers for Discrete Diffusion Models: Theory and Applications of High-Order AlgorithmsYinuo Ren, Haoxuan Chen, Yuchen Zhu, Wei Guo et al.NeurIPS 2025 · 51 citations
- Full-Atom Peptide Design based on Multi-modal Flow MatchingJiahan Li, Chaoran Cheng, Zuofan Wu, Ruihan Guo et al.ICML 2024 · 38 citations
- Generative Modelling of Structurally Constrained GraphsManuel Madeira, Clément Vignac, Dorina Thanou, Pascal FrossardNeurIPS 2024 · 20 citations
Builds on23
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 11,743 citations
- Improved Denoising Diffusion Probabilistic ModelsAlexander Quinn Nichol, Prafulla DhariwalICML 2021 · 5,234 citations
- Video Diffusion ModelsJonathan Ho, Tim Salimans, Alexey A. Gritsenko, William Chan et al.NeurIPS 2022 · 2,948 citations
Related papers
- Bridge-IF: Learning Inverse Protein Folding with Markov BridgesYiheng Zhu, Jialu Wu, Qiuyi Li, Jiahuan Yan et al.NeurIPS 2024 · 16 citations
- Generating Novel, Designable, and Diverse Protein Structures by Equivariantly Diffusing Oriented Residue CloudsYeqing Lin, Mohammed AlQuraishiICML 2023 · 105 citations
- SIPF: Sampling Method for Inverse Protein FoldingTianfan Fu, Jimeng SunKDD 2022 · 3 citations
- ProtInvTree: Deliberate Protein Inverse Folding with Reward-guided Tree SearchMengdi Liu, Xiaoxue Cheng, Zhangyang Gao, Hong Chang et al.NeurIPS 2025 · 10 citations
- Bridging Protein Sequences and Microscopy Images with Unified Diffusion ModelsDihan Zheng, Bo HuangICML 2025
