Fast Uncovering of Protein Sequence Diversity from Structure
Luca Alessandro Silva, Barthélémy Meynard-Piganeau, Carlo Lucibello, Christoph Feinauer
摘要
We present InvMSAFold, an inverse folding method for generating protein sequences that is optimized for diversity and speed. For a given structure, InvM-SAFold generates the parameters of a probability distribution over the space of sequences with pairwise interactions, capturing the amino acid covariances observed in Multiple Sequence Alignments (MSA) of homologous proteins. This allows for the efficient generation of highly diverse protein sequences while preserving structural and functional integrity. We show that this increased diversity in sampled sequences translates into greater variability in biochemical properties, highlighting the exciting potential of our method for applications such as protein design. The orders of magnitude improvement in sampling speed compared to existing methods unlocks new possibilities for high-throughput virtual screening.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper2
相关 Paper
- SIPF: Sampling Method for Inverse Protein FoldingTianfan Fu, Jimeng SunKDD 2022 · 被引用 3 次
- Graph Denoising Diffusion for Inverse Protein FoldingKai Yi, Bingxin Zhou, Yiqing Shen, Pietro Lió 等NeurIPS 2023 · 被引用 90 次
- Importance Weighted Expectation-Maximization for Protein Sequence DesignZhenqiao Song, Lei LiICML 2023 · 被引用 19 次
- All-atom inverse protein folding through discrete flow matchingKai Yi, Kiarash Jamali, Sjors H. W. ScheresICML 2025
- DS-ProGen: A Dual-Structure Deep Language Model for Functional Protein DesignYanting Li, Zikang Wang, Jiyue Jiang, Ziqian Lin 等AAAI 2026
