De novo Protein Design Using Geometric Vector Field Networks
Weian Mao, Muzhi Zhu, Zheng Sun, Shuaike Shen, Lin Yuanbo Wu, Hao Chen, Chunhua Shen
摘要
Innovations like protein diffusion have enabled significant progress in de novo protein design, which is a vital topic in life science. These methods typically depend on protein structure encoders to model residue backbone frames, where atoms do not exist. Most prior encoders rely on atom-wise features, such as angles and distances between atoms, which are not available in this context. Thus far, only several simple encoders, such as IPA (Jumper et al., 2021) , have been proposed for this scenario, exposing the frame modeling as a bottleneck. In this work, we proffer the Vector Field Network (VFN), which enables network layers to perform learnable vector computations between coordinates of frame-anchored virtual atoms, thus achieving a higher capability for modeling frames. The vector computation operates in a manner similar to a linear layer, with each input channel receiving 3D virtual atom coordinates instead of scalar values. The multiple feature vectors output by the vector computation are then used to update the residue representations and virtual atom coordinates via attention aggregation. Remarkably, VFN also excels in modeling both frames and atoms, as the real atoms can be treated as the virtual atoms for modeling, positioning VFN as a potential universal encoder. In protein diffusion (frame modeling), VFN exhibits an impressive performance advantage over IPA, excelling in terms of both designability (67.04% vs. 53.58%) and diversity (66.54% vs. 51.98%). In inverse folding (frame and atom modeling), VFN outperforms the previous SoTA model, PiFold (54.7% vs. 51.66%), on sequence recovery rate. We also propose a method of equipping VFN with the ESM model (Lin et al., 2023) , which significantly surpasses the previous ESM-based SoTA (62.67% vs. 55.65%), LM-Design (Zheng et al., 2023), by a substantial margin. * WM, ZS and MZ contributed equally. Work was done when WM was visiting Zhejiang University.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper9
- UniIF: Unified Molecule Inverse FoldingZhangyang Gao, Jue Wang, Cheng Tan, Lirong Wu 等NeurIPS 2024 · 被引用 17 次
- Bridge-IF: Learning Inverse Protein Folding with Markov BridgesYiheng Zhu, Jialu Wu, Qiuyi Li, Jiahuan Yan 等NeurIPS 2024 · 被引用 16 次
- ProtInvTree: Deliberate Protein Inverse Folding with Reward-guided Tree SearchMengdi Liu, Xiaoxue Cheng, Zhangyang Gao, Hong Chang 等NeurIPS 2025 · 被引用 10 次
- PDFBench: A Benchmark for De Novo Protein Design from FunctionJiahao Kuang, Nuowei Liu, Changzhi Sun, Jie Wang 等ICML 2026 · 被引用 10 次
- PRISM: Enhancing PRotein Inverse Folding through Fine- Grained Retrieval on Structure-Sequence Multimodal RepresentationsSazan Mahbub, Souvik Kundu, Eric P. XingICLR 2026 · 被引用 6 次
它引用的顶会 Paper9
- SE(3)-Transformers: 3D Roto-Translation Equivariant Attention NetworksFabian Fuchs, Daniel E. Worrall, Volker Fischer, Max WellingNeurIPS 2020 · 被引用 1,025 次
- Learning from Protein Structure with Geometric Vector PerceptronsBowen Jing, Stephan Eismann, Patricia Suriana, Raphael John Lamarre Townshend 等ICLR 2021 · 被引用 627 次
- Learning inverse folding from millions of predicted structuresChloe Hsu, Robert Verkuil, Jason Liu, Zeming Lin 等ICML 2022 · 被引用 560 次
- SE(3) diffusion model with application to protein backbone generationJason Yim, Brian L. Trippe, Valentin De Bortoli, Emile Mathieu 等ICML 2023 · 被引用 313 次
- ComENet: Towards Complete and Efficient Message Passing for 3D Molecular GraphsLimei Wang, Yi Liu, Yuchao Lin, Haoran Liu 等NeurIPS 2022 · 被引用 130 次
相关 Paper
- Generating Novel, Designable, and Diverse Protein Structures by Equivariantly Diffusing Oriented Residue CloudsYeqing Lin, Mohammed AlQuraishiICML 2023 · 被引用 105 次
- EVA: Geometric Inverse Design for Fast Protein Motif-Scaffolding with Coupled FlowYufei Huang, Yunshu Liu, Lirong Wu, Haitao Lin 等ICLR 2025
- CarbonNovo: Joint Design of Protein Structure and Sequence Using a Unified Energy-based ModelMilong Ren, Tian Zhu, Haicang ZhangICML 2024 · 被引用 13 次
- Sequence-Augmented SE(3)-Flow Matching For Conditional Protein GenerationGuillaume Huguet, James Vuckovic, Kilian Fatras, Eric Thibodeau-Laufer 等NeurIPS 2024 · 被引用 32 次
- DecompDiff: Diffusion Models with Decomposed Priors for Structure-Based Drug DesignJiaqi Guan, Xiangxin Zhou, Yuwei Yang, Yu Bao 等ICML 2023 · 被引用 115 次
