ProteinBench: A Holistic Evaluation of Protein Foundation Models
Fei Ye, Zaixiang Zheng, Dongyu Xue, Yuning Shen, Lihao Wang, Yiming Ma, Yan Wang, Xinyou Wang, Xiangxin Zhou, Quanquan Gu
摘要
Recent years have witnessed a surge in the development of protein foundation models, significantly improving performance in protein prediction and generative tasks ranging from 3D structure prediction and protein design to conformational dynamics. However, the capabilities and limitations associated with these models remain poorly understood due to the absence of a unified evaluation framework. To fill this gap, we introduce ProteinBench, a holistic evaluation framework designed to enhance the transparency of protein foundation models. Our approach consists of three key components: (i) A taxonomic classification of tasks that broadly encompass the main challenges in the protein domain, based on the relationships between different protein modalities; (ii) A multi-metric evaluation approach that assesses performance across four key dimensions: quality, novelty, diversity, and robustness; and (iii) In-depth analyses from various user objectives, providing a holistic view of model performance. Our comprehensive evaluation of protein foundation models reveals several key findings that shed light on their current capabilities and limitations. To promote transparency and facilitate further research, we release the evaluation dataset, code, and a leaderboard publicly for further analysis and a general modular toolkit. We intend for ProteinBench to be a living benchmark for establishing a standardized, in-depth evaluation framework for protein foundation models, driving their development and application while fostering collaboration within the field.
† Corresponding author. 1 In this study, we broaden the definition of protein foundation models to include any generative model aimed at addressing foundational problems in protein science.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper7
- Simultaneous Modeling of Protein Conformation and Dynamics via AutoregressionYuning Shen, Lihao Wang, Huizhuo Yuan, Yan Wang 等NeurIPS 2025 · 被引用 13 次
- Demystifying Multimodal Biomolecular Co-design With Intrinsic Geodesic CouplingKeyue Qiu, Xintong Wang, Zhilong Zhang, Hao Zhou 等ICML 2026
- TadA-Bench: A Million-Variant Benchmark for Future-Round Discovery Toward Agentic Protein EngineeringJin Gao, Juntu Zhao, Zirui Zeng, Jiaqi Shen 等ICML 2026
- Elucidating the Design Space of Multimodal Protein Language ModelsCheng-Yen Hsieh, Xinyou Wang, Daiheng Zhang, Dongyu Xue 等ICML 2025
- Rationalized All-Atom Protein Design with Unified Multi-Modal Bayesian FlowHanlin Wu, Yuxuan Song, Zhe Zhang, Zhilong Zhang 等NeurIPS 2025
它引用的顶会 Paper19
- Learning from Protein Structure with Geometric Vector PerceptronsBowen Jing, Stephan Eismann, Patricia Suriana, Raphael John Lamarre Townshend 等ICLR 2021 · 被引用 627 次
- Learning inverse folding from millions of predicted structuresChloe Hsu, Robert Verkuil, Jason Liu, Zeming Lin 等ICML 2022 · 被引用 560 次
- Antigen-Specific Antibody Design and Optimization with Diffusion-Based Generative Models for Protein StructuresShitong Luo, Yufeng Su, Xingang Peng, Sheng Wang 等NeurIPS 2022 · 被引用 331 次
- Generative Flows on Discrete State-Spaces: Enabling Multimodal Flows with Applications to Protein Co-DesignAndrew Campbell, Jason Yim, Regina Barzilay, Tom Rainforth 等ICML 2024 · 被引用 283 次
- Practical and Asymptotically Exact Conditional Sampling in Diffusion ModelsLuhuan Wu, Brian L. Trippe, Christian A. Naesseth, David M. Blei 等NeurIPS 2023 · 被引用 276 次
相关 Paper
- PDFBench: A Benchmark for De Novo Protein Design from FunctionJiahao Kuang, Nuowei Liu, Changzhi Sun, Jie Wang 等ICML 2026 · 被引用 10 次
- UniVBench: Towards Unified Evaluation for Video Foundation ModelsJianhui Wei, Xiaotian Zhang, Yichen Li, Yuan Wang 等CVPR 2026 · 被引用 12 次
- ProtDBench: A Unified Benchmark of Protein Binder Design and EvaluationCong Liu, Milong Ren, Jiaqi Guan, Chengyue Gong 等ICML 2026
- ProMiSE: Protein Multi-State Evaluation Benchmark in Biological ContextsBonjae Ku, Seeun Kim, Yubeen Kim, Hahnbeom Park 等ICML 2026
- Rethinking Text-based Protein Understanding: Retrieval or LLM?Juntong Wu, Zijing Liu, He Cao, Li Hao 等EMNLP 2025 · 被引用 7 次
