Property-Driven Protein Inverse Folding with Multi-Objective Preference Alignment
Junqi Liu, Xiaoyang Hou, Chence Shi, Xin Liu, Zhi Yang, Jian Tang
摘要
Protein sequence design must balance designability, defined as the ability to recover a target backbone, with multiple, often competing, developability properties such as solubility, thermostability, and expression. Existing approaches address these properties through post hoc mutation, inference-time biasing, or retraining on property-specific subsets, yet they are target dependent and demand substantial domain expertise or careful hyperparameter tuning. In this paper, we introduce Pro-tAlign, a multi-objective preference alignment framework that fine-tunes pretrained inverse folding models to satisfy diverse developability objectives while preserving structural fidelity. ProtAlign employs a semi-online Direct Preference Optimization strategy with a flexible preference margin to mitigate conflicts among competing objectives and constructs preference pairs using in silico property predictors. Applied to the widely used ProteinMPNN backbone, the resulting model MoMPNN enhances developability without compromising designability across tasks including sequence design for CATH 4.3 crystal structures, de novo generated backbones, and real-world binder design scenarios, making it an appealing framework for practical protein sequence design.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper18
- Direct Preference Optimization: Your Language Model is Secretly a Reward ModelRafael Rafailov, Archit Sharma, Eric Mitchell, Christopher D. Manning 等NeurIPS 2023 · 被引用 10,924 次
- Model soups: averaging weights of multiple fine-tuned models improves accuracy without increasing inference timeMitchell Wortsman, Gabriel Ilharco, Samir Yitzhak Gadre, Rebecca Roelofs 等ICML 2022 · 被引用 1,464 次
- SimPO: Simple Preference Optimization with a Reference-Free RewardYu Meng, Mengzhou Xia, Danqi ChenNeurIPS 2024 · 被引用 1,203 次
- Learning from Protein Structure with Geometric Vector PerceptronsBowen Jing, Stephan Eismann, Patricia Suriana, Raphael John Lamarre Townshend 等ICLR 2021 · 被引用 627 次
- Learning inverse folding from millions of predicted structuresChloe Hsu, Robert Verkuil, Jason Liu, Zeming Lin 等ICML 2022 · 被引用 560 次
相关 Paper
- DualMPNN: Harnessing Structural Alignments for High-Recovery Inverse Protein FoldingXuhui Liao, Qiyu Wang, Zhiqiang Liang, Liwei Xiao 等NeurIPS 2025 · 被引用 2 次
- Multi-state Protein Sequence Design with DynamicMPNNAlex Abrudan, Sebastian Pujalte Ojeda, Chaitanya K. Joshi, Matthew Greenig 等ICLR 2026 · 被引用 5 次
- Protein Inverse Folding From Structure FeedbackJunde Xu, Zijun Gao, Xinyi Zhou, Jie Hu 等NeurIPS 2025 · 被引用 9 次
- Advancing Protein Design via Multi-Agent Reinforcement Learning with Pareto-Based Collaborative OptimizationMingming Zhu, Jiahua Rao, Xiaoyu Chen, Qianmu Yuan 等AAAI 2026 · 被引用 1 次
- Discrete Diffusion Trajectory Alignment via Stepwise DecompositionJiaqi Han, Austin Wang, Minkai Xu, Wenda Chu 等ICLR 2026 · 被引用 11 次
