Preference-based Antibody Expression Ranking: Scaling with Large-scale Weak Supervision
Josh Sun, Morteza Babaie, Wenyang hou, Mark Crowley, David Young
Abstract
Antibody expression ranking is a critical task in antibody design, yet its modelling is severely hindered by the scarcity of labeled experimental data. To address this, we propose a unified preference-based learning framework that integrates scarce quantitative expression data with large-scale weak positive supervision from immunization data. We adapt Direct Preference Optimization (DPO) to protein language models by introducing a union-masked log-likelihood approximation and IMGT-based alignment, enabling efficient training on variable-length sequences. Evaluating on a diverse internal dataset of 1254 labeled sequences and 4 million unlabeled camelid-derived antibodies, we show that our method consistently outperforms baselines on most metrics. Our results demonstrate that preference learning can effectively learn from weak supervision, providing a scalable solution for antibody expressibility optimization in data-constrained settings. Project page: https://kisoji-biotechnology-inc.github.io/Preference-Expression-Ranking/.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 862d4578-6cd8-4ef9-a7cd-816f4933207bBuilds on5
- Direct Preference Optimization: Your Language Model is Secretly a Reward ModelRafael Rafailov, Archit Sharma, Eric Mitchell, Christopher D. Manning et al.NeurIPS 2023 · 10,924 citations
- Language models enable zero-shot prediction of the effects of mutations on protein functionJoshua Meier, Roshan Rao, Robert Verkuil, Jason Liu et al.NeurIPS 2021 · 969 citations
- Masked Language Model ScoringJulian Salazar, Davis Liang, Toan Q. Nguyen, Katrin KirchhoffACL 2020 · 167 citations
- Protein Inverse Folding From Structure FeedbackJunde Xu, Zijun Gao, Xinyi Zhou, Jie Hu et al.NeurIPS 2025 · 9 citations
- Protein Language Model Fitness is a Matter of PreferenceCade W. Gordon, Amy X. Lu, Pieter AbbeelICLR 2025
Related papers
- Pareto-Optimal Energy Alignment for Designing Nature-Like AntibodiesYibo Wen, Chenwei Xu, Jerry Yao-Chieh Hu, Kaize Ding et al.NeurIPS 2025
- Enhancing Safe and Controllable Protein Generation via Knowledge Preference OptimizationYuhao Wang, Keyan Ding, Kehua Feng, Zeyuan Wang et al.ACL 2025 · 2 citations
- Reprogramming Pretrained Language Models for Antibody Sequence InfillingIgor Melnyk, Vijil Chenthamarakshan, Pin-Yu Chen, Payel Das et al.ICML 2023 · 40 citations
- Uni-DPO: A Unified Paradigm for Dynamic Preference Optimization of LLMsShangpin Peng, Weinong Wang, Zhuotao Tian, Senqiao Yang et al.ICLR 2026 · 10 citations
- Antigen-Specific Antibody Design via Direct Energy-based Preference OptimizationXiangxin Zhou, Dongyu Xue, Ruizhe Chen, Zaixiang Zheng et al.NeurIPS 2024 · 48 citations
