EPO: Diverse and Realistic Protein Ensemble Generation via Energy Preference Optimization
Yuancheng Sun, Yuxuan Ren, Zhaoming Chen, Xu Han, Kang Liu, Qiwei Ye
Abstract
Accurate exploration of protein conformational ensembles is essential for uncovering function but remains hard because molecular-dynamics (MD) simulations suffer from high computational costs and energy-barrier trapping. This paper presents Energy Preference Optimization (EPO), an online refinement algorithm that turns a pretrained protein ensemble generator into an energy-aware sampler without extra MD trajectories. Specifically, EPO leverages stochastic differential equation sampling to explore the conformational landscape and incorporates a novel energy-ranking mechanism based on list-wise preference optimization. Crucially, EPO introduces a practical upper bound to efficiently approximate the intractable probability of long sampling trajectories in continuous-time generative models, making it easily adaptable to existing pretrained generators. On Tetrapeptides, AT-LAS, and Fast-Folding benchmarks, EPO successfully generates diverse and physically realistic ensembles, establishing a new state-of-the-art in nine evaluation metrics. These results demonstrate that energy-only preference signals can efficiently steer generative models toward thermodynamically consistent conformational ensembles, providing an alternative to long MD simulations and widening the applicability of learned potentials in structural biology and drug discovery.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 1e5002ff-514b-467b-b9db-0f52628b48a8Builds on15
- LoRA: Low-Rank Adaptation of Large Language ModelsEdward J. Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu et al.ICLR 2022 · 18,833 citations
- Direct Preference Optimization: Your Language Model is Secretly a Reward ModelRafael Rafailov, Archit Sharma, Eric Mitchell, Christopher D. Manning et al.NeurIPS 2023 · 10,924 citations
- Flow-GRPO: Training Flow Matching Models via Online RLJie Liu, Gongye Liu, Jiajun Liang, Yangguang Li et al.NeurIPS 2025 · 647 citations
- SE(3) diffusion model with application to protein backbone generationJason Yim, Brian L. Trippe, Valentin De Bortoli, Emile Mathieu et al.ICML 2023 · 313 citations
- AlphaFold Meets Flow Matching for Generating Protein EnsemblesBowen Jing, Bonnie Berger, Tommi S. JaakkolaICML 2024 · 229 citations
Related papers
- Aligning Protein Conformation Ensemble Generation with Physical FeedbackJiarui Lu, Xiaoyin Chen, Stephen Zhewen Lu, Aurélie C. Lozano et al.ICML 2025
- Protein Conformation Generation via Force-Guided SE(3) Diffusion ModelsYan Wang, Lihao Wang, Yuning Shen, Yiqun Wang et al.ICML 2024 · 65 citations
- ProTDyn: A Foundation Protein Language Model for Thermodynamics and Dynamics GenerationYikai Liu, Haoyang Zheng, Lining Mao, Yanbin Wang et al.ICLR 2026 · 5 citations
- Inference-time optimization for experiment-grounded protein ensemble generationSai Advaith Maddipatla, Anar Rzayev, Marco Pegoraro, Martin Pacesa et al.ICML 2026 · 3 citations
- Reinforcing Diffusion Models by Direct Group Preference OptimizationYihong Luo, Tianyang Hu, Jing TangICLR 2026 · 13 citations
