Enhancing Ligand Validity and Affinity in Structure-Based Drug Design with Multi-Reward Optimization
Seungbeom Lee, Munsun Jo, Jungseul Ok, Dongwoo Kim
Abstract
Deep learning-based structure-based drug design aims to generate ligand molecules with desirable properties for protein targets. While existing models have demonstrated competitive performance in generating ligand molecules, they primarily focus on learning the chemical distribution of training datasets, often lacking effective steerability to ensure the desired chemical quality of generated molecules. To address this issue, we propose a multi-reward optimization framework that finetunes generative models for attributes, such as binding affinity, validity, and drug-likeness, together. Specifically, we derive direct preference optimization for a Bayesian flow network, used as a backbone for molecule generation, and integrate a reward normalization scheme to adopt multiple objectives. Experimental results show that our method generates more realistic ligands than baseline models while achieving higher binding affinity, expanding the Pareto front empirically observed in previous studies.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext aeab0f7a-7570-4494-b28c-0759db11a7bdCited by top-tier papers1
Ask how each one uses itBuilds on13
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- Training language models to follow instructions with human feedbackLong Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida et al.NeurIPS 2022 · 24,707 citations
- Direct Preference Optimization: Your Language Model is Secretly a Reward ModelRafael Rafailov, Archit Sharma, Eric Mitchell, Christopher D. Manning et al.NeurIPS 2023 · 10,924 citations
- Gradient Surgery for Multi-Task LearningTianhe Yu, Saurabh Kumar, Abhishek Gupta, Sergey Levine et al.NeurIPS 2020 · 2,261 citations
- A 3D Generative Model for Structure-Based Drug DesignShitong Luo, Jiaqi Guan, Jianzhu Ma, Jian PengNeurIPS 2021 · 302 citations
Related papers
- Reinforced Genetic Algorithm for Structure-based Drug DesignTianfan Fu, Wenhao Gao, Connor W. Coley, Jimeng SunNeurIPS 2022 · 79 citations
- Molecule Generation For Target Protein Binding with Structural MotifsZaixi Zhang, Yaosen Min, Shuxin Zheng, Qi LiuICLR 2023
- Multi-domain Distribution Learning for De Novo Drug DesignArne Schneuing, Ilia Igashov, Adrian W. Dobbelstein, Thomas Castiglione et al.ICLR 2025
- GraphAF: a Flow-based Autoregressive Model for Molecular Graph GenerationChence Shi, Minkai Xu, Zhaocheng Zhu, Weinan Zhang et al.ICLR 2020 · 532 citations
- FlexSBDD: Structure-Based Drug Design with Flexible Protein ModelingZaixi Zhang, Mengdi Wang, Qi LiuNeurIPS 2024 · 19 citations
