Reinforced Active Learning for Large-Scale Virtual Screening with Learnable Policy Model
Yicong Chen, Jiahua Rao, Jiancong Xie, Dahao Xu, Zhen Wang, Yuedong Yang
摘要
Virtual Screening (VS) is vital for drug discovery but struggles with low hit rates and high computational costs. While Active Learning (AL) has shown promise in improving the efficiency of VS, traditional methods rely on inflexible and handcrafted heuristics, limiting adaptability in complex chemical spaces, particularly in balancing molecular diversity and selection accuracy. To overcome these challenges, we propose GLARE 1 , a reinforced active learning framework that reformulates VS as a Markov Decision Process (MDP). Using Group Relative Policy Optimization (GRPO), GLARE dynamically balances chemical diversity, biological relevance, and computational constraints, eliminating the need for inflexible heuristics. Experiments show GLARE outperforms state-of-the-art AL methods, with a 64.8% average improvement in Enrichment Factors (EF). Additionally, GLARE enhances the performance of VS foundation models like DrugCLIP, achieving up to an 8-fold improvement in EF 0.5% with as few as 15 active molecules. These results highlight the transformative potential of GLARE for adaptive and efficient drug discovery.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper10
- Deep Batch Active Learning by Diverse, Uncertain Gradient Lower BoundsJordan T. Ash, Chicheng Zhang, Akshay Krishnamurthy, John Langford 等ICLR 2020 · 被引用 974 次
- Pre-training Molecular Graph Representation with 3D GeometryShengchao Liu, Hanchen Wang, Weiyang Liu, Joan Lasenby 等ICLR 2022 · 被引用 440 次
- 3D Infomax improves GNNs for Molecular Property PredictionHannes Stärk, Dominique Beaini, Gabriele Corso, Prudencio Tossou 等ICML 2022 · 被引用 269 次
- TANKBind: Trigonometry-Aware Neural NetworKs for Drug-Protein Binding Structure PredictionWei Lu, Qifeng Wu, Jixian Zhang, Jiahua Rao 等NeurIPS 2022 · 被引用 254 次
- Batch Active Learning at ScaleGui Citovsky, Giulia DeSalvo, Claudio Gentile, Lazaros Karydas 等NeurIPS 2021 · 被引用 220 次
相关 Paper
- DrugCLIP: Contrasive Protein-Molecule Representation Learning for Virtual ScreeningBowen Gao, Bo Qiang, Haichuan Tan, Yinjun Jia 等NeurIPS 2023 · 被引用 45 次
- GraphAF: a Flow-based Autoregressive Model for Molecular Graph GenerationChence Shi, Minkai Xu, Zhaocheng Zhu, Weinan Zhang 等ICLR 2020 · 被引用 532 次
- Learning to Navigate The Synthetically Accessible Chemical Space Using Reinforcement LearningSai Krishna Gottipati, Boris Sattarov, Sufeng Niu, Yashaswi Pathak 等ICML 2020 · 被引用 127 次
- AANet: Virtual Screening under Structural Uncertainty via Alignment and AggregationWenyu Zhu, Jianhui Wang, Bowen Gao, Yinjun Jia 等NeurIPS 2025 · 被引用 2 次
- MARS: Markov Molecular Sampling for Multi-objective Drug DiscoveryYutong Xie, Chence Shi, Hao Zhou, Yuwei Yang 等ICLR 2021 · 被引用 186 次
