Searching for High-Value Molecules Using Reinforcement Learning and Transformers
Raj Ghugare, Santiago Miret, Adriana Hugessen, Mariano Phielipp, Glen Berseth
摘要
Reinforcement learning (RL) over text representations can be effective for finding high-value policies that can search over graphs. However, RL requires careful structuring of the search space and algorithm design to be effective in this challenge. Through extensive experiments, we explore how different design choices for text grammar and algorithmic choices for training can affect an RL policy's ability to generate molecules with desired properties. We arrive at a new RL-based molecular design algorithm (ChemRLformer) and perform a thorough analysis using 25 molecule design tasks, including computationally complex protein docking simulations. From this analysis, we discover unique insights in this problem space and show that ChemRLformer achieves state-of-the-art performance while being more straightforward than prior work by demystifying which design choices are actually helpful for text-based molecule design.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- Overcoming the Sim-to-Real Gap: Leveraging Simulation to Learn to Explore for Real-World RLAndrew Wagenmaker, Kevin Huang, Liyiming Ke, Kevin Jamieson 等NeurIPS 2024 · 被引用 45 次
- Knowledge-aware Reinforced Language Models for Protein Directed EvolutionYuhao Wang, Qiang Zhang, Ming Qin, Xiang Zhuang 等ICML 2024 · 被引用 4 次
- Regulatory DNA Sequence Design with Reinforcement LearningZhao Yang, Bing Su, Chuan Cao, Ji-Rong WenICLR 2025
它引用的顶会 Paper12
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah 等NeurIPS 2020 · 被引用 64,255 次
- Flow Network based Generative Models for Non-Iterative Diverse Candidate GenerationEmmanuel Bengio, Moksh Jain, Maksym Korablyov, Doina Precup 等NeurIPS 2021 · 被引用 565 次
- Stabilizing Transformers for Reinforcement LearningEmilio Parisotto, H. Francis Song, Jack W. Rae, Razvan Pascanu 等ICML 2020 · 被引用 464 次
- Mastering Visual Continuous Control: Improved Data-Augmented Reinforcement LearningDenis Yarats, Rob Fergus, Alessandro Lazaric, Lerrel PintoICLR 2022 · 被引用 457 次
- Multi-Objective Molecule Generation using Interpretable SubstructuresWengong Jin, Regina Barzilay, Tommi S. JaakkolaICML 2020 · 被引用 238 次
相关 Paper
- Hit and Lead Discovery with Explorative RL and Fragment-based Molecule GenerationSoojung Yang, Doyeong Hwang, Seul Lee, Seongok Ryu 等NeurIPS 2021 · 被引用 106 次
- Reinforced Molecular Optimization with Neighborhood-Controlled GrammarsChencheng Xu, Qiao Liu, Minlie Huang, Tao JiangNeurIPS 2020 · 被引用 23 次
- MolEditRL: Structure-Preserving Molecular Editing via Discrete Diffusion and Reinforcement LearningYuanxin Zhuang, Dazhong Shen, Ying SunICLR 2026 · 被引用 2 次
- Cliqueformer: Model-Based Optimization with Structured TransformersJakub Grudzien Kuba, Pieter Abbeel, Sergey LevineAAAI 2026 · 被引用 5 次
- Molformer: Motif-Based Transformer on 3D Heterogeneous Molecular GraphsFang Wu, Dragomir Radev, Stan Z. LiAAAI 2023 · 被引用 96 次
