Tree Search-Based Evolutionary Bandits for Protein Sequence Optimization
Jiahao Qiu, Hui Yuan, Jinghong Zhang, Wentao Chen, Huazheng Wang, Mengdi Wang
Abstract
While modern biotechnologies allow synthesizing new proteins and function measurements at scale, efficiently exploring a protein sequence space and engineering it remains a daunting task due to the vast sequence space of any given protein. Protein engineering is typically conducted through an iterative process of adding mutations to the wild-type or lead sequences, recombination of mutations, and running new rounds of screening. To enhance the efficiency of such a process, we propose a tree search-based bandit learning method, which expands a tree starting from the initial sequence with the guidance of a bandit machine learning model. Under simplified assumptions and a Gaussian Process prior, we provide theoretical analysis and a Bayesian regret bound, demonstrating that the combination of local search and bandit learning method can efficiently discover a near-optimal design. The full algorithm is compatible with a suite of randomized tree search heuristics, machine learning models, pre-trained embeddings, and bandit techniques. We test various instances of the algorithm across benchmark protein datasets using simulated screens. Experiment results demonstrate that the algorithm is both sample-efficient and able to find top designs using reasonably small mutation counts.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 8f3df4f2-13f7-4a28-a437-8ebdbcac62b4Builds on11
- Neural Contextual Bandits with UCB-based ExplorationDongruo Zhou, Lihong Li, Quanquan GuICML 2020 · 329 citations
- Reinforcement Learning in Feature Space: Matrix Bandit, Kernels, and Regret BoundLin Yang, Mengdi WangICML 2020 · 308 citations
- Model-based reinforcement learning for biological sequence designChristof Angermüller, David Dohan, David Belanger, Ramya Deshpande et al.ICLR 2020 · 159 citations
- Neural Thompson SamplingWeitong Zhang, Dongruo Zhou, Lihong Li, Quanquan GuICLR 2021 · 152 citations
- Accelerating Bayesian Optimization for Biological Sequence Design with Denoising AutoencodersSamuel Stanton, Wesley J. Maddox, Nate Gruver, Phillip M. Maffettone et al.ICML 2022 · 137 citations
Related papers
- Bandit Theory and Thompson Sampling-Guided Directed Evolution for Sequence OptimizationHui Yuan, Chengzhuo Ni, Huazheng Wang, Xuezhou Zhang et al.NeurIPS 2022 · 4 citations
- Designing Biological Sequences without Prior Knowledge Using Evolutionary Reinforcement LearningXi Zeng, Xiaotian Hao, Hongyao Tang, Zhentao Tang et al.AAAI 2024 · 2 citations
- Proximal Exploration for Model-guided Protein Sequence DesignZhizhou Ren, Jiahan Li, Fan Ding, Yuan Zhou et al.ICML 2022 · 52 citations
- ProtInvTree: Deliberate Protein Inverse Folding with Reward-guided Tree SearchMengdi Liu, Xiaoxue Cheng, Zhangyang Gao, Hong Chang et al.NeurIPS 2025 · 10 citations
- Tree ensemble kernels for Bayesian optimization with known constraints over mixed-feature spacesAlexander Thebelt, Calvin Tsay, Robert M. Lee, Nathan Sudermann-Merx et al.NeurIPS 2022 · 18 citations
