Bandit Theory and Thompson Sampling-Guided Directed Evolution for Sequence Optimization
Hui Yuan, Chengzhuo Ni, Huazheng Wang, Xuezhou Zhang, Le Cong, Csaba Szepesvári, Mengdi Wang
摘要
Directed Evolution (DE), a landmark wet-lab method originated in 1960s, enables discovery of novel protein designs via evolving a population of candidate sequences. Recent advances in biotechnology has made it possible to collect high-throughput data, allowing the use of machine learning to map out a protein's sequence-to-function relation. There is a growing interest in machine learning-assisted DE for accelerating protein optimization. Yet the theoretical understanding of DE, as well as the use of machine learning in DE, remains limited. In this paper, we connect DE with the bandit learning theory and make a first attempt to study regret minimization in DE. We propose a Thompson Sampling-guided Directed Evolution (TS-DE) framework for sequence optimization, where the sequence-to-function mapping is unknown and querying a single value is subject to costly and noisy measurements. TS-DE updates a posterior of the function based on collected measurements. It uses a posterior-sampled function estimate to guide the crossover recombination and mutation steps in DE. In the case of a linear model, we show that TS-DE enjoys a Bayesian regret of order , where is feature dimension, is population size and is number of rounds. This regret bound is nearly optimal, confirming that bandit learning can provably accelerate DE. It may have implications for more general sequence optimization and evolutionary algorithms.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper1
相关 Paper
- Steering Generative Models with Experimental Data for Protein Fitness OptimizationJason Yang, Wenda Chu, Daniel Khalil, Raul Astudillo 等NeurIPS 2025 · 被引用 13 次
- Knowledge-aware Reinforced Language Models for Protein Directed EvolutionYuhao Wang, Qiang Zhang, Ming Qin, Xiang Zhuang 等ICML 2024 · 被引用 4 次
- Proximal Exploration for Model-guided Protein Sequence DesignZhizhou Ren, Jiahan Li, Fan Ding, Yuan Zhou 等ICML 2022 · 被引用 52 次
- Thompson Sampling via Fine-Tuning of LLMsNicolas Menet, Aleksandar Terzic, Michael Hersche, Andreas Krause 等ICLR 2026 · 被引用 6 次
- Adaptive Sampling for DiscoveryZiping Xu, Eunjae Shim, Ambuj Tewari, Paul M. ZimmermanNeurIPS 2022 · 被引用 5 次
