TerraBind: Fast and Accurate Binding Affinity Prediction through Coarse Structural Representations
Matteo Rossi, Ryan Pederson, Miles Wang-Henderson, Benjamin Kaufman, Edward Williams, Carl Underkoffler, Owen Howell, Adrian Layer, Stephan Thaler, Narbe Mardirossian, John Parkhill
Abstract
We present TerraBind, a foundation model for protein-ligand structure and binding affinity prediction that achieves 26-fold faster inference than state-of-the-art methods while improving affinity prediction accuracy by ∼20%. Current deep learning approaches to structure-based drug design rely on expensive allatom diffusion to generate 3D coordinates, creating inference bottlenecks that render large-scale compound screening computationally intractable. We challenge this paradigm with a critical hypothesis: full all-atom resolution is unnecessary for accurate small molecule pose and binding affinity prediction. TerraBind tests this hypothesis through a coarse pocket-level representation (protein C β atoms and ligand heavy atoms only) within a multimodal architecture combining COATI-3 molecular encodings and ESM-2 protein embeddings that learns rich structural representations, which are used in a diffusion-free optimization module for pose generation and a binding affinity likelihood prediction module. On structure prediction benchmarks (FoldBench, PoseBusters, Runs N' Poses), TerraBind matches diffusion-based baselines in ligand pose accuracy. Crucially, TerraBind outperforms Boltz-2 by ∼20% in Pearson correlation for binding affinity prediction on both a public benchmark (CASP16) and a diverse proprietary dataset (18 biochemical/cell assays). We show that the affinity prediction module also provides well-calibrated affinity uncertainty estimates, addressing a critical gap in reliable compound prioritization for drug discovery. Furthermore, this module enables a continual learning framework and a hedged batch selection strategy that, in simulated drug discovery cycles, achieves 6× greater affinity improvement of selected molecules over greedy-based approaches.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext d90d7325-41aa-47d5-a3b5-e6178a9bf893Builds on3
- Efficiently sampling functions from Gaussian process posteriorsJames T. Wilson, Viacheslav Borovitskiy, Alexander Terenin, Peter Mostowsky et al.ICML 2020 · 186 citations
- Epistemic Neural NetworksIan Osband, Zheng Wen, Seyed Mohammad Asghari, Vikranth Dwaracherla et al.NeurIPS 2023 · 142 citations
- NeuralPLexer3: Accurate Biomolecular Complex Structure Prediction with Flow ModelsJarren Zhuoran Qiao, Feizhi Ding, Thomas Dresselhaus, Mia A. Rosenfeld et al.NeurIPS 2025 · 23 citations
Related papers
- E3Bind: An End-to-End Equivariant Network for Protein-Ligand DockingYangtian Zhang, Huiyu Cai, Chence Shi, Jian TangICLR 2023 · 13 citations
- 3D Equivariant Diffusion for Target-Aware Molecule Generation and Affinity PredictionJiaqi Guan, Wesley Wei Qian, Xingang Peng, Yufeng Su et al.ICLR 2023 · 79 citations
- DiffDock: Diffusion Steps, Twists, and Turns for Molecular DockingGabriele Corso, Hannes Stärk, Bowen Jing, Regina Barzilay et al.ICLR 2023 · 331 citations
- EquiBind: Geometric Deep Learning for Drug Binding Structure PredictionHannes Stärk, Octavian Ganea, Lagnajit Pattanaik, Regina Barzilay et al.ICML 2022 · 360 citations
- Towards All-Atom Foundation Models for Biomolecular Binding Affinity PredictionLiang Shi, Zuobai Zhang, Huiyu Cai, Santiago Miret et al.ICLR 2026
