Tanimoto Random Features for Scalable Molecular Machine Learning
Austin Tripp, Sergio Bacallado, Sukriti Singh, José Miguel Hernández-Lobato
摘要
The Tanimoto coefficient is commonly used to measure the similarity between molecules represented as discrete fingerprints, either as a distance metric or a positive definite kernel. While many kernel methods can be accelerated using random feature approximations, at present there is a lack of such approximations for the Tanimoto kernel. In this paper we propose two kinds of novel random features to allow this kernel to scale to large datasets, and in the process discover a novel extension of the kernel to real-valued vectors. We theoretically characterize these random features, and provide error bounds on the spectral norm of the Gram matrix. Experimentally, we show that these random features are effective at approximating the Tanimoto coefficient of real-world datasets and are useful for molecular property prediction and optimization tasks.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- Stochastic Gradient Descent for Gaussian Processes Done RightJihao Andreas Lin, Shreyas Padhy, Javier Antorán, Austin Tripp 等ICLR 2024 · 被引用 17 次
- Retro-fallback: retrosynthetic planning in an uncertain worldAustin Tripp, Krzysztof Maziarz, Sarah Lewis, Marwin H. S. Segler 等ICLR 2024 · 被引用 14 次
- Flexible Kernels for Protein Property PredictionMartin Jankowiak, Yerdos Ordabayev, Rudraksh Tuwani, Henry Ward 等ICML 2026
- Variance-Reducing Couplings for Random FeaturesIsaac Reid, Stratis Markou, Krzysztof Marcin Choromanski, Richard E. Turner 等ICLR 2025
- Return of the Latent Space COWBOYS: Re-thinking the use of VAEs for Bayesian Optimisation of Structured SpacesHenry B. Moss, Sebastian W. Ober, Tom DietheICML 2025
它引用的顶会 Paper4
- Efficiently sampling functions from Gaussian process posteriorsJames T. Wilson, Viacheslav Borovitskiy, Alexander Terenin, Peter Mostowsky 等ICML 2020 · 被引用 186 次
- Oblivious Sketching of High-Degree Polynomial KernelsThomas D. Ahle, Michael Kapralov, Jakob Bæk Tejs Knudsen, Rasmus Pagh 等SODA 2020 · 被引用 42 次
- Scaling Neural Tangent Kernels via Sketching and Random FeaturesAmir Zandieh, Insu Han, Haim Avron, Neta Shoham 等NeurIPS 2021 · 被引用 42 次
- Bayesian Algorithm Execution: Estimating Computable Properties of Black-box Functions Using Mutual InformationWillie Neiswanger, Ke Alexander Wang, Stefano ErmonICML 2021 · 被引用 40 次
相关 Paper
- TRF: Learning Kernels with Tuned Random FeaturesAlistair Shilton, Sunil Gupta, Santu Rana, Arun Kumar Anjanapura Venkatesh 等AAAI 2022
- Taming graph kernels with random featuresKrzysztof Marcin ChoromanskiICML 2023 · 被引用 21 次
- Weisfeiler and Leman Go Walking: Random Walk Kernels RevisitedNils M. KriegeNeurIPS 2022 · 被引用 22 次
- Learning with Optimized Random Features: Exponential Speedup by Quantum Machine Learning without Sparsity and Low-Rank AssumptionsHayata Yamasaki, Sathyawageeswar Subramanian, Sho Sonoda, Masato KoashiNeurIPS 2020 · 被引用 23 次
- Measuring dissimilarity with diffeomorphism invarianceThéophile Cantelobre, Carlo Ciliberto, Benjamin Guedj, Alessandro RudiICML 2022 · 被引用 1 次
