Lune

ICML2026Top-tier venue

Tighter Regret Lower Bound for Gaussian Process Bandits with Squared Exponential Kernel in Hypersphere

Shogo Iwazaki

2026Year
5Citations

Abstract

We study an algorithm-independent, worst-case lower bound for the Gaussian process (GP) bandit problem in the frequentist setting, where the reward function is fixed and has a bounded norm in the known reproducing kernel Hilbert space (RKHS). Specifically, we focus on the squared exponential (SE) kernel, one of the most widely used kernel functions in GP bandits. One of the remaining open questions for this problem is the gap in the dimension-dependent logarithmic factors between upper and lower bounds. This paper partially resolves this open question under a hyperspherical input domain. We show that any algorithm suffers Ω(T(ln⁡T)d(ln⁡ln⁡T)−d)\Omega(\sqrt{T (\ln T)^{d} (\ln \ln T)^{-d}}) cumulative regret, where TT and dd represent the total number of steps and the dimension of the hyperspherical domain, respectively. Regarding the simple regret, we show that any algorithm requires Ω(ϵ−2(ln⁡1ϵ)d(ln⁡ln⁡1ϵ)−d)\Omega(\epsilon^{-2}(\ln \frac{1}{\epsilon})^d (\ln \ln \frac{1}{\epsilon})^{-d}) time steps to find an ϵ\epsilon-optimal point. We also provide the improved O((ln⁡T)d+1(ln⁡ln⁡T)−d)O((\ln T)^{d+1}(\ln \ln T)^{-d}) upper bound on the maximum information gain for the SE kernel. Our results guarantee the optimality of the existing best algorithm up to dimension-independent logarithmic factors under a hyperspherical input domain.

Ask about this paper

Your agent reads all of it.

Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.

Questions to start from

Your agent calls

Luneget_paper_fulltext

Ask in Lune

Free to start. No credit card required.

Builds on11

Related papers

Dusk over the sea between two cliffs drawn in fine vertical lines