arXiv · 2605.10299
Nearly-Optimal Algorithm for Adversarial Kernelized Bandits
Abstract
This paper studies kernelized bandits (also known as Gaussian process bandits) in an adversarial environment, where the reward functions in a known reproducing kernel Hilbert space (RKHS) may be adversarially chosen at each round. We show that the exponential-weight algorithm achieves $\tilde{O}(\sqrt{T \gamma_T})$ adversarial regret, where $T$ and $\gamma_T$ denote the number of total rounds and the maximum information gain, respectively. For squared exponential (SE) and $\nu$-Mat\'ern kernels, we also show algorithm-independent lower bounds that guarantee the optimality of our algorithm up to polylogarithmic factors. Furthermore, we present a computationally efficient variant of our algorithm using Nystr\"om approximation while maintaining nearly optimal regret guarantees.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Shogo Iwazaki. 2026-05-11. Nearly-Optimal Algorithm for Adversarial Kernelized Bandits. https://arxiv.org/abs/2605.10299
Cite the original work for its findings. Save a collection to share your selection of sources.