Gaussian process bandits with adaptive discretization

DOI10.1214/18-EJS1497MaRDI QIDQ1711556zbMATH OpenOpenAlexWikidataFDO

Publication date 18 January 2019

Published in Electronic Journal of Statistics (Search for Journal in Brave)

Full work available at URL https://arxiv.org/abs/1712.01447, https://projecteuclid.org/euclid.ejs/1543892564

Gaussian processes Bayesian optimization bandits

Bayesian inference (62F15) Learning and adaptive systems in artificial intelligence (68T05) Gaussian processes (60G15) Bayesian problems; characterization of Bayes procedures (62C10)

Abstract: In this paper, the problem of maximizing a black-box function

f : m a t h c a l X o m a t h b b R

is studied in the Bayesian framework with a Gaussian Process (GP) prior. In particular, a new algorithm for this problem is proposed, and high probability bounds on its simple and cumulative regret are established. The query point selection rule in most existing methods involves an exhaustive search over an increasingly fine sequence of uniform discretizations of

m a t h c a l X

. The proposed algorithm, in contrast, adaptively refines

m a t h c a l X

which leads to a lower computational complexity, particularly when

m a t h c a l X

is a subset of a high dimensional Euclidean space. In addition to the computational gains, sufficient conditions are identified under which the regret bounds of the new algorithm improve upon the known results. Finally an extension of the algorithm to the case of contextual bandits is proposed, and high probability bounds on the contextual regret are presented.

Recommendations

Cites work

Cited in

(11)

This page was built for publication: Gaussian process bandits with adaptive discretization

Report a bug (only for logged in users!)Click here to report a bug for this page (MaRDI item Q1711556)