Search papers, labs, and topics across Lattice.
This paper addresses the limitations of existing regret analyses for parallel Gaussian process (GP) bandit optimization, specifically the multiplicative factor related to batch size in known upper bounds for GP batched upper confidence bound and GP batched Thompson sampling. By eliminating the need for an initial uncertainty sampling phase, the authors demonstrate that it is possible to achieve improved regret upper bounds without this ineffective preliminary step. Their findings reveal that in the noiseless setting, the regret bounds significantly outperform those in the noisy setting, highlighting a critical distinction in performance across different environments.
Achieving improved regret bounds in parallel GP bandit optimization without the need for an initial uncertainty sampling phase could redefine efficiency in practical applications.
This paper studies the regret analysis for parallel Gaussian process (GP) bandit optimization. The known regret upper bounds for the widely used GP batched upper confidence bound and GP batched Thompson sampling (GP-BTS) suffer from a multiplicative factor with respect to the batch size $Q$. To avoid this degradation, existing analyses require a polynomial number of uncertainty sampling (US) for $Q$ at the beginning of optimization. However, this initial US phase is often ineffective in practice. This paper shows that the regret upper bound without the multiplicative factor on $Q$ can be achieved without the initial US phase, using GP-BTS as an example. Furthermore, we show much better regret upper bounds in the noiseless setting than in the noisy setting, as in the sequential GP bandit setting.