A Robust Phased Elimination Algorithm for Corruption-Tolerant Gaussian   Process Bandits

Ilija Bogunovic; Zihan Li; Andreas Krause; Jonathan Scarlett

arXiv:2202.01850·stat.ML·March 30, 2022

A Robust Phased Elimination Algorithm for Corruption-Tolerant Gaussian Process Bandits

Ilija Bogunovic, Zihan Li, Andreas Krause, Jonathan Scarlett

PDF

Open Access 1 Video

TL;DR

This paper introduces RGP-PE, a robust algorithm for Gaussian process bandit optimization that effectively handles adversarial corruptions, providing tighter regret bounds and demonstrating empirical robustness.

Contribution

The paper proposes a novel robust elimination algorithm for corrupted GP bandits with improved theoretical regret bounds and empirical validation of robustness.

Findings

01

Regret bound improves to O(C γ_T^{3/2}) from O(C √T γ_T)

02

Algorithm demonstrates robustness against various adversarial attacks

03

First empirical study of robustness in corrupted GP bandit setting

Abstract

We consider the sequential optimization of an unknown, continuous, and expensive to evaluate reward function, from noisy and adversarially corrupted observed rewards. When the corruption attacks are subject to a suitable budget $C$ and the function lives in a Reproducing Kernel Hilbert Space (RKHS), the problem can be posed as corrupted Gaussian process (GP) bandit optimization. We propose a novel robust elimination-type algorithm that runs in epochs, combines exploration with infrequent switching to select a small subset of actions, and plays each action for multiple time instants. Our algorithm, Robust GP Phased Elimination (RGP-PE), successfully balances robustness to corruptions with exploration and exploitation such that its performance degrades minimally in the presence (or absence) of adversarial corruptions. When $T$ is the number of samples and $γ_{T}$ is the maximal…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Videos

A Robust Phased Elimination Algorithm for Corruption-Tolerant Gaussian Process Bandits· slideslive

Taxonomy

TopicsAdvanced Bandit Algorithms Research · Adversarial Robustness in Machine Learning · Gaussian Processes and Bayesian Inference

MethodsGaussian Process