Problem-Complexity Adaptive Model Selection for Stochastic Linear   Bandits

Avishek Ghosh; Abishek Sankararaman; Kannan Ramchandran

arXiv:2006.02612·stat.ML·June 17, 2020·1 cites

Problem-Complexity Adaptive Model Selection for Stochastic Linear Bandits

Avishek Ghosh, Abishek Sankararaman, Kannan Ramchandran

PDF

Open Access

TL;DR

This paper introduces adaptive algorithms for stochastic linear bandits that automatically adjust to unknown problem complexities like parameter norm and sparsity, achieving near-optimal regret bounds.

Contribution

The paper presents the first algorithms that adaptively select models based on unknown complexity measures such as parameter norm and sparsity in linear bandits.

Findings

01

ALB achieves regret of O(∥θ*∥√T) without prior knowledge of ∥θ*∥.

02

ALB attains regret of O(d*√T) when sparsity d* is unknown.

03

Experimental results confirm theoretical guarantees on synthetic and real data.

Abstract

We consider the problem of model selection for two popular stochastic linear bandit settings, and propose algorithms that adapts to the unknown problem complexity. In the first setting, we consider the $K$ armed mixture bandits, where the mean reward of arm $i \in [K]$ , is $μ_{i} + ⟨ α_{i, t}, θ^{*} ⟩$ , with $α_{i, t} \in R^{d}$ being the known context vector and $μ_{i} \in [- 1, 1]$ and $θ^{*}$ are unknown parameters. We define $∥ θ^{*} ∥$ as the problem complexity and consider a sequence of nested hypothesis classes, each positing a different upper bound on $∥ θ^{*} ∥$ . Exploiting this, we propose Adaptive Linear Bandit (ALB), a novel phase based algorithm that adapts to the true problem complexity, $∥ θ^{*} ∥$ . We show that ALB achieves regret scaling of $O (∥ θ^{*} ∥ T)$ , where $∥ θ^{*} ∥$ is apriori unknown. As a corollary, when…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

TopicsAdvanced Bandit Algorithms Research · Machine Learning and Algorithms · Optimization and Search Problems