A Complete Characterization of Learnability for Stochastic Noisy Bandits

Steve Hanneke; Kun Wang

arXiv:2410.09597·cs.LG·January 20, 2025

A Complete Characterization of Learnability for Stochastic Noisy Bandits

Steve Hanneke, Kun Wang

PDF

Open Access 4 Reviews

TL;DR

This paper provides a comprehensive characterization of when and how efficiently one can learn optimal actions in stochastic noisy bandit problems with unknown reward functions, addressing a key open question in the field.

Contribution

It offers the first complete characterization of learnability for noisy bandit classes, describes the full range of optimal query complexities, and introduces a new DEC variant for this setting.

Findings

01

Learnability is decidable for classes with arbitrary noise.

02

Full spectrum of optimal query complexities characterized.

03

Adaptivity may be necessary for optimal performance.

Abstract

We study the stochastic noisy bandit problem with an unknown reward function $f^{*}$ in a known function class $F$ . Formally, a model $M$ maps arms $π$ to a probability distribution $M (π)$ of reward. A model class $M$ is a collection of models. For each model $M$ , define its mean reward function $f^{M} (π) = E_{r \sim M (π)} [r]$ . In the bandit learning problem, we proceed in rounds, pulling one arm $π$ each round and observing a reward sampled from $M (π)$ . With knowledge of $M$ , supposing that the true model $M \in M$ , the objective is to identify an arm $\overset{π}{^}$ of near-maximal mean reward $f^{M} (\overset{π}{^})$ with high probability in a bounded number of rounds. If this is possible, then the model class is said to be learnable. Importantly, a result of \cite{hanneke2023bandit} shows there exist model classes for which learnability is…

Peer Reviews

Decision·ALT 2025

Reviewer 01Rating · AcceptConfidence 3

Reviewer 02Rating 6Confidence 3

Reviewer 03Rating 6Confidence 5

Reviewer 04Rating 7Confidence 3

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

TopicsAdvanced Bandit Algorithms Research · Data Stream Mining Techniques · Distributed Sensor Networks and Detection Algorithms