Regret Analysis of Sleeping Competing Bandits

Shinnosuke Uba; Yutaro Yamaguchi

arXiv:2603.19700·cs.LG·March 23, 2026

Regret Analysis of Sleeping Competing Bandits

Shinnosuke Uba, Yutaro Yamaguchi

PDF

Open Access

TL;DR

This paper introduces Sleeping Competing Bandits, extending the competing bandits framework to scenarios with variable availability, and provides regret bounds and an optimal algorithm for this setting.

Contribution

It formulates Sleeping Competing Bandits, derives regret bounds, and proposes an asymptotically optimal algorithm for this new model.

Findings

01

Proposed an algorithm with regret bound of O(NK log T_i / Δ^2).

02

Established a regret lower bound of Ω(N(K-N+1) log T_i / Δ^2).

03

Algorithm is asymptotically optimal when K is larger than N.

Abstract

The Competing Bandits framework is a recently emerging area that integrates multi-armed bandits in online learning with stable matching in game theory. While conventional models assume that all players and arms are constantly available, in real-world problems, their availability can vary arbitrarily over time. In this paper, we formulate this setting as Sleeping Competing Bandits. To analyze this problem, we naturally extend the regret definition used in existing competing bandits and derive regret bounds for the proposed model. We propose an algorithm that simultaneously achieves an asymptotic regret bound of $O (N K lo g T_{i} / Δ^{2})$ under reasonable assumptions, where $N$ is the number of players, $K$ is the number of arms, $T_{i}$ is the number of rounds of each player $p_{i}$ , and $Δ$ is the minimum reward gap. We also provide a regret lower bound of…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

TopicsAdvanced Bandit Algorithms Research · Game Theory and Applications · Optimization and Search Problems