Relaxed Indexability and Index Policy for Partially Observable Restless   Bandits

Keqin Liu

arXiv:2107.11939·math.OC·April 18, 2025·Manag. Sci.

Relaxed Indexability and Index Policy for Partially Observable Restless Bandits

Keqin Liu

PDF

Open Access

TL;DR

This paper introduces a relaxed indexability concept for partially observable restless bandits, enabling near-optimal, low-complexity algorithms for resource-constrained decision-making in stochastic environments.

Contribution

It generalizes Whittle's indexability to partially observable settings, facilitating efficient online computation of approximate indices for complex RMAB problems.

Findings

01

Achieves near-optimal performance with low complexity

02

Extends indexability to partially observable models

03

Provides efficient online index computation

Abstract

This paper addresses an important class of restless multi-armed bandit (RMAB) problems that finds broad application in operations research, stochastic optimization, and reinforcement learning. There are $N$ independent Markov processes that may be operated, observed and offer rewards. Due to the resource constraint, we can only choose a subset of $M (M < N)$ processes to operate and accrue reward determined by the states of selected processes. We formulate the problem as a partially observable RMAB with an infinite state space and design an algorithm that achieves a near-optimal performance with low complexity. Our algorithm is based on a generalization of Whittle's original idea of indexability. Referred to as the relaxed indexability, the extended definition leads to the efficient online verifications and computations of the approximate Whittle index under the proposed algorithmic…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

TopicsAdvanced Bandit Algorithms Research · Smart Grid Energy Management · Optimization and Search Problems