Knowledge Gradient for Selection with Covariates: Consistency and   Computation

Liang Ding; L. Jeff Hong; Haihui Shen; Xiaowei Zhang

arXiv:1906.05098·math.ST·January 17, 2022·5 cites

Knowledge Gradient for Selection with Covariates: Consistency and Computation

Liang Ding, L. Jeff Hong, Haihui Shen, Xiaowei Zhang

PDF

Open Access 1 Repo

TL;DR

This paper extends the knowledge gradient method to ranking and selection problems with covariates, proving its consistency and proposing a stochastic gradient algorithm for efficient computation.

Contribution

It introduces a knowledge gradient-based sampling policy for covariate-dependent selection, proving its almost sure consistency and providing a practical computation method.

Findings

01

The policy is consistent under minimal assumptions.

02

The stochastic gradient algorithm effectively computes the policy.

03

Numerical experiments demonstrate the method's performance.

Abstract

Knowledge gradient is a design principle for developing Bayesian sequential sampling policies to solve optimization problems. In this paper we consider the ranking and selection problem in the presence of covariates, where the best alternative is not universal but depends on the covariates. In this context, we prove that under minimal assumptions, the sampling policy based on knowledge gradient is consistent, in the sense that following the policy the best alternative as a function of the covariates will be identified almost surely as the number of samples grows. We also propose a stochastic gradient ascent algorithm for computing the sampling policy and demonstrate its performance via numerical experiments.

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Code & Models

Repositories

shenhaihui/ikg
noneOfficial

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

TopicsAdvanced Bandit Algorithms Research · Machine Learning and Algorithms · Gaussian Processes and Bayesian Inference