Performative Prediction with Bandit Feedback: Learning through   Reparameterization

Yatong Chen; Wei Tang; Chien-Ju Ho; Yang Liu

arXiv:2305.01094·cs.LG·August 14, 2024·1 cites

Performative Prediction with Bandit Feedback: Learning through Reparameterization

Yatong Chen, Wei Tang, Chien-Ju Ho, Yang Liu

PDF

Open Access 1 Repo

TL;DR

This paper introduces a reparameterization approach for performative prediction that relaxes common assumptions, enabling convex reformulation and provable regret guarantees using a two-level zeroth-order optimization method.

Contribution

It develops a novel reparameterization framework and a two-level optimization procedure for performative prediction without requiring convexity or known distribution mappings.

Findings

01

Transform non-convex performative risk into a convex form under mild conditions.

02

Achieve sublinear regret bounds in the number of performative samples.

03

Method applicable in high-dimensional settings with polynomial regret dependence.

Abstract

Performative prediction, as introduced by Perdomo et al, is a framework for studying social prediction in which the data distribution itself changes in response to the deployment of a model. Existing work in this field usually hinges on three assumptions that are easily violated in practice: that the performative risk is convex over the deployed model, that the mapping from the model to the data distribution is known to the model designer in advance, and the first-order information of the performative risk is available. In this paper, we initiate the study of performative prediction problems that do not require these assumptions. Specifically, we develop a reparameterization framework that reparametrizes the performative prediction objective as a function of the induced data distribution. We then develop a two-level zeroth-order optimization procedure, where the first level performs…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Code & Models

Repositories

ucsc-real/pp-bandit-feedback
noneOfficial

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

TopicsAdvanced Bandit Algorithms Research · Stochastic Gradient Optimization Techniques · Gaussian Processes and Bayesian Inference