Performative Prediction with Bandit Feedback: Learning through Reparameterization

Chen, Yatong; Tang, Wei; Ho, Chien-Ju; Liu, Yang

Computer Science > Machine Learning

arXiv:2305.01094 (cs)

[Submitted on 1 May 2023 (v1), last revised 12 Aug 2024 (this version, v4)]

Title:Performative Prediction with Bandit Feedback: Learning through Reparameterization

Authors:Yatong Chen, Wei Tang, Chien-Ju Ho, Yang Liu

View PDF

Abstract:Performative prediction, as introduced by Perdomo et al, is a framework for studying social prediction in which the data distribution itself changes in response to the deployment of a model. Existing work in this field usually hinges on three assumptions that are easily violated in practice: that the performative risk is convex over the deployed model, that the mapping from the model to the data distribution is known to the model designer in advance, and the first-order information of the performative risk is available. In this paper, we initiate the study of performative prediction problems that do not require these assumptions. Specifically, we develop a reparameterization framework that reparametrizes the performative prediction objective as a function of the induced data distribution. We then develop a two-level zeroth-order optimization procedure, where the first level performs iterative optimization on the distribution parameter space, and the second level learns the model that induces a particular target distribution at each iteration. Under mild conditions, this reparameterization allows us to transform the non-convex objective into a convex one and achieve provable regret guarantees. In particular, we provide a regret bound that is sublinear in the total number of performative samples taken and is only polynomial in the dimension of the model parameter.

Subjects:	Machine Learning (cs.LG); Machine Learning (stat.ML)
Cite as:	arXiv:2305.01094 [cs.LG]
	(or arXiv:2305.01094v4 [cs.LG] for this version)
	https://doi.org/10.48550/arXiv.2305.01094

Submission history

From: Yatong Chen [view email]
[v1] Mon, 1 May 2023 21:31:29 UTC (409 KB)
[v2] Mon, 8 May 2023 05:34:57 UTC (408 KB)
[v3] Tue, 24 Oct 2023 20:16:24 UTC (264 KB)
[v4] Mon, 12 Aug 2024 20:59:55 UTC (440 KB)

Computer Science > Machine Learning

Title:Performative Prediction with Bandit Feedback: Learning through Reparameterization

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Machine Learning

Title:Performative Prediction with Bandit Feedback: Learning through Reparameterization

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators