Pith. sign in

REVIEW 1 cited by

Performative Prediction with Bandit Feedback: Learning through Reparameterization

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2305.01094 v4 pith:HRAPGU2J submitted 2023-05-01 cs.LG stat.ML

classification cs.LGstat.ML
keywords performativemodeldistributionpredictiondatareparameterizationassumptionsconvex
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Performative prediction, as introduced by Perdomo et al, is a framework for studying social prediction in which the data distribution itself changes in response to the deployment of a model. Existing work in this field usually hinges on three assumptions that are easily violated in practice: that the performative risk is convex over the deployed model, that the mapping from the model to the data distribution is known to the model designer in advance, and the first-order information of the performative risk is available. In this paper, we initiate the study of performative prediction problems that do not require these assumptions. Specifically, we develop a reparameterization framework that reparametrizes the performative prediction objective as a function of the induced data distribution. We then develop a two-level zeroth-order optimization procedure, where the first level performs iterative optimization on the distribution parameter space, and the second level learns the model that induces a particular target distribution at each iteration. Under mild conditions, this reparameterization allows us to transform the non-convex objective into a convex one and achieve provable regret guarantees. In particular, we provide a regret bound that is sublinear in the total number of performative samples taken and is only polynomial in the dimension of the model parameter.

Discussion (0). Sign in to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Performative Risk Control: Calibrating Models for Reliable Deployment under Performativity

    stat.ML 2025-05 conditional novelty 6.0 of 10

    An iterative threshold calibration method, Performative Risk Control, gives finite-sample guarantees that risk stays controlled under performative (self-influencing) distribution shifts.

Pith tools