REVIEW 3 cited by
Decision-Dependent Stochastic Optimization: The Role of Distribution Dynamics
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Distribution shifts have long been regarded as troublesome external forces that a decision-maker should either counteract or conform to. An intriguing feedback phenomenon termed decision dependence arises when the deployed decision affects the environment and alters the data-generating distribution. In the realm of performative prediction, this is encoded by distribution maps parameterized by decisions due to strategic behaviors. In contrast, we formalize an endogenous distribution shift as a feedback process featuring nonlinear dynamics that couple the evolving distribution with the decision. Stochastic optimization in this dynamic regime provides a fertile ground to examine the various roles played by dynamics in the composite problem structure. To this end, we develop an online algorithm that achieves optimal decision-making by both adapting to and shaping the dynamic distribution. Throughout the paper, we adopt a distributional perspective and demonstrate how this view facilitates characterizations of distribution dynamics and the optimality and generalization performance of the proposed algorithm. We showcase the theoretical results in an opinion dynamics context, where an opportunistic party maximizes the affinity of a dynamic polarized population, and in a recommender system scenario, featuring performance optimization with discrete distributions in the probability simplex.
Forward citations
Cited by 3 Pith papers
-
Lipschitz continuity of expected value under decision-dependent uncertainty with moving support
Uniform expectations over Lipschitz-moving convex supports are locally Lipschitz under full-dimensional or constant-dimension conditions, and a counterexample shows Lipschitz support alone is insufficient.
-
Online Feedback Optimization for Constrained Stochastic Problems with Decision-Dependent Distributions: Extended Version
Develops projected primal-dual OFO algorithm for decision-dependent stochastic optimization and bounds mean-square tracking error with four interpretable terms.
-
Foundations of Reinforcement Learning and Control:Connections and New Perspectives
A SAC-trained Half-Cheetah policy paired with a low-level model-reference adaptive controller recovers running performance after a change in joint damping, where the fixed learned policy alone fails.
Discussion (0). Sign in to comment.