Pith. sign in

REVIEW 3 cited by

Hidden Incentives for Auto-Induced Distributional Shift

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2009.09153 v1 pith:O5D5ZQZD submitted 2020-09-19 cs.LG cs.AIstat.ML

classification cs.LGcs.AIstat.ML
keywords learningshiftauto-inducedcontentdistributionalhi-adsmachinealgorithm
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Decisions made by machine learning systems have increasing influence on the world, yet it is common for machine learning algorithms to assume that no such influence exists. An example is the use of the i.i.d. assumption in content recommendation. In fact, the (choice of) content displayed can change users' perceptions and preferences, or even drive them away, causing a shift in the distribution of users. We introduce the term auto-induced distributional shift (ADS) to describe the phenomenon of an algorithm causing a change in the distribution of its own inputs. Our goal is to ensure that machine learning systems do not leverage ADS to increase performance when doing so could be undesirable. We demonstrate that changes to the learning algorithm, such as the introduction of meta-learning, can cause hidden incentives for auto-induced distributional shift (HI-ADS) to be revealed. To address this issue, we introduce `unit tests' and a mitigation strategy for HI-ADS, as well as a toy environment for modelling real-world issues with HI-ADS in content recommendation, where we demonstrate that strong meta-learners achieve gains in performance via ADS. We show meta-learning and Q-learning both sometimes fail unit tests, but pass when using our mitigation strategy.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 3 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Joint Scoring Rules: Zero-Sum Competition Avoids Performative Prediction

    cs.LG 2024-12 conditional novelty 7.0 of 10

    Joint zero-sum scoring of multiple conditional predictors lets a principal deterministically take their most preferred action without performative manipulation.

  2. FairSense: Long-Term Fairness Analysis of ML-Enabled Systems

    cs.LG 2025-01 conditional novelty 6.0 of 10

    FairSense uses Monte-Carlo simulation and sensitivity analysis to identify which design and environmental parameters drive long-term unfairness in ML-enabled systems, demonstrated on three case studies.

  3. AI Value Alignment for Evolving Social Norms

    cs.CY 2026-07 conditional novelty 5.5 of 10

    Static AI alignment to historical user values produces value lock-in and can collapse distinct social norms into a maladaptive consensus, so alignment should be dynamic and adaptive.

Pith tools