Pith. sign in

REVIEW 1 cited by

Correcting for Interference in Experiments: A Case Study at Douyin

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2305.02542 v1 pith:ZFD6WW4O submitted 2023-05-04 stat.ME cs.LGstat.APstat.ML

classification stat.MEcs.LGstat.APstat.ML
keywords interferencedouyineffectestimatorexperimentstreatmentbiascreators
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Interference is a ubiquitous problem in experiments conducted on two-sided content marketplaces, such as Douyin (China's analog of TikTok). In many cases, creators are the natural unit of experimentation, but creators interfere with each other through competition for viewers' limited time and attention. "Naive" estimators currently used in practice simply ignore the interference, but in doing so incur bias on the order of the treatment effect. We formalize the problem of inference in such experiments as one of policy evaluation. Off-policy estimators, while unbiased, are impractically high variance. We introduce a novel Monte-Carlo estimator, based on "Differences-in-Qs" (DQ) techniques, which achieves bias that is second-order in the treatment effect, while remaining sample-efficient to estimate. On the theoretical side, our contribution is to develop a generalized theory of Taylor expansions for policy evaluation, which extends DQ theory to all major MDP formulations. On the practical side, we implement our estimator on Douyin's experimentation platform, and in the process develop DQ into a truly "plug-and-play" estimator for interference in real-world settings: one which provides robust, low-bias, low-variance treatment effect estimates; admits computationally cheap, asymptotically exact uncertainty quantification; and reduces MSE by 99\% compared to the best existing alternatives in our applications.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Deep-learning Causal Retrieval Optimization for Efficient e-commerce Distribution in Pinterest

    cs.IR 2026-07 conditional novelty 5.0 of 10

    A multi-task causal uplift model decides when to trigger shopping candidate generators in early retrieval, cutting triggers by up to 85% with neutral shopping sessions and positive engagement.

Pith tools