Pith. sign in

REVIEW 1 cited by

Accurate Inference for Adaptive Linear Models

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 1712.06695 v5 pith:CJFHJMD3 submitted 2017-12-18 stat.ML cs.LG

classification stat.MLcs.LG
keywords dataadaptivemathbfpolicybiascollectiondecorrelationdevelop
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
abstract

Estimators computed from adaptively collected data do not behave like their non-adaptive brethren. Rather, the sequential dependence of the collection policy can lead to severe distributional biases that persist even in the infinite data limit. We develop a general method -- $\mathbf{W}$-decorrelation -- for transforming the bias of adaptive linear regression estimators into variance. The method uses only coarse-grained information about the data collection policy and does not need access to propensity scores or exact knowledge of the policy. We bound the finite-sample bias and variance of the $\mathbf{W}$-estimator and develop asymptotically correct confidence intervals based on a novel martingale central limit theorem. We then demonstrate the empirical benefits of the generic $\mathbf{W}$-decorrelation procedure in two different adaptive data settings: the multi-armed bandit and the autoregressive time series.

Discussion (0). Sign in to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Stabilizing Bandits using Regularization: Precise Regret and A Quantitative Central Limit Theorem

    stat.ML 2026-03 conditional novelty 6.0 of 10

    Log-barrier regularized stochastic mirror descent yields Lai–Wei stable bandit sampling, valid Wald intervals, near-optimal regret up to logs, and asymptotic normality under o(√T) corruption.

Pith tools