Pith. sign in

REVIEW 1 cited by

Optimization-Driven Adaptive Experimentation

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2408.04570 v2 pith:FY4FRGKG submitted 2024-08-08 cs.LG

classification cs.LG
keywords adaptiveconstraintsmethodsobjectivesacrossbespokeexperimentsframework
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
abstract

Real-world experiments involve batched & delayed feedback, non-stationarity, multiple objectives & constraints, and (often some) personalization. Tailoring adaptive methods to address these challenges on a per-problem basis is infeasible, and static designs remain the de facto standard. Focusing on short-horizon ($\le 10$) adaptive experiments, we move away from bespoke algorithms and present a mathematical programming formulation that can flexibly incorporate a wide range of objectives, constraints, and statistical procedures. We formulating a dynamic program based on central limit approximations, which enables the use of scalable optimization methods based on auto-differentiation and GPU parallelization. To evaluate our framework, we implement a simple heuristic planning method ("solver") and benchmark it across hundreds of problem instances involving non-stationarity, personalization, and multiple objectives & constraints. Unlike bespoke methods (e.g., Thompson sampling variants), our mathematical programming framework provides consistent gains over static randomized control trials and exhibits robust performance across problem instances.

Discussion (0). Sign in to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Optimization of Epsilon-Greedy Exploration

    cs.LG 2025-06 conditional novelty 6.0 of 10

    A gradient-based framework tunes epsilon-greedy exploration schedules by minimizing Bayesian regret, matching or beating heuristics in batched recommendation benchmarks.

Pith tools