REVIEW 1 cited by
Lock in Feedback in Sequential Experiments
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
abstract
We often encounter situations in which an experimenter wants to find, by sequential experimentation, $x_{max} = \arg\max_{x} f(x)$, where $f(x)$ is a (possibly unknown) function of a well controllable variable $x$. Taking inspiration from physics and engineering, we have designed a new method to address this problem. In this paper, we first introduce the method in continuous time, and then present two algorithms for use in sequential experiments. Through a series of simulation studies, we show that the method is effective for finding maxima of unknown functions by experimentation, even when the maximum of the functions drifts or when the signal to noise ratio is low.
Forward citations
Cited by 1 Pith paper
-
Exploring Offline Policy Evaluation for the Continuous-Armed Bandit Problem
A δ-window extension of Li et al.'s offline bandit evaluation lets logged actions near a policy's choice count, giving a biased but rank-preserving (at coarse level) way to compare continuous-armed bandit policies.
Discussion (0). Continue with ORCID to comment.