Pith. sign in

REVIEW

Learning to search efficiently for causally near-optimal treatments

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2007.00973 v2 pith:QTSNU6Y4 submitted 2020-07-02 cs.LG stat.ML

classification cs.LGstat.ML
keywords searchlearningtreatmentalgorithmdatafindingmethodsmodel-free
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Finding an effective medical treatment often requires a search by trial and error. Making this search more efficient by minimizing the number of unnecessary trials could lower both costs and patient suffering. We formalize this problem as learning a policy for finding a near-optimal treatment in a minimum number of trials using a causal inference framework. We give a model-based dynamic programming algorithm which learns from observational data while being robust to unmeasured confounding. To reduce time complexity, we suggest a greedy algorithm which bounds the near-optimality constraint. The methods are evaluated on synthetic and real-world healthcare data and compared to model-free reinforcement learning. We find that our methods compare favorably to the model-free baseline while offering a more transparent trade-off between search time and treatment efficacy.

Discussion (0). Continue with ORCID to comment.

Pith tools