Pith. sign in

REVIEW 2 cited by

Policy Learning with Observational Data

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 1702.02896 v6 pith:ZPRBGETN submitted 2017-02-09 math.ST cs.LGecon.EMstat.MLstat.TH

classification math.STcs.LGecon.EMstat.MLstat.TH
keywords dataobservationalpolicycausalconstraintsformtreatmenttreatments
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

In many areas, practitioners seek to use observational data to learn a treatment assignment policy that satisfies application-specific constraints, such as budget, fairness, simplicity, or other functional form constraints. For example, policies may be restricted to take the form of decision trees based on a limited set of easily observable individual characteristics. We propose a new approach to this problem motivated by the theory of semiparametrically efficient estimation. Our method can be used to optimize either binary treatments or infinitesimal nudges to continuous treatments, and can leverage observational data where causal effects are identified using a variety of strategies, including selection on observables and instrumental variables. Given a doubly robust estimator of the causal effect of assigning everyone to treatment, we develop an algorithm for choosing whom to treat, and establish strong guarantees for the asymptotic utilitarian regret of the resulting policy.

Discussion (0). Sign in to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. A Hierarchy of Policy Learning Problems

    stat.ML 2026-07 conditional novelty 6.0 of 10

    Policy existence reduces to improving-policy learning which reduces to optimal-policy learning; under a natural monotonicity condition the first two are separated by a polynomial sample-complexity gap.

  2. From Observational Data to Clinical Recommendations: A Causal Framework for Estimating Patient-level Treatment Effects and Learning Policies

    stat.ML 2025-07 conditional novelty 5.0 of 10

    A causal framework for learning treatment policies with deferral, applied to diuretic dosing in acute heart failure with kidney injury.

Pith tools