Under covariate shift, with outcomes only in the source domain, the paper derives a doubly robust semiparametric efficient estimator of the target reward and uses it to learn a treatment policy.
Predicting the efficacy of future training programs using past experiences at other locations
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.LG 1years
2025 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
Optimal Policy Adaptation under Covariate Shift
Under covariate shift, with outcomes only in the source domain, the paper derives a doubly robust semiparametric efficient estimator of the target reward and uses it to learn a treatment policy.