REVIEW 3 cited by
Predictive Uncertainty Quantification with Missing Covariates
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Predictive uncertainty quantification is crucial in decision-making problems. We investigate how to adequately quantify predictive uncertainty with missing covariates. A bottleneck is that missing values induce heteroskedasticity on the response's predictive distribution given the observed covariates. Thus, we focus on building predictive sets for the response that are valid conditionally to the missing values pattern. We show that this goal is impossible to achieve informatively in a distribution-free fashion, and we propose useful restrictions on the distribution class. Motivated by these hardness results, we characterize how missing values and predictive uncertainty intertwine. Particularly, we rigorously formalize the idea that the more missing values, the higher the predictive uncertainty. Then, we introduce a generalized framework, coined CP-MDA-Nested*, outputting predictive sets in both regression and classification. Under independence between the missing value pattern and both the features and the response (an assumption justified by our hardness results), these predictive sets are valid conditionally to any pattern of missing values. Moreover, it provides great flexibility in the trade-off between statistical variability and efficiency. Finally, we experimentally assess the performances of CP-MDA-Nested* beyond its scope of theoretical validity, demonstrating promising outcomes in more challenging configurations than independence.
Forward citations
Cited by 3 Pith papers
-
Conformal Prediction for Regression with Clipped Outcomes
New conformal methods ClipCQR and ClipCQR+ provide tight marginal and improved conditional coverage for regression with doubly clipped outcomes, converging to a snapped oracle under consistency.
-
Multiply Robust Conformal Risk Control with Coarsened Data
A conformal risk control framework using efficient influence functions yields distribution-free prediction sets for outcomes trained on coarsened, missing, or censored data.
-
Robust Conformal Outlier Detection under Contaminated Reference Data
Label-Trim, an active-labeling method that verifies and removes suspicious points from a contaminated calibration set, recovers power lost to contamination in conformal outlier detection while keeping type-I error nea...
Discussion (0). Continue with ORCID to comment.