REVIEW 3 cited by
Deep pNML: Predictive Normalized Maximum Likelihood for Deep Neural Networks
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
The Predictive Normalized Maximum Likelihood (pNML) scheme has been recently suggested for universal learning in the individual setting, where both the training and test samples are individual data. The goal of universal learning is to compete with a ``genie'' or reference learner that knows the data values, but is restricted to use a learner from a given model class. The pNML minimizes the associated regret for any possible value of the unknown label. Furthermore, its min-max regret can serve as a pointwise measure of learnability for the specific training and data sample. In this work we examine the pNML and its associated learnability measure for the Deep Neural Network (DNN) model class. As shown, the pNML outperforms the commonly used Empirical Risk Minimization (ERM) approach and provides robustness against adversarial attacks. Together with its learnability measure it can detect out of distribution test examples, be tolerant to noisy labels and serve as a confidence measure for the ERM. Finally, we extend the pNML to a ``twice universal'' solution, that provides universality for model class selection and generates a learner competing with the best one from all model classes.
Forward citations
Cited by 3 Pith papers
-
Skillful joint probabilistic weather forecasting from marginals
FGN, a neural weather model trained only on per-location forecast scores, produces more accurate global ensemble forecasts than GenCast and captures realistic spatial correlations.
-
Functional Risk Minimization
FRM replaces output-space losses with function-space losses, fitting a per-data-point function and approximating the resulting objective with Taylor/Laplace expansions, yielding weighted least squares with a Jacobian-...
-
Quantifying the Prediction Uncertainty of Machine Learning Models for Individual Data
A per-sample confidence score derived from the pNML min-max regret is applied to linear regression and neural networks, and improves OOD detection, adversarial robustness, and active learning.
Discussion (0). Continue with ORCID to comment.