Pith. sign in

REVIEW

What is Flagged in Uncertainty Quantification? Latent Density Models for Uncertainty Categorization

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2207.05161 v2 pith:OK5MRLP4 submitted 2022-07-11 cs.LG cs.AI

classification cs.LGcs.AI
keywords examplesuncertaintymethodsdensityquantificationflaggedframeworkmisclassification
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Uncertainty Quantification (UQ) is essential for creating trustworthy machine learning models. Recent years have seen a steep rise in UQ methods that can flag suspicious examples, however, it is often unclear what exactly these methods identify. In this work, we propose a framework for categorizing uncertain examples flagged by UQ methods in classification tasks. We introduce the confusion density matrix -- a kernel-based approximation of the misclassification density -- and use this to categorize suspicious examples identified by a given uncertainty method into three classes: out-of-distribution (OOD) examples, boundary (Bnd) examples, and examples in regions of high in-distribution misclassification (IDM). Through extensive experiments, we show that our framework provides a new and distinct perspective for assessing differences between uncertainty quantification methods, thereby forming a valuable assessment benchmark.

Discussion (0). Sign in to comment.

Pith tools