Pith. sign in

REVIEW 1 cited by

Testing robustness of predictions of trained classifiers against naturally occurring perturbations

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2204.10046 v2 pith:URHAWCIF submitted 2022-04-21 cs.LG

classification cs.LG
keywords robustnessapproachperturbationspredictionsthusapplicationindividualinput
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Correctly quantifying the robustness of machine learning models is a central aspect in judging their suitability for specific tasks, and ultimately, for generating trust in them. We address the problem of finding the robustness of individual predictions. We show both theoretically and with empirical examples that a method based on counterfactuals that was previously proposed for this is insufficient, as it is not a valid metric for determining the robustness against perturbations that occur ``naturally'', outside specific adversarial attack scenarios. We propose a flexible approach that models possible perturbations in input data individually for each application. This is then combined with a probabilistic approach that computes the likelihood that a ``real-world'' perturbation will change a prediction, thus giving quantitative information of the robustness of individual predictions of the trained machine learning model. The method does not require access to the internals of the classifier and thus in principle works for any black-box model. It is, however, based on Monte-Carlo sampling and thus only suited for input spaces with small dimensions. We illustrate our approach on the Iris and the Ionosphere datasets, on an application predicting fog at an airport, and on analytically solvable cases.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Establishing and Evaluating Trustworthy AI: Overview and Research Challenges

    cs.LG 2024-11 conditional novelty 3.0 of 10

    A semi-structured literature review synthesizing six trustworthy AI requirements and their evaluation methods, plus cross-cutting research challenges.

Pith tools