Pith. sign in

REVIEW 1 cited by

Assessing the Local Interpretability of Machine Learning Models

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 1902.03501 v2 pith:HDNARGBM submitted 2019-02-09 cs.LG cs.HCstat.ML

classification cs.LGcs.HCstat.ML
keywords interpretabilitylocalmodelinterpretablelearningmachinemodelssimulatability
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

The increasing adoption of machine learning tools has led to calls for accountability via model interpretability. But what does it mean for a machine learning model to be interpretable by humans, and how can this be assessed? We focus on two definitions of interpretability that have been introduced in the machine learning literature: simulatability (a user's ability to run a model on a given input) and "what if" local explainability (a user's ability to correctly determine a model's prediction under local changes to the input, given knowledge of the model's original prediction). Through a user study with 1,000 participants, we test whether humans perform well on tasks that mimic the definitions of simulatability and "what if" local explainability on models that are typically considered locally interpretable. To track the relative interpretability of models, we employ a simple metric, the runtime operation count on the simulatability task. We find evidence that as the number of operations increases, participant accuracy on the local interpretability tasks decreases. In addition, this evidence is consistent with the common intuition that decision trees and logistic regression models are interpretable and are more interpretable than neural networks.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. An Empirical Examination of the Evaluative AI Framework

    cs.HC 2024-11 conditional novelty 6.0 of 10

    A pre-registered experiment found that an AI providing only pro and con evidence, without recommendations, did not improve decision performance and was used shallowly by participants.

Pith tools