Pith. sign in

REVIEW 1 cited by

Towards Benchmarking Explainable Artificial Intelligence Methods

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2208.12120 v1 pith:6IFDYF43 submitted 2022-08-25 cs.AI cs.LG

classification cs.AIcs.LG
keywords methodsexplainabilityneuralartificialdecisionsintelligencelearningthey
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

The currently dominating artificial intelligence and machine learning technology, neural networks, builds on inductive statistical learning. Neural networks of today are information processing systems void of understanding and reasoning capabilities, consequently, they cannot explain promoted decisions in a humanly valid form. In this work, we revisit and use fundamental philosophy of science theories as an analytical lens with the goal of revealing, what can be expected, and more importantly, not expected, from methods that aim to explain decisions promoted by a neural network. By conducting a case study we investigate a selection of explainability method's performance over two mundane domains, animals and headgear. Through our study, we lay bare that the usefulness of these methods relies on human domain knowledge and our ability to understand, generalise and reason. The explainability methods can be useful when the goal is to gain further insights into a trained neural network's strengths and weaknesses. If our aim instead is to use these explainability methods to promote actionable decisions or build trust in ML-models they need to be less ambiguous than they are today. In this work, we conclude from our study, that benchmarking explainability methods, is a central quest towards trustworthy artificial intelligence and machine learning.

Discussion (0). Sign in to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Towards generating more interpretable counterfactuals via concept vectors: a preliminary study on chest X-rays

    eess.IV 2025-06 conditional novelty 4.0 of 10

    Concept vectors in an autoencoder's latent space produce stable counterfactual explanations for large chest X-ray pathologies, but the method does not beat the Latent Shift baseline.

Pith tools