Pith. sign in

REVIEW 4 cited by

Stop Explaining Black Box Machine Learning Models for High Stakes Decisions and Use Interpretable Models Instead

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 1811.10154 v3 pith:EH4MUYBI submitted 2018-11-26 stat.ML cs.LG

classification stat.MLcs.LG
keywords modelsblackinterpretableexplaininglearningmachineboxescreating
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Black box machine learning models are currently being used for high stakes decision-making throughout society, causing problems throughout healthcare, criminal justice, and in other domains. People have hoped that creating methods for explaining these black box models will alleviate some of these problems, but trying to \textit{explain} black box models, rather than creating models that are \textit{interpretable} in the first place, is likely to perpetuate bad practices and can potentially cause catastrophic harm to society. There is a way forward -- it is to design models that are inherently interpretable. This manuscript clarifies the chasm between explaining black boxes and using inherently interpretable models, outlines several key reasons why explainable black boxes should be avoided in high-stakes decisions, identifies challenges to interpretable machine learning, and provides several example applications where interpretable models could potentially replace black box models in criminal justice, healthcare, and computer vision.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 4 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. OpenAlex reports about 44 citations worldwide. Full citation record

  1. Interpretable machine learning of halo gas density profiles: a sensitivity analysis of cosmological hydrodynamical simulations

    astro-ph.GA 2025-12 conditional novelty 6.0 of 10

    Random forests reproduce simulated halo gas density profiles to roughly 80-90% accuracy, and Sobol analysis of those forests ranks halo mass and central gas mass as the dominant predictors across EAGLE, IllustrisTNG, ...

  2. AlphaEval: A Comprehensive and Efficient Evaluation Framework for Formula Alpha Mining

    cs.AI 2025-08 conditional novelty 6.0 of 10

    AlphaEval scores alpha mining models on prediction, stability, robustness, logic, and diversity, replacing backtests with fast parallel metrics that the paper claims align with backtest outcomes.

  3. Foundation Models for Astrophysics

    astro-ph.IM 2026-08 conditional novelty 3.0 of 10

    Astronomical 'foundation models' largely reuse transformers and self-supervised pretraining, but evidence of transfer to new instruments, populations, or tasks remains rare; the paper argues such evidence, not archite...

  4. The State of Post-Hoc Local XAI Techniques for Image Processing: Challenges and Motivations

    cs.CV 2025-01 conditional novelty 1.0 of 10

    A review of post-hoc local XAI techniques for images, covering motivations, challenges, and suggested future directions.

Pith tools