Pith. sign in

Interpretability and Explainability: A Machine Learning Zoo Mini-tour

1 Pith paper cite this work. Polarity classification is still indexing.

1 Pith paper citing it
abstract

In this review, we examine the problem of designing interpretable and explainable machine learning models. Interpretability and explainability lie at the core of many machine learning and statistical applications in medicine, economics, law, and natural sciences. Although interpretability and explainability have escaped a clear universal definition, many techniques motivated by these properties have been developed over the recent 30 years with the focus currently shifting towards deep learning methods. In this review, we emphasise the divide between interpretability and explainability and illustrate these two different research directions with concrete examples of the state-of-the-art. The review is intended for a general machine learning audience with interest in exploring the problems of interpretation and explanation beyond logistic regression or random forest variable importance. This work is not an exhaustive literature survey, but rather a primer focusing selectively on certain lines of research which the authors found interesting or informative.

citation-role summary

background 1

citation-polarity summary

fields

cs.LG 1

years

2025 1

verdicts

REJECT 1

roles

background 1

polarities

unclear 1

representative citing papers

Multi-criteria Rank-based Aggregation for Explainable AI

cs.LG · 2025-05-30 · reject · novelty 6.0

A multi-criteria rank-based aggregation method that combines LIME, SHAP, and ANCHOR explanations, weighted by new rank-based complexity, faithfulness, and stability metrics, is proposed and tested on five datasets.

citing papers explorer

Showing 1 of 1 citing paper.

  • Multi-criteria Rank-based Aggregation for Explainable AI cs.LG · 2025-05-30 · reject · none · ref 7 · internal anchor

    A multi-criteria rank-based aggregation method that combines LIME, SHAP, and ANCHOR explanations, weighted by new rank-based complexity, faithfulness, and stability metrics, is proposed and tested on five datasets.