Pith. sign in

REVIEW 1 cited by

Explaining Machine Learning Classifiers through Diverse Counterfactual Explanations

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 1905.07697 v2 pith:OKPNNQTN submitted 2019-05-19 cs.LG cs.CYstat.ML

classification cs.LGcs.CYstat.ML
keywords counterfactualsexplanationscounterfactualdiverseframeworkgeneratinglearninglocal
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Post-hoc explanations of machine learning models are crucial for people to understand and act on algorithmic predictions. An intriguing class of explanations is through counterfactuals, hypothetical examples that show people how to obtain a different prediction. We posit that effective counterfactual explanations should satisfy two properties: feasibility of the counterfactual actions given user context and constraints, and diversity among the counterfactuals presented. To this end, we propose a framework for generating and evaluating a diverse set of counterfactual explanations based on determinantal point processes. To evaluate the actionability of counterfactuals, we provide metrics that enable comparison of counterfactual-based methods to other local explanation methods. We further address necessary tradeoffs and point to causal implications in optimizing for counterfactuals. Our experiments on four real-world datasets show that our framework can generate a set of counterfactuals that are diverse and well approximate local decision boundaries, outperforming prior approaches to generating diverse counterfactuals. We provide an implementation of the framework at https://github.com/microsoft/DiCE.

Discussion (0). Sign in to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Tabular Diffusion based Actionable Counterfactual Explanations for Network Intrusion Detection

    cs.LG 2025-07 conditional novelty 4.0 of 10

    A diffusion-based counterfactual explanation method for network intrusion detection, with distilled fast sampling and decision-tree global rules, is evaluated against six baselines on three NIDS datasets.

Pith tools