Pith. sign in

REVIEW 1 cited by

Evaluation of post-hoc interpretability methods in time-series classification

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2202.05656 v2 pith:5IAZHE7V submitted 2022-02-11 cs.LG cs.AIstat.ML

classification cs.LGcs.AIstat.ML
keywords interpretabilitymethodspost-hocresultsquantitativeseveralclassificationcritical
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Post-hoc interpretability methods are critical tools to explain neural-network results. Several post-hoc methods have emerged in recent years, but when applied to a given task, they produce different results, raising the question of which method is the most suitable to provide correct post-hoc interpretability. To understand the performance of each method, quantitative evaluation of interpretability methods is essential. However, currently available frameworks have several drawbacks which hinders the adoption of post-hoc interpretability methods, especially in high-risk sectors. In this work, we propose a framework with quantitative metrics to assess the performance of existing post-hoc interpretability methods in particular in time series classification. We show that several drawbacks identified in the literature are addressed, namely dependence on human judgement, retraining, and shift in the data distribution when occluding samples. We additionally design a synthetic dataset with known discriminative features and tunable complexity. The proposed methodology and quantitative metrics can be used to understand the reliability of interpretability methods results obtained in practical applications. In turn, they can be embedded within operational workflows in critical fields that require accurate interpretability results for e.g., regulatory policies.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. ST-Tree with Interpretability for Multivariate Time Series Classification

    cs.LG 2024-11 conditional novelty 4.0 of 10

    ST-Tree couples a Swin Transformer feature extractor with a prototype-based neural tree, reporting average accuracy of 0.789 on 10 UEA multivariate time series datasets, with visualizations of node prototypes as evide...

Pith tools