Pith. sign in

REVIEW 1 cited by

Benchmarking Deep Learning Interpretability in Time Series Predictions

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2010.13924 v1 pith:JJDWEGZN submitted 2020-10-26 cs.LG stat.ML

classification cs.LGstat.ML
keywords timeimportancemethodssaliencyfeatureseriesdatafeatures
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Saliency methods are used extensively to highlight the importance of input features in model predictions. These methods are mostly used in vision and language tasks, and their applications to time series data is relatively unexplored. In this paper, we set out to extensively compare the performance of various saliency-based interpretability methods across diverse neural architectures, including Recurrent Neural Network, Temporal Convolutional Networks, and Transformers in a new benchmark of synthetic time series data. We propose and report multiple metrics to empirically evaluate the performance of saliency methods for detecting feature importance over time using both precision (i.e., whether identified features contain meaningful signals) and recall (i.e., the number of features with signal identified as important). Through several experiments, we show that (i) in general, network architectures and saliency methods fail to reliably and accurately identify feature importance over time in time series data, (ii) this failure is mainly due to the conflation of time and feature domains, and (iii) the quality of saliency maps can be improved substantially by using our proposed two-step temporal saliency rescaling (TSR) approach that first calculates the importance of each time step before calculating the importance of each feature at a time step.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. ST-Tree with Interpretability for Multivariate Time Series Classification

    cs.LG 2024-11 conditional novelty 4.0 of 10

    ST-Tree couples a Swin Transformer feature extractor with a prototype-based neural tree, reporting average accuracy of 0.789 on 10 UEA multivariate time series datasets, with visualizations of node prototypes as evide...

Pith tools