Pith. sign in

REVIEW 1 cited by

Gradient-Based Interpretability Methods and Binarized Neural Networks

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2106.12569 v1 pith:WZL2M5KY submitted 2021-06-23 cs.CV cs.AIcs.LG

classification cs.CVcs.AIcs.LG
keywords networksbinarizedbnnsinterpretabilitymapsnetworkneuralproduces
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Binarized Neural Networks (BNNs) have the potential to revolutionize the way that deep learning is carried out in edge computing platforms. However, the effectiveness of interpretability methods on these networks has not been assessed. In this paper, we compare the performance of several widely used saliency map-based interpretabilty techniques (Gradient, SmoothGrad and GradCAM), when applied to Binarized or Full Precision Neural Networks (FPNNs). We found that the basic Gradient method produces very similar-looking maps for both types of network. However, SmoothGrad produces significantly noisier maps for BNNs. GradCAM also produces saliency maps which differ between network types, with some of the BNNs having seemingly nonsensical explanations. We comment on possible reasons for these differences in explanations and present it as an example of why interpretability techniques should be tested on a wider range of network types.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. CRITS: Convolutional Rectifier for Interpretable Time Series Classification

    cs.LG 2025-05 conditional novelty 5.0 of 10

    CRITS is an intrinsically interpretable time series classifier whose local saliency maps are the exact per-sample weights of the model, obtained without gradients, perturbations, or upsampling.

Pith tools