Pith. sign in

REVIEW 2 cited by

Utilizing Explainable AI for Quantization and Pruning of Deep Neural Networks

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2008.09072 v1 pith:2UDE2CLC submitted 2020-08-20 cs.CV cs.AIcs.LG

classification cs.CVcs.AIcs.LG
keywords explainablepruningmethodsquantizationcompressiondeepdnnsclassification
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

For many applications, utilizing DNNs (Deep Neural Networks) requires their implementation on a target architecture in an optimized manner concerning energy consumption, memory requirement, throughput, etc. DNN compression is used to reduce the memory footprint and complexity of a DNN before its deployment on hardware. Recent efforts to understand and explain AI (Artificial Intelligence) methods have led to a new research area, termed as explainable AI. Explainable AI methods allow us to understand better the inner working of DNNs, such as the importance of different neurons and features. The concepts from explainable AI provide an opportunity to improve DNN compression methods such as quantization and pruning in several ways that have not been sufficiently explored so far. In this paper, we utilize explainable AI methods: mainly DeepLIFT method. We use these methods for (1) pruning of DNNs; this includes structured and unstructured pruning of \ac{CNN} filters pruning as well as pruning weights of fully connected layers, (2) non-uniform quantization of DNN weights using clustering algorithm; this is also referred to as Weight Sharing, and (3) integer-based mixed-precision quantization; this is where each layer of a DNN may use a different number of integer bits. We use typical image classification datasets with common deep learning image classification models for evaluation. In all these three cases, we demonstrate significant improvements as well as new insights and opportunities from the use of explainable AI in DNN compression.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Relevance-driven Input Dropout: an Explanation-guided Regularization Technique

    cs.LG 2025-05 conditional novelty 6.0 of 10

    RelDrop, which occludes the most attribution-relevant input regions during training, improves generalization and occlusion robustness for image and point cloud classification.

  2. Compressing Deep Neural Networks Using Explainable AI

    cs.LG 2025-07 reject novelty 4.0 of 10

    Using LRP scores, the method prunes negative-relevance filters and quantizes remaining weights by per-layer median, yielding 90.5% accuracy at 1.2 MB on a synthetic 4-class task.

Pith tools