REVIEW 3 cited by
Axiom-based Grad-CAM: Towards Accurate Visualization and Explanation of CNNs
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
To have a better understanding and usage of Convolution Neural Networks (CNNs), the visualization and interpretation of CNNs has attracted increasing attention in recent years. In particular, several Class Activation Mapping (CAM) methods have been proposed to discover the connection between CNN's decision and image regions. In spite of the reasonable visualization, lack of clear and sufficient theoretical support is the main limitation of these methods. In this paper, we introduce two axioms -- Conservation and Sensitivity -- to the visualization paradigm of the CAM methods. Meanwhile, a dedicated Axiom-based Grad-CAM (XGrad-CAM) is proposed to satisfy these axioms as much as possible. Experiments demonstrate that XGrad-CAM is an enhanced version of Grad-CAM in terms of conservation and sensitivity. It is able to achieve better visualization performance than Grad-CAM, while also be class-discriminative and easy-to-implement compared with Grad-CAM++ and Ablation-CAM. The code is available at https://github.com/Fu0511/XGrad-CAM.
Forward citations
Cited by 3 Pith papers
-
Divisive Decisions: Improving Salience-Based Training for Generalization in Binary Classification Tasks
Teaching binary classifiers to contrast their true-class and false-class attention maps during training yields modest out-of-domain generalization gains, but results are inconsistent and lack significance testing.
-
What Pixels Are Enough? SEAMS: Sufficiency Saliency via MSE-Preservation Soft-Masks
SEAMS optimizes a compact soft mask so a frozen model’s class score, CLS embedding, or tokens stay nearly the same when non-mask pixels are replaced by self-generated distractor and blur.
-
Explainable Artificial Intelligence in Biomedical Image Analysis: A Comprehensive Survey
A broad modality-aware survey of explainable AI methods for biomedical imaging, covering heatmap, concept, text, and latent-space approaches plus tools, metrics, and vision-language models.
Discussion (0). Continue with ORCID to comment.