Pith. sign in

REVIEW 1 cited by

Saliency is a Possible Red Herring When Diagnosing Poor Generalization

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 1910.00199 v3 pith:FJ7XAVGG submitted 2019-10-01 cs.CV cs.LGeess.IV

classification cs.CVcs.LGeess.IV
keywords generalizationfeaturestrainingattributionimagepoorsaliencyexpert
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Poor generalization is one symptom of models that learn to predict target variables using spuriously-correlated image features present only in the training distribution instead of the true image features that denote a class. It is often thought that this can be diagnosed visually using attribution (aka saliency) maps. We study if this assumption is correct. In some prediction tasks, such as for medical images, one may have some images with masks drawn by a human expert, indicating a region of the image containing relevant information to make the prediction. We study multiple methods that take advantage of such auxiliary labels, by training networks to ignore distracting features which may be found outside of the region of interest. This mask information is only used during training and has an impact on generalization accuracy depending on the severity of the shift between the training and test distributions. Surprisingly, while these methods improve generalization performance in the presence of a covariate shift, there is no strong correspondence between the correction of attribution towards the features a human expert has labelled as important and generalization performance. These results suggest that the root cause of poor generalization may not always be spatially defined, and raise questions about the utility of masks as "attribution priors" as well as saliency maps for explainable predictions.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Towards generating more interpretable counterfactuals via concept vectors: a preliminary study on chest X-rays

    eess.IV 2025-06 conditional novelty 4.0 of 10

    Concept vectors in an autoencoder's latent space produce stable counterfactual explanations for large chest X-ray pathologies, but the method does not beat the Latent Shift baseline.

Pith tools