Pith. sign in

REVIEW 1 cited by

ShortcutProbe: Probing Prediction Shortcuts for Learning Robust Models

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2505.13910 v2 pith:4QLNU3M4 submitted 2025-05-20 cs.LG

classification cs.LG
keywords spuriousbiasmodelpredictionframeworkgrouplabelslearning
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Deep learning models often achieve high performance by inadvertently learning spurious correlations between targets and non-essential features. For example, an image classifier may identify an object via its background that spuriously correlates with it. This prediction behavior, known as spurious bias, severely degrades model performance on data that lacks the learned spurious correlations. Existing methods on spurious bias mitigation typically require a variety of data groups with spurious correlation annotations called group labels. However, group labels require costly human annotations and often fail to capture subtle spurious biases such as relying on specific pixels for predictions. In this paper, we propose a novel post hoc spurious bias mitigation framework without requiring group labels. Our framework, termed ShortcutProbe, identifies prediction shortcuts that reflect potential non-robustness in predictions in a given model's latent space. The model is then retrained to be invariant to the identified prediction shortcuts for improved robustness. We theoretically analyze the effectiveness of the framework and empirically demonstrate that it is an efficient and practical tool for improving a model's robustness to spurious bias on diverse datasets.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Improving Group Robustness on Spurious Correlation via Evidential Alignment

    cs.LG 2025-06 conditional novelty 6.0 of 10

    Evidential Alignment improves worst-group accuracy by upweighting a biased model's high-uncertainty errors and retraining the last layer with a calibration set, without group annotations.

Pith tools