REVIEW 17 cited by
Last Layer Re-Training is Sufficient for Robustness to Spurious Correlations
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Neural network classifiers can largely rely on simple spurious features, such as backgrounds, to make predictions. However, even in these cases, we show that they still often learn core features associated with the desired attributes of the data, contrary to recent findings. Inspired by this insight, we demonstrate that simple last layer retraining can match or outperform state-of-the-art approaches on spurious correlation benchmarks, but with profoundly lower complexity and computational expenses. Moreover, we show that last layer retraining on large ImageNet-trained models can also significantly reduce reliance on background and texture information, improving robustness to covariate shift, after only minutes of training on a single GPU.
Forward citations
Cited by 17 Pith papers
-
Subgroups Matter for Robust Bias Mitigation
Subgroup choice determines whether bias mitigation helps or hurts, and the minimum KL divergence to the unbiased test distribution predicts success.
-
Adapting Vision Foundation Models with Cascaded Semantics
A visual prompt tuning method that injects fixed color, texture, and shape features plus cascaded self-attention maps into a frozen ViT improves accuracy across 34 datasets with 0.74% trainable parameters.
-
Beyond Metadata: CAPRA for Hidden Subgroup Analysis under Missing Metadata in Medical Imaging
CAPRA calibrates image-derived semantic proxy axes on a small labeled split into a reusable subgroup interface for failure auditing and domain-dependent robust transfer without deployment metadata.
-
Spatially Grounded Concept-Based Image Classification
SEG-MIL-CBM uses CLIP-guided segmentation with attention-based multiple instance learning to build a concept bottleneck model that produces spatially grounded explanations and improves worst-group accuracy on spurious...
-
AIM: Amending Inherent Interpretability via Self-Supervised Masking
AIM uses multi-stage feature guidance for self-supervised masking to improve both interpretability (EPG) and accuracy on vision benchmarks.
-
Controllable Feature Whitening for Hyperparameter-Free Bias Mitigation
Controllable Feature Whitening decorrelates target and bias features via a covariance-based whitening transform, reducing spurious-correlation reliance without adversarial training.
-
Machine Learning from Explanations
A two-stage optimization pipeline that alternates label loss with a KL divergence between feature maps of masked and unmasked inputs improves accuracy and robustness in small-data classification.
-
The Pitfalls of Memorization: When Memorization Hurts Generalization
Memorization-aware training (MAT) shifts logits using calibrated held-out predictions from an XRM auxiliary model, improving worst-group accuracy under subpopulation shift.
-
Token-Based Detection of Spurious Correlations in Vision Transformers
A token-discarding method for vision transformers measures whether predictions rely on features outside the object's bounding box, identifying spurious correlations and problematic ImageNet classes.
-
Deontological Keyword Bias: The Impact of Modal Expressions on Normative Judgments of Language Models
LLMs systematically treat modal words like 'must' as evidence of obligation even in non-obligatory contexts, more strongly than humans, and a few-shot plus reasoning prompt can lower the rate of such judgments.
-
Bridging Distribution Shift and AI Safety: Conceptual and Methodological Synergies
The paper proposes a one-to-one mapping between six causes of distribution shift and several AI safety issues, arguing for mutual method transfer through aligned definitions.
-
Focusing Image Generation to Mitigate Spurious Correlations
SCGS generates new training images by inpainting over a classifier's misattended background regions, reducing spurious-correlation reliance without group labels.
-
Re-evaluating Group Robustness via Adaptive Class-Specific Scaling
Class-specific score scaling at test time lets a vanilla ERM model match or surpass debiasing methods and yields a scalar robust-average accuracy trade-off metric.
-
Navigating Shortcuts, Spurious Correlations, and Confounders: From Origins via Detection to Mitigation
A unifying taxonomy and formal definition that connects shortcut learning, spurious correlations, Clever Hans behavior, and confounders across detection, mitigation, and datasets.
-
Scalable Out-of-distribution Robustness in the Presence of Unobserved Confounders
Under weak overlap, a latent confounder can be approximately identified from a single proxy or multiple unlabeled sources, and a reweighted mixture-of-experts model adapts to confounder shift.
-
Weight Averaging for Out-of-Distribution Generalization and Few-Shot Domain Adaptation
Gradient-similarity-regularized weight averaging and WA+SAM fine-tuning are tested on OOD and few-shot domain adaptation benchmarks, with mixed results that do not support the claimed improvements.
-
Learning Causality for Modern Machine Learning
A thesis compiling six papers that use causal invariance to improve graph neural networks' out-of-distribution generalization, interpretability, and robustness.
Discussion (0). Continue with ORCID to comment.