REVIEW 7 cited by
Last Layer Re-Training is Sufficient for Robustness to Spurious Correlations
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Neural network classifiers can largely rely on simple spurious features, such as backgrounds, to make predictions. However, even in these cases, we show that they still often learn core features associated with the desired attributes of the data, contrary to recent findings. Inspired by this insight, we demonstrate that simple last layer retraining can match or outperform state-of-the-art approaches on spurious correlation benchmarks, but with profoundly lower complexity and computational expenses. Moreover, we show that last layer retraining on large ImageNet-trained models can also significantly reduce reliance on background and texture information, improving robustness to covariate shift, after only minutes of training on a single GPU.
Forward citations
Cited by 7 Pith papers
-
Beyond Metadata: CAPRA for Hidden Subgroup Analysis under Missing Metadata in Medical Imaging
CAPRA calibrates image-derived semantic proxy axes on a small labeled split into a reusable subgroup interface for failure auditing and domain-dependent robust transfer without deployment metadata.
-
Spatially Grounded Concept-Based Image Classification
SEG-MIL-CBM uses CLIP-guided segmentation with attention-based multiple instance learning to build a concept bottleneck model that produces spatially grounded explanations and improves worst-group accuracy on spurious...
-
AIM: Amending Inherent Interpretability via Self-Supervised Masking
AIM uses multi-stage feature guidance for self-supervised masking to improve both interpretability (EPG) and accuracy on vision benchmarks.
-
Controllable Feature Whitening for Hyperparameter-Free Bias Mitigation
Controllable Feature Whitening decorrelates target and bias features via a covariance-based whitening transform, reducing spurious-correlation reliance without adversarial training.
-
Machine Learning from Explanations
A two-stage optimization pipeline that alternates label loss with a KL divergence between feature maps of masked and unmasked inputs improves accuracy and robustness in small-data classification.
-
Detecting Regional Spurious Correlations in Vision Transformers via Token Discarding
A token-discarding method for vision transformers measures whether predictions rely on features outside the object's bounding box, identifying spurious correlations and problematic ImageNet classes.
-
Learning Causality for Modern Machine Learning
A thesis compiling six papers that use causal invariance to improve graph neural networks' out-of-distribution generalization, interpretability, and robustness.
Discussion (0). Sign in to comment.