REVIEW 2 cited by
CLAD: A Contrastive Learning based Approach for Background Debiasing
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Convolutional neural networks (CNNs) have achieved superhuman performance in multiple vision tasks, especially image classification. However, unlike humans, CNNs leverage spurious features, such as background information to make decisions. This tendency creates different problems in terms of robustness or weak generalization performance. Through our work, we introduce a contrastive learning-based approach (CLAD) to mitigate the background bias in CNNs. CLAD encourages semantic focus on object foregrounds and penalizes learning features from irrelavant backgrounds. Our method also introduces an efficient way of sampling negative samples. We achieve state-of-the-art results on the Background Challenge dataset, outperforming the previous benchmark with a margin of 4.1\%. Our paper shows how CLAD serves as a proof of concept for debiasing of spurious features, such as background and texture (in supplementary material).
Forward citations
Cited by 2 Pith papers
-
Bringing the Context Back into Object Recognition, Robustly
Localizing the foreground before classification and fusing its classifier output with the full-image prediction improves accuracy and robustness to background shifts in supervised and zero-shot VLM recognition.
-
Efficient Calisthenics Skills Classification through Foreground Instance Selection and Depth Estimation
Using YOLO athlete cropping and Depth Anything V2 depth maps, a CNN classifies calisthenics skills with 0.837 accuracy, slightly above a 0.815 OpenPose skeleton baseline.
Discussion (0). Continue with ORCID to comment.