Pith. sign in

REVIEW 1 cited by

On Mutual Information in Contrastive Learning for Visual Representations

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2005.13149 v2 pith:YHT77FAZ submitted 2020-05-27 cs.LG cs.CVstat.ML

classification cs.LGcs.CVstat.ML
keywords learningdetectioninformationmutualalgorithmscontrastivenegativerepresentations
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

In recent years, several unsupervised, "contrastive" learning algorithms in vision have been shown to learn representations that perform remarkably well on transfer tasks. We show that this family of algorithms maximizes a lower bound on the mutual information between two or more "views" of an image where typical views come from a composition of image augmentations. Our bound generalizes the InfoNCE objective to support negative sampling from a restricted region of "difficult" contrasts. We find that the choice of negative samples and views are critical to the success of these algorithms. Reformulating previous learning objectives in terms of mutual information also simplifies and stabilizes them. In practice, our new objectives yield representations that outperform those learned with previous approaches for transfer to classification, bounding box detection, instance segmentation, and keypoint detection. % experiments show that choosing more difficult negative samples results in a stronger representation, outperforming those learned with IR, LA, and CMC in classification, bounding box detection, instance segmentation, and keypoint detection. The mutual information framework provides a unifying comparison of approaches to contrastive learning and uncovers the choices that impact representation learning.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. An Augmentation-Aware Theory for Self-Supervised Contrastive Learning

    cs.LG 2025-05 conditional novelty 6.0 of 10

    An augmentation-aware error bound for contrastive learning decomposes the augmentation gap into a minimum same-class distance and a maximum same-image distance, with a trade-off driven by augmentation strength.

Pith tools