Pith. sign in

REVIEW 1 cited by

Dynamically Scaled Temperature in Self-Supervised Contrastive Learning

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2308.01140 v2 pith:C62ZKZ7M submitted 2023-08-02 cs.LG cs.CV

classification cs.LGcs.CV
keywords samplestemperaturecontrastiveself-supervisedalgorithmsdynamicallyfunctioninfonce
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

In contemporary self-supervised contrastive algorithms like SimCLR, MoCo, etc., the task of balancing attraction between two semantically similar samples and repulsion between two samples of different classes is primarily affected by the presence of hard negative samples. While the InfoNCE loss has been shown to impose penalties based on hardness, the temperature hyper-parameter is the key to regulating the penalties and the trade-off between uniformity and tolerance. In this work, we focus our attention on improving the performance of InfoNCE loss in self-supervised learning by proposing a novel cosine similarity dependent temperature scaling function to effectively optimize the distribution of the samples in the feature space. We also provide mathematical analyses to support the construction of such a dynamically scaled temperature function. Experimental evidence shows that the proposed framework outperforms the contrastive loss-based SSL algorithms.

Discussion (0). Sign in to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

  1. Self-supervised Radio Representation Learning: Can we Learn Multiple Tasks?

    eess.SP 2025-09 conditional novelty 5.0 of 10

    Self-supervised contrastive pretraining on unlabeled IQ radio data transfers to both direction-finding and modulation classification, cutting label needs by up to 99.9%.

Pith tools