REVIEW 10 cited by
Contrastive Learning with Hard Negative Samples
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
How can you sample good negative examples for contrastive learning? We argue that, as with metric learning, contrastive learning of representations benefits from hard negative samples (i.e., points that are difficult to distinguish from an anchor point). The key challenge toward using hard negatives is that contrastive methods must remain unsupervised, making it infeasible to adopt existing negative sampling strategies that use true similarity information. In response, we develop a new family of unsupervised sampling methods for selecting hard negative samples where the user can control the hardness. A limiting case of this sampling results in a representation that tightly clusters each class, and pushes different classes as far apart as possible. The proposed method improves downstream performance across multiple modalities, requires only few additional lines of code to implement, and introduces no computational overhead.
Forward citations
Cited by 10 Pith papers
-
Fusion Embedding: A Unified Embedding Space for Text, Image, Video, and Audio
Trained connectors and audio-only gated adapters integrate audio into a frozen vision-language embedding space, preserving base outputs bit-exactly and yielding emergent audio-image retrieval.
-
DisProtEdit: Exploring Disentangled Representations for Multi-Attribute Protein Editing
DisProtEdit learns disentangled protein representations from separate structural and functional text descriptions, enabling controllable single- and multi-attribute protein editing via latent interpolation.
-
Discriminative Axis, Not Data Volume: What a Contrastive Corpus Teaches an Audio Embedding
A contrastive audio embedding learns an attribute only when in-batch negatives cannot be separated without it, so corpus structure, not size or caption vocabulary, controls what is encoded.
-
Generalizable Object Re-Identification via Visual In-Context Prompting
VICP uses an LLM to generate per-category visual prompts for a frozen DINOv2, enabling few-shot generalization to unseen object categories in re-identification without parameter updates.
-
QuRe: Query-Relevant Retrieval through Hard Negative Sampling in Composed Image Retrieval
QuRe trains composed image retrieval models with a pairwise reward objective on hard negatives found between sharp relevance-score drops, and adds a human-preference benchmark for evaluating retrieval relevance.
-
Multimodal Semantic-Aware Contrastive Learning For False Negative Mitigation in 3D Medical Imaging
MseaCL, a semantic-aware contrastive pretraining method for 3D brain MRI and reports, reports a 0.226 external AUC gain over instance-based contrastive learning for pediatric brain tumor molecular classification.
-
A Generalized Learning Framework for Self-Supervised Contrastive Learning
A single framework unifies BYOL, Barlow Twins, and SwAV, plus a plug-in calibration method, ADC, that improves learned representations by preserving input-space distances.
-
The Impact of Semantic Pairs on Self-Supervised Representation Learning
Semantic positive pairs (same-class, different images) consistently beat augmented-view pairs for self-supervised pretraining in controlled ImageNet-subset experiments, with the largest gains for SimCLR.
-
eMargin: Revisiting Contrastive Learning with Margin-Based Separation
An adaptive margin added to InfoNCE improves time series clustering metrics but hurts linear-probe classification, exposing a disconnect between clustering scores and downstream utility.
-
Subject Invariant Contrastive Learning for Human Activity Recognition
A subject-reweighted contrastive loss improves cross-subject generalization for human activity recognition across unimodal, multimodal, and supervised settings.
Discussion (0). Sign in to comment.