Pith. sign in

REVIEW 4 cited by

Know Your Self-supervised Learning: A Survey on Image-based Generative and Discriminative Training

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2305.13689 v1 pith:Y7KRSIMD submitted 2023-05-23 cs.CV cs.AI

classification cs.CVcs.AI
keywords learningdiscriminativeimage-basedgenerativeself-supervisedyearscomputerframeworks
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
abstract

Although supervised learning has been highly successful in improving the state-of-the-art in the domain of image-based computer vision in the past, the margin of improvement has diminished significantly in recent years, indicating that a plateau is in sight. Meanwhile, the use of self-supervised learning (SSL) for the purpose of natural language processing (NLP) has seen tremendous successes during the past couple of years, with this new learning paradigm yielding powerful language models. Inspired by the excellent results obtained in the field of NLP, self-supervised methods that rely on clustering, contrastive learning, distillation, and information-maximization, which all fall under the banner of discriminative SSL, have experienced a swift uptake in the area of computer vision. Shortly afterwards, generative SSL frameworks that are mostly based on masked image modeling, complemented and surpassed the results obtained with discriminative SSL. Consequently, within a span of three years, over $100$ unique general-purpose frameworks for generative and discriminative SSL, with a focus on imaging, were proposed. In this survey, we review a plethora of research efforts conducted on image-oriented SSL, providing a historic view and paying attention to best practices as well as useful software packages. While doing so, we discuss pretext tasks for image-based SSL, as well as techniques that are commonly used in image-based SSL. Lastly, to aid researchers who aim at contributing to image-focused SSL, we outline a number of promising research directions.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 4 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Color Flow Imaging Microscopy Improves Identification of Stress Sources of Protein Aggregates in Biopharmaceuticals

    cs.CV 2025-01 conditional novelty 6.0 of 10

    Deep learning trained on color flow imaging microscopy images identifies stress sources of protein aggregates more accurately than on grayscale images.

  2. An efficient unsupervised classification model for galaxy morphology: Voting clustering based on coding from ConvNeXt large model

    astro-ph.GA 2024-12 conditional novelty 5.0 of 10

    An unsupervised pipeline using ConvNeXt encoding, PCA, and multi-model voting classifies about 53% of COSMOS galaxies into 20 clusters, later merged into five morphology types.

  3. Identifying Critical Tokens for Accurate Predictions in Transformer-based Medical Imaging Models

    cs.CV 2025-01 conditional novelty 4.0 of 10

    Token Insight iteratively removes the image token that most reduces a vision transformer's polyp-class confidence until the prediction flips, exposing the patches that drive the decision.

  4. Handwritten Text Recognition: A Survey

    cs.CV 2025-02 conditional novelty 3.0 of 10

    A survey of handwritten text recognition that categorizes methods by reading-order complexity and compares reported performance on the IAM database.

Pith tools