REVIEW 2 cited by
Overcoming Dimensional Collapse in Self-supervised Contrastive Learning for Medical Image Segmentation
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Self-supervised learning (SSL) approaches have achieved great success when the amount of labeled data is limited. Within SSL, models learn robust feature representations by solving pretext tasks. One such pretext task is contrastive learning, which involves forming pairs of similar and dissimilar input samples, guiding the model to distinguish between them. In this work, we investigate the application of contrastive learning to the domain of medical image analysis. Our findings reveal that MoCo v2, a state-of-the-art contrastive learning method, encounters dimensional collapse when applied to medical images. This is attributed to the high degree of inter-image similarity shared between the medical images. To address this, we propose two key contributions: local feature learning and feature decorrelation. Local feature learning improves the ability of the model to focus on the local regions of the image, while feature decorrelation removes the linear dependence among the features. Our experimental findings demonstrate that our contributions significantly enhance the model's performance in the downstream task of medical segmentation, both in the linear evaluation and full fine-tuning settings. This work illustrates the importance of effectively adapting SSL techniques to the characteristics of medical imaging tasks. The source code will be made publicly available at: https://github.com/CAMMA-public/med-moco
Forward citations
Cited by 2 Pith papers
-
Applying Machine Learning Tools for Urban Resilience Against Floods
The paper applies standard machine learning tools to the Climate Disaster Resilience Index to predict 2025 flood resilience in Tehran's District 6, but the predictions are unvalidated and based on a very small dataset.
-
Quantized and Interpretable Learning Scheme for Deep Neural Networks in Classification Task
Saliency-guided training combined with PACT quantization keeps MNIST and CIFAR-10 accuracy near parity with a quantized baseline, while the claimed efficiency and interpretability gains are not directly measured.
Discussion (0). Continue with ORCID to comment.