A CLIP-style distillation from microscopy to transcriptomics, combined with a perturbation-embedding augmentation, improves biological relationship recall on out-of-distribution datasets while preserving interpretability.
DDA: Dimensionality Driven Augmentation Search for Contrastive Learning in Laparoscopic Surgery
1 Pith paper cite this work. Polarity classification is still indexing.
abstract
Self-supervised learning (SSL) has potential for effective representation learning in medical imaging, but the choice of data augmentation is critical and domain-specific. It remains uncertain if general augmentation policies suit surgical applications. In this work, we automate the search for suitable augmentation policies through a new method called Dimensionality Driven Augmentation Search (DDA). DDA leverages the local dimensionality of deep representations as a proxy target, and differentiably searches for suitable data augmentation policies in contrastive learning. We demonstrate the effectiveness and efficiency of DDA in navigating a large search space and successfully identifying an appropriate data augmentation policy for laparoscopic surgery. We systematically evaluate DDA across three laparoscopic image classification and segmentation tasks, where it significantly improves over existing baselines. Furthermore, DDA's optimised set of augmentations provides insight into domain-specific dependencies when applying contrastive learning in medical applications. For example, while hue is an effective augmentation for natural images, it is not advantageous for laparoscopic images.
fields
cs.LG 1years
2025 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
A Cross Modal Knowledge Distillation & Data Augmentation Recipe for Improving Transcriptomics Representations through Morphological Features
A CLIP-style distillation from microscopy to transcriptomics, combined with a perturbation-embedding augmentation, improves biological relationship recall on out-of-distribution datasets while preserving interpretability.