FaRMamba adds a multi-scale frequency transform module and a reconstruction auxiliary encoder to a Vision Mamba segmentation encoder, reporting improved Dice and mIoU on CAMUS, Mouse-cochlea, and Kvasir-Seg datasets.
Self-Supervised Alignment Learning for Medical Image Segmentation
1 Pith paper cite this work. Polarity classification is still indexing.
abstract
Recently, self-supervised learning (SSL) methods have been used in pre-training the segmentation models for 2D and 3D medical images. Most of these methods are based on reconstruction, contrastive learning and consistency regularization. However, the spatial correspondence of 2D slices from a 3D medical image has not been fully exploited. In this paper, we propose a novel self-supervised alignment learning framework to pre-train the neural network for medical image segmentation. The proposed framework consists of a new local alignment loss and a global positional loss. We observe that in the same 3D scan, two close 2D slices usually contain similar anatomic structures. Thus, the local alignment loss is proposed to make the pixel-level features of matched structures close to each other. Experimental results show that the proposed alignment learning is competitive with existing self-supervised pre-training approaches on CT and MRI datasets, under the setting of limited annotations.
fields
cs.CV 1years
2025 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
FaRMamba: Frequency-based learning and Reconstruction aided Mamba for Medical Segmentation
FaRMamba adds a multi-scale frequency transform module and a reconstruction auxiliary encoder to a Vision Mamba segmentation encoder, reporting improved Dice and mIoU on CAMUS, Mouse-cochlea, and Kvasir-Seg datasets.