Pith. sign in

REVIEW 10 cited by

SAM-Med3D: Towards General-purpose Segmentation Models for Volumetric Medical Images

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2310.15161 v3 pith:D4N4NXF4 submitted 2023-10-23 cs.CV

SAM-Med3D: Towards General-purpose Segmentation Models for Volumetric Medical Images

classification cs.CV
keywords medicalsam-med3dsegmentationdatasetanatomicaldiversegeneral-purposeimages
verification ladder T0 review T1 audit T2 compute T3 formal T4 reserved
0 comments
read the original abstract

Existing volumetric medical image segmentation models are typically task-specific, excelling at specific target but struggling to generalize across anatomical structures or modalities. This limitation restricts their broader clinical use. In this paper, we introduce SAM-Med3D for general-purpose segmentation on volumetric medical images. Given only a few 3D prompt points, SAM-Med3D can accurately segment diverse anatomical structures and lesions across various modalities. To achieve this, we gather and process a large-scale 3D medical image dataset, SA-Med3D-140K, from a blend of public sources and licensed private datasets. This dataset includes 22K 3D images and 143K corresponding 3D masks. Then SAM-Med3D, a promptable segmentation model characterized by the fully learnable 3D structure, is trained on this dataset using a two-stage procedure and exhibits impressive performance on both seen and unseen segmentation targets. We comprehensively evaluate SAM-Med3D on 16 datasets covering diverse medical scenarios, including different anatomical structures, modalities, targets, and zero-shot transferability to new/unseen tasks. The evaluation shows the efficiency and efficacy of SAM-Med3D, as well as its promising application to diverse downstream tasks as a pre-trained model. Our approach demonstrates that substantial medical resources can be utilized to develop a general-purpose medical AI for various potential applications. Our dataset, code, and models are available at https://github.com/uni-medical/SAM-Med3D.

discussion (0)

Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.

Forward citations

Cited by 10 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

  1. TSegAgent: Zero-Shot Tooth Segmentation via Geometry-Aware Vision-Language Agents

    cs.CV 2026-03 unverdicted novelty 7.0

    TSegAgent achieves accurate zero-shot tooth segmentation on 3D dental scans via geometry-aware vision-language reasoning without task-specific training.

  2. SLIP: Segmentation with Low-latency Interactive Prompting for 3D Medical Images

    cs.CV 2026-07 conditional novelty 6.0

    SLIP decouples image encoding from prompt refinement via a patch memory bank, achieving 0.06s latency and reversible prompting for interactive 3D medical segmentation.

  3. VesselSim: learning 3D blood vessel segmentation without expert annotations

    cs.CV 2026-05 unverdicted novelty 6.0

    VesselSim trains a 3D vessel segmentation model exclusively on 16,500 synthetic angiographic volumes generated by stochastic branching simulation and achieves competitive zero-shot performance on real clinical dataset...

  4. ESICA: A Scalable Framework for Text-Guided 3D Medical Image Segmentation

    cs.CV 2026-04 unverdicted novelty 6.0

    ESICA delivers state-of-the-art accuracy on a five-modality 3D medical segmentation benchmark while offering a compact variant with far fewer parameters.

  5. CrossPan: A Comprehensive Benchmark for Cross-Sequence Pancreas MRI Segmentation and Generalization

    cs.CV 2026-04 unverdicted novelty 6.0

    CrossPan benchmark shows cross-sequence MRI domain shifts cause pancreas segmentation models to fail catastrophically, establishing sequence generalization as the primary barrier to clinical deployment over center var...

  6. PGE-SAM: Prompt-Guided Feature Enhancement for Interactive Segmentation under Degradation

    cs.CV 2026-06 unverdicted novelty 5.0

    PGE-SAM adds a Prompt Guidance Generator, multi-scale feature interaction, and foreground reconstruction loss to SAM for better interactive segmentation on degraded images, plus a new DM-Seg benchmark.

  7. Align then Refine: Text-Guided 3D Prostate Lesion Segmentation

    cs.CV 2026-04 unverdicted novelty 5.0

    A text-guided multi-encoder U-Net with alignment loss, heatmap calibration, and confidence-gated cross-attention refiner sets new state-of-the-art 3D prostate lesion segmentation performance on the PI-CAI dataset.

  8. TSegAgent: Zero-Shot Tooth Segmentation via Geometry-Aware Vision-Language Agents

    cs.CV 2026-03 unverdicted novelty 5.0

    TSegAgent performs zero-shot tooth instance segmentation and identification on 3D dental scans via multi-view foundation models plus explicit dental-arch geometric reasoning.

  9. LETT-NeXt: A Lightweight RECIST-Guided Model for 3D CT Lesion Segmentation

    cs.CV 2026-06 unverdicted novelty 4.0

    LETT-NeXt uses RECIST line prompts in a cropped MedNeXt-v2 encoder-decoder to predict 3D lesion masks, reaching DSC 73.9 on hidden test data for a CVPR 2026 segmentation competition.

  10. AMO-ENE: Attention-based Multi-Omics Fusion Model for Outcome Prediction in Extra Nodal Extension and HPV-associated Oropharyngeal Cancer

    eess.IV 2026-04 unverdicted novelty 4.0

    An attention-based fusion model combining semi-supervised CT segmentation, radiomics, and clinical features predicts metastatic recurrence, overall survival, and disease-free survival in HPV+ oropharyngeal cancer with...