LoRA-tuned Grounding DINO feeding boxes to a frozen SAM2 achieves strong text-prompted, box-free multi-organ ultrasound segmentation across 18 public datasets including three unseen domains.
Mcv-unet: A modified convolution & transformer hybrid encoder-decoder network with multi-scale information fusion for ultrasound image semantic segmentation,
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.CV 1years
2025 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
Grounding DINO-US-SAM: Text-Prompted Multi-Organ Segmentation in Ultrasound with LoRA-Tuned Vision-Language Models
LoRA-tuned Grounding DINO feeding boxes to a frozen SAM2 achieves strong text-prompted, box-free multi-organ ultrasound segmentation across 18 public datasets including three unseen domains.