PR-MaGIC refines prompts in in-context segmentation via test-time gradient flow from the mask decoder plus top-1 selection, yielding better masks across benchmarks without training.
hub
Medical sam adapter: Adapting seg- 10 ment anything model for medical image segmentation
14 Pith papers cite this work. Polarity classification is still indexing.
hub tools
citation-role summary
citation-polarity summary
roles
background 2polarities
background 2representative citing papers
Presents MMIO benchmark and RTVP method achieving state-of-the-art 42.2% AP in zero-shot industrial defect detection.
DeCoDrift stabilizes decoder coupling in closed-loop foundation segmentation by constraining prompt updates without retraining or ground truth.
SAMamba3D adapts a frozen SAM encoder with Mamba volumetric context and cross-scale features to match or exceed 3D baselines on diverse sandstone and carbonate datasets while reducing case-specific retraining.
SGPer combines DINOv2 semantic priors converted to dense prompts with SAM geometric priors through disease-sensitive adapters and dynamic consistency filtering to deliver robust limited-data wheat disease segmentation.
Presents COMMA, a coordinate-aware Mamba network for 3D vessel segmentation that uses global and local branches, along with a new 570-case labeled dataset.
SAM 2 delivers more accurate video segmentation with 3x fewer user interactions and 6x faster image segmentation than the original SAM by training a streaming-memory transformer on the largest video segmentation dataset collected to date.
Empirical tests show adapters (2-3 per block) and LoRA on deformable attention achieve competitive instance segmentation with 1-6% parameters tuned versus 40-55% for full fine-tuning.
ANAUS introduces anatomy-anchored self-supervision with LP-SAM delineation and dual policies (inter-view anatomy alignment plus core-region prediction) to distill invariant ultrasound representations, claiming SOTA results on six datasets.
DiffKT3D transfers priors from video diffusion models to 3D radiotherapy dose prediction via modality-specific embeddings and clinically guided RL, reducing voxel MAE from 2.07 to 1.93 and claiming SOTA over the GDP-HMM challenge winner.
DRD introduces a reprogramming module and CKA-based distillation to enable efficient, robust adaptation of medical foundation models to downstream 2D/3D classification and segmentation tasks, outperforming prior PEFT and KD methods on 18 tasks.
A text-guided multi-encoder U-Net with alignment loss, heatmap calibration, and confidence-gated cross-attention refiner sets new state-of-the-art 3D prostate lesion segmentation performance on the PI-CAI dataset.
GleSAM++ improves SAM robustness on degraded images by using generative enhancement, feature alignment, and adaptive degradation prediction while adding few parameters.
LoRA-adapted SAM 3 with hard-negative mining and phase-coherent filtering achieves median Dice 0.968 on pulmonary structures from 4DCT using seven annotated volumes.
citing papers explorer
-
PR-MaGIC: Prompt Refinement Via Mask Decoder Gradient Flow For In-Context Segmentation
PR-MaGIC refines prompts in in-context segmentation via test-time gradient flow from the mask decoder plus top-1 selection, yielding better masks across benchmarks without training.
-
Zero-Shot Learning in Industrial Scenarios: New Large-Scale Benchmark, Challenges and Baseline
Presents MMIO benchmark and RTVP method achieving state-of-the-art 42.2% AP in zero-shot industrial defect detection.
-
DeCoDrift: Stabilizing Decoder Coupling in Closed-Loop Foundation Segmentation
DeCoDrift stabilizes decoder coupling in closed-loop foundation segmentation by constraining prompt updates without retraining or ground truth.
-
SAMamba3D: adapting Segment Anything for generalizable 3D segmentation of multiphase pore-scale images
SAMamba3D adapts a frozen SAM encoder with Mamba volumetric context and cross-scale features to match or exceed 3D baselines on diverse sandstone and carbonate datasets while reducing case-specific retraining.
-
Learning to Synergize Semantic and Geometric Priors for Limited-Data Wheat Disease Segmentation
SGPer combines DINOv2 semantic priors converted to dense prompts with SAM geometric priors through disease-sensitive adapters and dynamic consistency filtering to deliver robust limited-data wheat disease segmentation.
-
COMMA: Coordinate-aware Modulated Mamba Network for 3D Dispersed Vessel Segmentation
Presents COMMA, a coordinate-aware Mamba network for 3D vessel segmentation that uses global and local branches, along with a new 570-case labeled dataset.
-
SAM 2: Segment Anything in Images and Videos
SAM 2 delivers more accurate video segmentation with 3x fewer user interactions and 6x faster image segmentation than the original SAM by training a streaming-memory transformer on the largest video segmentation dataset collected to date.
-
Parameter-Efficient Fine-Tuning of Large Pretrained Models for Instance Segmentation Tasks
Empirical tests show adapters (2-3 per block) and LoRA on deformable attention achieve competitive instance segmentation with 1-6% parameters tuned versus 40-55% for full fine-tuning.
-
Anatomy-Anchored Self-Supervision: Distilling Vision Foundation Models for Invariant Ultrasound Representation
ANAUS introduces anatomy-anchored self-supervision with LP-SAM delineation and dual policies (inter-view anatomy alignment plus core-region prediction) to distill invariant ultrasound representations, claiming SOTA results on six datasets.
-
Any2Any 3D Diffusion Models with Knowledge Transfer: A Radiotherapy Planning Study
DiffKT3D transfers priors from video diffusion models to 3D radiotherapy dose prediction via modality-specific embeddings and clinically guided RL, reducing voxel MAE from 2.07 to 1.93 and claiming SOTA over the GDP-HMM challenge winner.
-
Deep Reprogramming Distillation for Medical Foundation Models
DRD introduces a reprogramming module and CKA-based distillation to enable efficient, robust adaptation of medical foundation models to downstream 2D/3D classification and segmentation tasks, outperforming prior PEFT and KD methods on 18 tasks.
-
Align then Refine: Text-Guided 3D Prostate Lesion Segmentation
A text-guided multi-encoder U-Net with alignment loss, heatmap calibration, and confidence-gated cross-attention refiner sets new state-of-the-art 3D prostate lesion segmentation performance on the PI-CAI dataset.
-
Towards Any-Quality Image Segmentation via Generative and Adaptive Latent Space Enhancement
GleSAM++ improves SAM robustness on degraded images by using generative enhancement, feature alignment, and adaptive degradation prediction while adding few parameters.
-
Parameter-Efficient Adaptation of SAM 3 for Automated ITV Generation from 4DCT Images
LoRA-adapted SAM 3 with hard-negative mining and phase-coherent filtering achieves median Dice 0.968 on pulmonary structures from 4DCT using seven annotated volumes.