Pith. sign in

REVIEW 9 cited by

SAM2-Adapter: Evaluating & Adapting Segment Anything 2 in Downstream Tasks: Camouflage, Shadow, Medical Image Segmentation, and More

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2408.04579 v2 pith:6WXNNOTZ submitted 2024-08-08 cs.CV

SAM2-Adapter: Evaluating & Adapting Segment Anything 2 in Downstream Tasks: Camouflage, Shadow, Medical Image Segmentation, and More

classification cs.CV
keywords sam2-adaptersegmentationmodelstasksanythingimagemedicalsam2
verification ladder T0 review T1 audit T2 compute T3 formal T4 reserved
0 comments
read the original abstract

The advent of large models, also known as foundation models, has significantly transformed the AI research landscape, with models like Segment Anything (SAM) achieving notable success in diverse image segmentation scenarios. Despite its advancements, SAM encountered limitations in handling some complex low-level segmentation tasks like camouflaged object and medical imaging. In response, in 2023, we introduced SAM-Adapter, which demonstrated improved performance on these challenging tasks. Now, with the release of Segment Anything 2 (SAM2), a successor with enhanced architecture and a larger training corpus, we reassess these challenges. This paper introduces SAM2-Adapter, the first adapter designed to overcome the persistent limitations observed in SAM2 and achieve new state-of-the-art (SOTA) results in specific downstream tasks including medical image segmentation, camouflaged (concealed) object detection, and shadow detection. SAM2-Adapter builds on the SAM-Adapter's strengths, offering enhanced generalizability and composability for diverse applications. We present extensive experimental results demonstrating SAM2-Adapter's effectiveness. We show the potential and encourage the research community to leverage the SAM2 model with our SAM2-Adapter for achieving superior segmentation outcomes. Code, pre-trained models, and data processing protocols are available at http://tianrun-chen.github.io/SAM-Adaptor/

discussion (0)

Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.

Forward citations

Cited by 9 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

  1. Modality-Agnostic Prompt Learning for Multi-Modal Camouflaged Object Detection

    cs.CV 2026-04 unverdicted novelty 7.0

    A framework uses modality-agnostic prompts to adapt SAM for multi-modal camouflaged object detection, with a mask refine module for better boundaries.

  2. When Does Resolution Help a Frozen Backbone? Global Attention at Resolution Predicts Scalable Adaptation for Camouflaged and Marine Animal Segmentation

    cs.CV 2026-07 conditional novelty 6.5

    Global attention over a high-resolution token set, not capacity or pretraining, determines whether LoRA adapters convert resolution into accuracy on fine-grained segmentation.

  3. Refining Context-Entangled Content Segmentation via Curriculum Selection and Anti-Curriculum Promotion

    cs.CV 2026-02 conditional novelty 6.0

    A curriculum-then-anti-curriculum training schedule, ending with spectral low-pass fine-tuning, improves context-entangled segmentation across several datasets and backbones.

  4. From Reconstruction to Decision: A Post-Encoder Plug-in Adapter for Curvilinear Segmentation

    cs.CV 2026-06 unverdicted novelty 5.0

    PEPA is a post-encoder adapter combining target-conditioned snake upsampling and adaptive differentiable thresholding that improves topological metrics over region overlap when added to frozen-encoder curvilinear segm...

  5. M$^4$-SAM: Multi-Modal Mixture-of-Experts with Memory-Augmented SAM for RGB-D Video Salient Object Detection

    cs.CV 2026-05 unverdicted novelty 5.0

    M⁴-SAM equips SAM2 with modality-aware MoE-LoRA, gated multi-level fusion, and pseudo-guided initialization to reach state-of-the-art on RGB-D video salient object detection.

  6. Weight Group-wise Post-Training Quantization for Medical Foundation Model

    cs.CV 2026-04 unverdicted novelty 5.0

    Permutation-COMQ is a new post-training quantization algorithm that reorders weights within layers and uses only dot-product and rounding steps to deliver the highest reported accuracy for 2-, 4-, and 8-bit medical fo...

  7. Multimodal SAM-adapter for Semantic Segmentation

    cs.CV 2025-09 conditional novelty 5.0

    A side-tuning adapter injects RGB-plus-auxiliary-sensor fused features into SAM's encoder, reaching state-of-the-art semantic segmentation on DeLiVER, FMB, and MUSES.

  8. DifferSeg: Towards Diverse Multimodal Binary Segmentation via Differential Perception and Frequency Guidance

    cs.CV 2026-06 unverdicted novelty 4.0

    DifferSeg introduces learnable differential operators for modality fusion and cross-frequency decoder interactions, claiming superior performance over 67 prior methods on 29 datasets across 18 tasks.

  9. BED-SAM2: Boundary-Enhanced-Depth SAM2 via Monocular Geometric Priors

    cs.CV 2026-05 unverdicted novelty 4.0

    BED-SAM2 enhances the SAM2 vision model by integrating monocular geometric priors to improve boundary delineation in object segmentation tasks.