Pith. sign in

REVIEW 1 cited by

CamoFormer: Masked Separable Attention for Camouflaged Object Detection

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2212.06570 v1 pith:YJJK6OCM submitted 2022-12-10 cs.CV

classification cs.CV
keywords camouflagedcamoformerdetectionobjectattentionbackgroundmaskedmethods
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

How to identify and segment camouflaged objects from the background is challenging. Inspired by the multi-head self-attention in Transformers, we present a simple masked separable attention (MSA) for camouflaged object detection. We first separate the multi-head self-attention into three parts, which are responsible for distinguishing the camouflaged objects from the background using different mask strategies. Furthermore, we propose to capture high-resolution semantic representations progressively based on a simple top-down decoder with the proposed MSA to attain precise segmentation results. These structures plus a backbone encoder form a new model, dubbed CamoFormer. Extensive experiments show that CamoFormer surpasses all existing state-of-the-art methods on three widely-used camouflaged object detection benchmarks. There are on average around 5% relative improvements over previous methods in terms of S-measure and weighted F-measure.

Discussion (0). Sign in to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. SAMwave: Wavelet-Driven Feature Enrichment for Effective Adaptation of Segment Anything Model

    cs.CV 2025-07 conditional novelty 6.0 of 10

    Wavelet high-frequency features and real or complex adapters improve SAM and SAM2 on several low-level vision tasks, though gains depend on the wavelet family.

Pith tools