Pith. sign in

REVIEW 2 cited by

Spider: A Unified Framework for Context-dependent Concept Segmentation

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2405.01002 v2 pith:EQKUEISV submitted 2024-05-02 cs.CV cs.LG

Spider: A Unified Framework for Context-dependent Concept Segmentation

classification cs.CV cs.LG
keywords spidertaskscontext-dependentconceptconceptsunderstandingcamouflageddifferent
verification ladder T0 review T1 audit T2 compute T3 formal T4 reserved
0 comments
Share X Bluesky LinkedIn Reddit HN
read the original abstract

Different from the context-independent (CI) concepts such as human, car, and airplane, context-dependent (CD) concepts require higher visual understanding ability, such as camouflaged object and medical lesion. Despite the rapid advance of many CD understanding tasks in respective branches, the isolated evolution leads to their limited cross-domain generalisation and repetitive technique innovation. Since there is a strong coupling relationship between foreground and background context in CD tasks, existing methods require to train separate models in their focused domains. This restricts their real-world CD concept understanding towards artificial general intelligence (AGI). We propose a unified model with a single set of parameters, Spider, which only needs to be trained once. With the help of the proposed concept filter driven by the image-mask group prompt, Spider is able to understand and distinguish diverse strong context-dependent concepts to accurately capture the Prompter's intention. Without bells and whistles, Spider significantly outperforms the state-of-the-art specialized models in 8 different context-dependent segmentation tasks, including 4 natural scenes (salient, camouflaged, and transparent objects and shadow) and 4 medical lesions (COVID-19, polyp, breast, and skin lesion with color colonoscopy, CT, ultrasound, and dermoscopy modalities). Besides, Spider shows obvious advantages in continuous learning. It can easily complete the training of new tasks by fine-tuning parameters less than 1\% and bring a tolerable performance degradation of less than 5\% for all old tasks. The source code will be publicly available at \href{https://github.com/Xiaoqi-Zhao-DLUT/Spider-UniCDSeg}{Spider-UniCDSeg}.

discussion (0)

Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

  1. Hierarchical Consistency Learning for Test-time Adaptation in Camouflage Perception

    cs.CV 2026-05 unverdicted novelty 5.0

    Proposes HCL framework with HRR, TAG, and PCC modules for test-time adaptation in camouflaged object detection, claiming consistent outperformance on benchmarks under distribution shifts.

  2. DifferSeg: Towards Diverse Multimodal Binary Segmentation via Differential Perception and Frequency Guidance

    cs.CV 2026-06 unverdicted novelty 4.0

    DifferSeg introduces learnable differential operators for modality fusion and cross-frequency decoder interactions, claiming superior performance over 67 prior methods on 29 datasets across 18 tasks.