Pith. sign in

ISNet: Integrate image-level and semantic-level context for semantic seg- mentation,

1 Pith paper cite this work. Polarity classification is still indexing.

1 Pith paper citing it

citation-role summary

background 1

citation-polarity summary

fields

cs.MM 1

years

2024 1

verdicts

CONDITIONAL 1

roles

background 1

polarities

unclear 1

representative citing papers

Towards Open-Vocabulary Video Semantic Segmentation

cs.MM · 2024-12-12 · conditional · novelty 5.0

OV2VSS, a CLIP-based baseline with spatial-temporal fusion, random-frame enhancement, and video text encoding, segments novel categories in video and beats image-based methods on VSPW and Cityscapes zero-shot.

citing papers explorer

Showing 1 of 1 citing paper.

  • Towards Open-Vocabulary Video Semantic Segmentation cs.MM · 2024-12-12 · conditional · none · ref 2

    OV2VSS, a CLIP-based baseline with spatial-temporal fusion, random-frame enhancement, and video text encoding, segments novel categories in video and beats image-based methods on VSPW and Cityscapes zero-shot.