FGAseg combines a pixel-text alignment transformer, a text-pixel alignment loss, and similarity-based pseudo-masks to achieve state-of-the-art open-vocabulary segmentation on multiple benchmarks.
Title resolution pending
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.CV 1years
2025 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
FGAseg: Fine-Grained Pixel-Text Alignment for Open-Vocabulary Semantic Segmentation
FGAseg combines a pixel-text alignment transformer, a text-pixel alignment loss, and similarity-based pseudo-masks to achieve state-of-the-art open-vocabulary segmentation on multiple benchmarks.