SAB3R unifies 3D reconstruction and open-vocabulary segmentation in a single feed-forward network trained by distilling CLIP and DINOv2 features into MASt3R.
Learn- ing to generate text-grounded mask for open-world semantic segmentation from only image-text pairs
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
citation-role summary
baseline 1
citation-polarity summary
fields
cs.CV 1years
2025 1verdicts
CONDITIONAL 1roles
baseline 1polarities
baseline 1representative citing papers
citing papers explorer
-
SAB3R: Semantic-Augmented Backbone in 3D Reconstruction
SAB3R unifies 3D reconstruction and open-vocabulary segmentation in a single feed-forward network trained by distilling CLIP and DINOv2 features into MASt3R.