ExACT combines a Vision Exemplar-based Calibrator and Structure-Aware Refiner to improve training-free visual grounding of language descriptions in remote sensing images using frozen MLLMs and SAM.
Vi- sion and language reference prompt into sam for few-shot segmentation
2 Pith papers cite this work. Polarity classification is still indexing.
fields
cs.CV 2verdicts
UNVERDICTED 2representative citing papers
V2-SAM adapts SAM2 to cross-view object correspondence with geometry-aware and appearance-based prompt generators plus a post-hoc cyclic consistency selector, reporting new state-of-the-art results on Ego-Exo4D, DAVIS-2017, and HANDAL-X.
citing papers explorer
-
ExACT: Exemplar-Driven Calibrated Refinement for Training-Free Visual Grounding in Remote Sensing Images
ExACT combines a Vision Exemplar-based Calibrator and Structure-Aware Refiner to improve training-free visual grounding of language descriptions in remote sensing images using frozen MLLMs and SAM.
-
V$^{2}$-SAM: Marrying SAM2 with Multi-Prompt Experts for Cross-View Object Correspondence
V2-SAM adapts SAM2 to cross-view object correspondence with geometry-aware and appearance-based prompt generators plus a post-hoc cyclic consistency selector, reporting new state-of-the-art results on Ego-Exo4D, DAVIS-2017, and HANDAL-X.