R2SM provides the first benchmark pairing modal and amodal text prompts with matching masks, letting models learn when to segment only visible parts versus complete occluded shapes.
LISA: reasoning segmentation via large language model
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
citation-role summary
background 1
citation-polarity summary
fields
cs.CV 1years
2025 1verdicts
CONDITIONAL 1roles
background 1polarities
background 1representative citing papers
citing papers explorer
-
R2SM: Referring and Reasoning for Selective Masks
R2SM provides the first benchmark pairing modal and amodal text prompts with matching masks, letting models learn when to segment only visible parts versus complete occluded shapes.