Fun3DU is a training-free pipeline that uses chain-of-thought language reasoning and vision-language pointing to segment functional objects in 3D, beating open-vocabulary baselines on SceneFun3D.
Scannet: Richly- annotated 3d reconstructions of indoor scenes
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.CV 1years
2024 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
Functionality understanding and segmentation in 3D scenes
Fun3DU is a training-free pipeline that uses chain-of-thought language reasoning and vision-language pointing to segment functional objects in 3D, beating open-vocabulary baselines on SceneFun3D.