Pith. sign in

Creating such a system is one of the main goals of computer vision and can enable autonomous robots, cars, and numerous other applications[6]

1 Pith paper cite this work. Polarity classification is still indexing.

1 Pith paper citing it

citation-role summary

background 1

citation-polarity summary

fields

cs.CV 1

years

2024 1

verdicts

CONDITIONAL 1

roles

background 1

polarities

background 1

representative citing papers

Do large language vision models understand 3D shapes?

cs.CV · 2024-12-14 · conditional · novelty 6.0

A large synthetic benchmark shows vision-language models match 3D shapes well across single changes like rotation or texture, but fail when rotation and texture change together, trailing humans by a wide margin.

citing papers explorer

Showing 1 of 1 citing paper.

  • Do large language vision models understand 3D shapes? cs.CV · 2024-12-14 · conditional · none · ref 1

    A large synthetic benchmark shows vision-language models match 3D shapes well across single changes like rotation or texture, but fail when rotation and texture change together, trailing humans by a wide margin.