← back to paper
arxiv: 2509.25339 · 2 revisions
VisualOverload: Probing Visual Understanding of VLMs in Really Dense Scenes