Pith. sign in

hub Mixed citations

Omnispatial: Towards comprehensive spatial reasoning benchmark for vision language models

Mixed citation behavior. Most common role is background (62%).

33 Pith papers citing it
1 external citations · external index
Background 62% of classified citations

hub tools

citation-role summary

background 5 dataset 2 baseline 1

citation-polarity summary

years

2026 29 2025 4

representative citing papers

SATURN: Symbolic Spatial Reasoning for Multi-Perspective Grounding

cs.CV · 2026-06-21 · unverdicted · novelty 7.0

SATURN reconstructs approximate 3D scenes, derives soft perspective-aware predicates, and executes them symbolically to achieve stable performance on complex multi-perspective spatial grounding tasks where VLMs degrade.

Why MLLMs Struggle to Determine Object Orientations

cs.CV · 2026-04-14 · accept · novelty 7.0

Orientation information is recoverable from MLLM visual encoder embeddings via linear regression, contradicting the hypothesis that failures originate in the encoders.

SCP: Spatial Causal Prediction in Video

cs.CV · 2026-03-04 · unverdicted · novelty 7.0

SCP defines a new benchmark task for predicting spatial causal outcomes beyond direct observation and shows that 23 leading models lag far behind humans on it.

citing papers explorer

Showing 33 of 33 citing papers.