Pith. sign in

Paper Citation Record · LEDGER

MM-Spatial: Exploring 3D Spatial Understanding in Multimodal LLMs

As of 23 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 9 inbound Pith citation observations for arXiv:2503.13111.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2503.13111 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 9 of 9 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00

measured 9 of 9 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:55:28.571725Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T23:29:02.895704Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 2ba24ee8-7de6-493b-a41d-1afd98bdbb87 · inbound

SVQA-R1: Reinforcing Spatial Reasoning in MLLMs via View-Consistent Reward Optimization cites this paper.

SVQA-R1: Reinforcing Spatial Reasoning in MLLMs via View-Consistent Reward Optimization MM-Spatial: Exploring 3D Spatial Understanding in Multimodal LLMs

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T11:55:28.571725Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:55:28.571725Z digest=sha256:3d63f9e1581d5121d0cef5a178bd8b75f8a4b512d186cc3a47bb4766a980ed06

Observation c03f60c5-0ec2-4492-94af-b179664f1177 · inbound

BMMR: A Large-Scale Bilingual Multimodal Multi-Discipline Reasoning Dataset cites this paper.

BMMR: A Large-Scale Bilingual Multimodal Multi-Discipline Reasoning Dataset MM-Spatial: Exploring 3D Spatial Understanding in Multimodal LLMs

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-06T20:15:52.729171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:15:52.729171Z digest=sha256:8d270a3b61f776ff4b87159228d6343757587a0f6608e5ad2483dd04f863a86b

Observation 754c7e6f-e5aa-4259-9b5c-544b0f9bf65d · inbound

SpatialThinker: Reinforcing Scene Graph-Grounded Spatial Reasoning via Dense Rewards cites this paper.

SpatialThinker: Reinforcing Scene Graph-Grounded Spatial Reasoning via Dense Rewards MM-Spatial: Exploring 3D Spatial Understanding in Multimodal LLMs

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-03T23:08:49.211667Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T23:08:49.211667Z digest=sha256:f66ca3fde4cf56610b17be76ae315210e94a111a9f02b29ac41dda0c8c0979c4

Observation b8c5681f-aa06-436a-84c2-10c96b4ed535 · inbound

Lost in Space? Vision-Language Models Struggle with Relative Camera Pose Estimation cites this paper.

Lost in Space? Vision-Language Models Struggle with Relative Camera Pose Estimation MM-Spatial: Exploring 3D Spatial Understanding in Multimodal LLMs

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-16T09:32:41.265073Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-16T09:31:39.548562Z digest=sha256:1c2a3ee9e16b2a7e1d13f031171c4073a73b8b105db1a76615de62335c65f30d

Observation c71c2b9b-785e-4ec4-a444-dbe134b89562 · inbound

Multimodal Language Models Cannot Spot Spatial Inconsistencies cites this paper.

Multimodal Language Models Cannot Spot Spatial Inconsistencies MM-Spatial: Exploring 3D Spatial Understanding in Multimodal LLMs

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-13T23:13:24.844162Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-05-13T23:12:27.333405Z digest=sha256:23b4df586dd34a8ed95c604ae6ed9bc4b7d551fdbe123dc9941ac469f1e1078f

Observation 9bafafa0-f3a2-4bfa-b857-04b2dc2f190f · inbound

World2VLM: Distilling World Model Imagination into VLMs for Dynamic Spatial Reasoning cites this paper.

World2VLM: Distilling World Model Imagination into VLMs for Dynamic Spatial Reasoning MM-Spatial: Exploring 3D Spatial Understanding in Multimodal LLMs

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-12T09:46:27.005280Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-07T09:20:36.243949Z digest=sha256:174eb691d8b28d54a15c02654b4b310fc1553a3b50d9d3f1f5da9a7a3adf91ab

Observation a44f4414-e6b3-4238-aa7c-40bf0546cf0e · inbound

SpaCE: Rethinking Spatial Capacity and Generalization in Multi-Frame Multimodal Large Language Models cites this paper.

SpaCE: Rethinking Spatial Capacity and Generalization in Multi-Frame Multimodal Large Language Models MM-Spatial: Exploring 3D Spatial Understanding in Multimodal LLMs

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T23:29:02.897269Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-06-26T22:08:47.285928Z digest=sha256:bc1e8d61b79f8cae908c6560855254c6825576ef7a6e7553fc0fccd7973f9e09

Observation 6a3060fd-8b62-4dba-9c37-fdcf7f12f176 · inbound

Decodable Is Not Grounded: A Vision-Ablation Arbiter for VLM Spatial Reasoning cites this paper.

Decodable Is Not Grounded: A Vision-Ablation Arbiter for VLM Spatial Reasoning MM-Spatial: Exploring 3D Spatial Understanding in Multimodal LLMs

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-07-01T09:35:41.079970Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-07-01T06:23:00.372251Z digest=sha256:aaf590829ffeb494adbd868dd5d855b07a90450c55ff9382e28b64bdfaebbac0

Observation ab675188-d761-4c2c-b357-f5b59bc398cb · inbound

GReFEM: Multimodal LLMs as Zero-Shot Semantic Assistants for Physics-Guided 3D Mesh Refinement cites this paper.

GReFEM: Multimodal LLMs as Zero-Shot Semantic Assistants for Physics-Guided 3D Mesh Refinement MM-Spatial: Exploring 3D Spatial Understanding in Multimodal LLMs

Reference 51

Resolution
unresolved
no resolver link, observed 2026-07-13T06:42:45.558324Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-13T06:42:45.558324Z digest=sha256:1d94cec833d547eab3ad12f49e8edf14271251de8f9de64472dc32d4bdf9d3f5