Pith. sign in

Paper Citation Record · LEDGER

VLM4D: Towards Spatiotemporal Awareness in Vision Language Models

As of 5 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2508.02095.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.02095 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-04T10:37:41.277166Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-25T05:46:39.977556Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 52903efc-fbc8-49af-b197-4bdc8b134722 · inbound

SpaceVista: All-Scale Visual Spatial Reasoning from mm to km cites this paper.

SpaceVista: All-Scale Visual Spatial Reasoning from mm to km VLM4D: Towards Spatiotemporal Awareness in Vision Language Models

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-04T10:37:41.277166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T10:37:41.277166Z digest=sha256:84983fa26c717ef66b7592aa62c11ab86dc18bb9c19d345c1b929b253b9622a2

Observation b9d7bddd-8785-4d5c-bb5e-1031eda8cec5 · inbound

World Simulation with Video Foundation Models for Physical AI cites this paper.

World Simulation with Video Foundation Models for Physical AI VLM4D: Towards Spatiotemporal Awareness in Vision Language Models

Reference 100

Resolution
verified exact
arxiv_id, observed 2026-05-12T23:01:13.883037Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-12T23:01:13.546110Z digest=sha256:2f9b95fc8a1bfe09a823e94132cefc6205dcc5cc75f0131c906f3f57c855eb75

Observation b428e44c-bccd-4e54-903b-c1e3f9ef82b3 · inbound

SpatialBench: Benchmarking Multimodal Large Language Models for Spatial Cognition cites this paper.

SpatialBench: Benchmarking Multimodal Large Language Models for Spatial Cognition VLM4D: Towards Spatiotemporal Awareness in Vision Language Models

Reference 74

Resolution
verified exact
arxiv_id, observed 2026-05-17T04:59:04.226528Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-17T04:54:59.903644Z digest=sha256:654fb12054a71cb4ba5fcf358a2182546f936fac9be32e3eb94ea49866ccb877

Observation 261994e6-167f-4bdc-b274-17dfea95c1c4 · inbound

The TIME Machine: On The Power of Motion for Efficient Perception cites this paper.

The TIME Machine: On The Power of Motion for Efficient Perception VLM4D: Towards Spatiotemporal Awareness in Vision Language Models

Reference 49

Resolution
verified exact
arxiv_id, observed 2026-05-25T05:46:39.980679Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-25T05:40:33.752341Z digest=sha256:90a0faaf1967f6bfb259b3ac93e0e545ecb4a4f347e2d834a49d2c9ddb6cebd6

Observation febf3a8a-1ac7-4b17-b709-d0e8ce2faf8b · inbound

The TIME Machine: On The Power of Motion for Efficient Perception cites this paper.

The TIME Machine: On The Power of Motion for Efficient Perception VLM4D: Towards Spatiotemporal Awareness in Vision Language Models

Reference 49

Resolution
unresolved
no resolver link, observed 2026-07-15T11:06:23.089564Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-15T11:06:23.089564Z digest=sha256:6bf6815aceb59c946ef6eb9dfdbc1357891bc98e8cac4b73ea53f4a974161523

Observation 493af086-78f3-47ac-ba1f-8cd2174e622e · inbound

The TIME Machine: On The Power of Motion for Efficient Perception cites this paper.

The TIME Machine: On The Power of Motion for Efficient Perception VLM4D: Towards Spatiotemporal Awareness in Vision Language Models

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-02T13:27:52.231964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T13:27:52.231964Z digest=sha256:a78d60387f969908cb045aa5f25c44cde4e2c822e0d7d301f8f4e839ddbf16c1