Pith. sign in

Paper Citation Record · LEDGER

Mind's Eye of LLMs: Visualization-of-Thought Elicits Spatial Reasoning in Large Language Models

As of 17 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 10 inbound Pith citation observations for arXiv:2404.03622.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2404.03622 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 10 of 10 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 10 of 10 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T12:17:05.982323Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

2
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 0044c37b-b761-4c17-922a-f350461adaa7 · inbound

VideoCogQA: A Controllable Benchmark for Evaluating Cognitive Abilities in Video-Language Models cites this paper.

VideoCogQA: A Controllable Benchmark for Evaluating Cognitive Abilities in Video-Language Models Mind's Eye of LLMs: Visualization-of-Thought Elicits Spatial Reasoning in Large Language Models

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-12T21:07:09.643393Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:07:09.643393Z digest=sha256:39624ac4799f0b682b07b3a00619f37fe05dd608dfbec1a2085cf17754327eda

Observation 142c6a96-f632-490f-b4e5-d58b7d8eaf93 · inbound

CoMT: A Novel Benchmark for Chain of Multi-modal Thought on Large Vision-Language Models cites this paper.

CoMT: A Novel Benchmark for Chain of Multi-modal Thought on Large Vision-Language Models Mind's Eye of LLMs: Visualization-of-Thought Elicits Spatial Reasoning in Large Language Models

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-11T13:38:14.120001Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T13:38:14.120001Z digest=sha256:db141efd4742ee637d141cdca0f082e806ca9b2e2f0185509c3d61954b906b9e

Observation 9e875c7d-0722-46b8-aaf1-4a8b6bb602b9 · inbound

Do Multimodal Language Models Really Understand Direction? A Benchmark for Compass Direction Reasoning cites this paper.

Do Multimodal Language Models Really Understand Direction? A Benchmark for Compass Direction Reasoning Mind's Eye of LLMs: Visualization-of-Thought Elicits Spatial Reasoning in Large Language Models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-11T10:29:05.578679Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T10:29:05.578679Z digest=sha256:179fefbbf0e6f9edbcc491cbe049951d76c28f5212a479801f9d89ececc850d8

Observation a8fd1120-f322-4431-8841-f3f0918269f6 · inbound

Object-Driven Narrative in AR: A Scenario-Metaphor Framework with VLM Integration cites this paper.

Object-Driven Narrative in AR: A Scenario-Metaphor Framework with VLM Integration Mind's Eye of LLMs: Visualization-of-Thought Elicits Spatial Reasoning in Large Language Models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-16T12:17:05.982323Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:17:05.982323Z digest=sha256:b798a7778c72de579047045116fc385e909bfa73ec8c910213c4b3a97282abdb

Observation 4599051a-8521-447e-914d-b797dd73a1be · inbound

Grounded Reinforcement Learning for Visual Reasoning cites this paper.

Grounded Reinforcement Learning for Visual Reasoning Mind's Eye of LLMs: Visualization-of-Thought Elicits Spatial Reasoning in Large Language Models

Reference 72

Resolution
verified exact
arxiv_id, observed 2026-05-22T01:05:52.133035Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-22T01:05:18.801388Z digest=sha256:af6f79f700a92bcbaba283550a2f54552394990ee7fbc08ccc7337180af9387f

Observation 5666202e-5beb-4160-8ff1-4c59f3288bbb · inbound

Enhancing Spatial Reasoning in Vision-Language Models via Chain-of-Thought Prompting and Reinforcement Learning cites this paper.

Enhancing Spatial Reasoning in Vision-Language Models via Chain-of-Thought Prompting and Reinforcement Learning Mind's Eye of LLMs: Visualization-of-Thought Elicits Spatial Reasoning in Large Language Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T19:54:48.116134Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:54:48.116134Z digest=sha256:8e0265dac28f7df430b5e75722c03b001cd05b5199b6dea94061088d9d1805be

Observation 1f538ca5-df5c-4728-8f1e-c26e3ae4beff · inbound

Think in Games: Learning to Reason in Games via Reinforcement Learning with Large Language Models cites this paper.

Think in Games: Learning to Reason in Games via Reinforcement Learning with Large Language Models Mind's Eye of LLMs: Visualization-of-Thought Elicits Spatial Reasoning in Large Language Models

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-05T14:23:26.747722Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T14:23:26.747722Z digest=sha256:b561e1c5bfcd5098f2aa21367bd81e252d9b0e716edbf2ac16868d348e9fb885

Observation 50412294-2af5-4b14-981b-4061771e67aa · inbound

AtomWorld: A Benchmark for Evaluating Spatial Reasoning in Large Language Models on Crystalline Materials cites this paper.

AtomWorld: A Benchmark for Evaluating Spatial Reasoning in Large Language Models on Crystalline Materials Mind's Eye of LLMs: Visualization-of-Thought Elicits Spatial Reasoning in Large Language Models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-04T11:26:06.616827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T11:26:06.616827Z digest=sha256:fbf8b8e98f7ec5874f031edc01ff0f72c30be95f374232d43dcbca43dab6a9bb

Observation 282cc3c0-46aa-46a1-8545-13340e51ec49 · inbound

Imagine Before You Predict: Interleaved Latent Visual Reasoning for Video Event Prediction cites this paper.

Imagine Before You Predict: Interleaved Latent Visual Reasoning for Video Event Prediction Mind's Eye of LLMs: Visualization-of-Thought Elicits Spatial Reasoning in Large Language Models

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-07-02T11:56:55.434831Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-06-28T02:46:50.373450Z digest=sha256:d6f7d574e8b15fa30474d6dc8bbbf5d4247a59b7e2c6d152ee918438bc1d801c

Observation c60ed340-612f-47c5-ba7a-4fbb069ba5cc · inbound

Lost in Aggregation: A Multi-Scale Diagnostic Benchmark for LLM Spatial Navigation cites this paper.

Lost in Aggregation: A Multi-Scale Diagnostic Benchmark for LLM Spatial Navigation Mind's Eye of LLMs: Visualization-of-Thought Elicits Spatial Reasoning in Large Language Models

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-06-26T10:39:18.347904Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-06-26T10:37:24.946718Z digest=sha256:7915a3c2419772cc022614497afba7e2b170b504a0304fafcc62aade28731f1f