Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T21:12:12.530827Z
Paper Citation Record · LEDGER
As of 19 August 2026, this Paper Citation Record lists 41 of 41 outbound references and 0 inbound Pith citation observations for arXiv:2505.10453.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T21:12:12.530827Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
41 of 41 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 61e1e4e4-c887-4805-92bb-6492584a9f6c · outbound
Vision language models have difficulty recognizing virtual objects The imaginative mind.Human brain map- ping, 37(11):4197–4211, 2016
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation f91c4ed3-7143-4b53-91ee-ca32f1fbabd8 · outbound
Vision language models have difficulty recognizing virtual objects Mapping the imaginative mind: Charting new paths forward.Current Directions in Psychological Science, 30(1):82–89, 2021
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 9a92cd4f-23b8-4d0e-9bb5-bf5ca575c1d8 · outbound
Vision language models have difficulty recognizing virtual objects Humans predict liquid dynamics using probabilistic simulation
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 6f4f7625-2b7b-4975-b1bb-fcbc65845f1b · outbound
Vision language models have difficulty recognizing virtual objects Non- commitment in mental imagery.Cognition, 238:105498,
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 30cd9e77-9830-4a90-b16e-64e50895bf85 · outbound
Vision language models have difficulty recognizing virtual objects Infusing perception with imagination.Per- ceptual imagination and perceptual memory, pages 133– 160, 2018
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 9bd19b81-4e66-40ab-8a77-898be6cdd1ac · outbound
Vision language models have difficulty recognizing virtual objects SpatialBot: Precise Spatial Understanding with Vision Language Models
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d5cd8748-daf4-48bc-bf0c-0e753bab194b · outbound
Vision language models have difficulty recognizing virtual objects The artist as neuroscientist.Nature, 434 (7031):301–307, 2005
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 7f57a4cb-e8d9-48c4-9c6a-c05b41646c12 · outbound
Vision language models have difficulty recognizing virtual objects Spatialvlm: Endow- ing vision-language models with spatial reasoning capabili- ties
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b457c641-7738-4ea9-bcd7-2e409852a23c · outbound
Vision language models have difficulty recognizing virtual objects Large language models are visual reasoning coordinators.Ad- vances in Neural Information Processing Systems, 36, 2024
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 2fb3764b-8f06-4347-b9f1-de1549c1d379 · outbound
Vision language models have difficulty recognizing virtual objects SpatialRGPT: Grounded Spatial Reasoning in Vision Language Models
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9ad80f3b-ef34-4abd-a3f0-b9d4cf5ec100 · outbound
Vision language models have difficulty recognizing virtual objects What makes mental modeling difficult? normative data for the multidimensional relational reasoning task.Frontiers in psychology, 12:668256, 2021
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 9bfc39e8-cd9b-4a0e-83a1-9f0711354e6b · outbound
Vision language models have difficulty recognizing virtual objects A survey on multimodal large lan- guage models for autonomous driving
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0a817b15-cc02-4583-9785-ba8b4851fe32 · outbound
Vision language models have difficulty recognizing virtual objects Objaverse: A universe of annotated 3d objects
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e734e45a-98b1-4815-8ca0-3c556d5ec889 · outbound
Vision language models have difficulty recognizing virtual objects Psychology press, 2014
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 164277a5-3b5f-4f02-a53e-2c9829f9401c · outbound
Vision language models have difficulty recognizing virtual objects Scene-LLM: Extending Language Model for 3D Visual Understanding and Reasoning
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8711b5e9-2115-4fc6-8791-60ae69e25e4a · outbound
Vision language models have difficulty recognizing virtual objects Exploring the frontier of vision- language models: A survey of current methodologies and future directions.arXiv preprint arXiv:2404.07214, 2024
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4eb97da6-8ce5-497d-a1d8-6740840ccc4d · outbound
Vision language models have difficulty recognizing virtual objects Mental animation: Inferring motion from static displays of mechanical systems.Journal of experi- mental psychology: learning, memory, and cognition, 18(5): 1084, 1992
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation a5664f86-3b87-45e0-aafa-3b3a0e1a0ad8 · outbound
Vision language models have difficulty recognizing virtual objects Components of spatial intelligence
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation d0e1af8d-6469-49a4-8168-ac44de2f4fdf · outbound
Vision language models have difficulty recognizing virtual objects Correctness comparison of chatgpt-4, gemini, claude-3, and copilot for spatial tasks.Transactions in GIS, 2024
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation aec52deb-391d-4a65-9802-a6b75c2ac3fa · outbound
Vision language models have difficulty recognizing virtual objects Vcoder: Ver- satile vision encoders for multimodal large language models
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 10bb8773-193f-4eb9-86dd-819740fed691 · outbound
Vision language models have difficulty recognizing virtual objects Imagery, visualization, and think- ing.Perception and cognition at century’s end, pages 441– 467, 1998
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation fde784fd-c61c-48d3-9335-170a4de56184 · outbound
Vision language models have difficulty recognizing virtual objects Harrison, Wallace E
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation ac046d2e-a6a4-4f84-a67b-ab23c3895bf1 · outbound
Vision language models have difficulty recognizing virtual objects Kinematic mental sim- ulations in abduction and deduction.proceedings of the na- tional academy of sciences, 110(42):16766–16771, 2013
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 46fb6041-f54a-47e4-810c-97446058807d · outbound
Vision language models have difficulty recognizing virtual objects Mit Press, 2013
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 76e15394-aa56-4769-8216-912b6c7bd466 · outbound
Vision language models have difficulty recognizing virtual objects Visual imagery can impede reasoning.Memory & cognition, 30:363–371,
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 21114088-c913-4f85-ad9b-7ae1baa0ece0 · outbound
Vision language models have difficulty recognizing virtual objects tracking
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 141662ea-b9a3-4ea4-9dd4-2f1f2f6aff36 · outbound
Vision language models have difficulty recognizing virtual objects What matters when building vision-language models?
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation da5c2a36-8920-42df-8271-ae264f5165ff · outbound
Vision language models have difficulty recognizing virtual objects BLIP: Bootstrapping language-image pre-training for unified vision-language understanding and generation
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation b94335cc-8ad5-47a5-a88f-f0f2c730c53e · outbound
Vision language models have difficulty recognizing virtual objects A Survey of Attacks on Large Vision-Language Models: Resources, Advances, and Future Trends
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 807d42e8-5609-4e4e-8e72-8cb10b988d6e · outbound
Vision language models have difficulty recognizing virtual objects Visual spa- tial reasoning.Transactions of the Association for Computa- tional Linguistics, 11:635–651, 2023
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 82f5e51e-8551-45a3-874b-97ddd95a782e · outbound
Vision language models have difficulty recognizing virtual objects A Survey on Hallucination in Large Vision-Language Models
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 94fa8e14-06a5-4714-bcf7-f397c6c1dd3c · outbound
Vision language models have difficulty recognizing virtual objects 5 Enhancing visual reasoning with autonomous imagination in multimodal large language models.arXiv preprint arXiv:2411.18142, 2024
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5c04e663-9ad8-49bf-8f0d-1d495d31001a · outbound
Vision language models have difficulty recognizing virtual objects The human imagination: the cognitive neu- roscience of visual mental imagery.Nature reviews neuro- science, 20(10):624–634, 2019
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 310301e0-6938-4449-949b-61240fddea3d · outbound
Vision language models have difficulty recognizing virtual objects Mental rotation of three-dimensional objects.Science, 171(3972):701–703,
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f95d7ee2-e416-4371-8d95-f61529d9a29c · outbound
Vision language models have difficulty recognizing virtual objects Llama 2: Open Foundation and Fine-Tuned Chat Models
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 45b7ea2c-5398-4e21-a6d6-9dc468a9a7c5 · outbound
Vision language models have difficulty recognizing virtual objects Visuospatial reasoning
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation e4a4ded8-de26-4fe7-9487-b3961f580a70 · outbound
Vision language models have difficulty recognizing virtual objects Learning physical parameters from dynamic scenes.Cognitive psychology, 104:57–82,
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 75f919fd-5d57-4978-bd47-fd163b16edfa · outbound
Vision language models have difficulty recognizing virtual objects Multimodal large language models: A sur- vey
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 53f8aa8d-70b3-4747-b4f4-65376980ecc1 · outbound
Vision language models have difficulty recognizing virtual objects Unresolved cited work
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 79ff480f-2082-47aa-86dc-3b8e118755d7 · outbound
Vision language models have difficulty recognizing virtual objects Vision-language models for vision tasks: A survey.IEEE Transactions on Pattern Analysis and Machine Intelligence,
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 4531f783-7981-4220-9a1f-54e597539126 · outbound
Vision language models have difficulty recognizing virtual objects ImagineNav: Prompting Vision-Language Models as Embodied Navigator through Scene Imagination
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.