Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-11T00:56:06.243890Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 58 of 58 outbound references and 1 inbound Pith citation observation for arXiv:2605.07148.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-11T00:56:06.243890Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-07-01T06:23:00.372251Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-07-01T09:35:41.079576Z
58 of 58 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 08c3477f-89d2-4975-8896-0434b2e46ef6 · outbound
Uncovering and Shaping the Latent Representation of 3D Scene Topology in Vision-Language Models Cognitive maps in rats and men.Psychological review, 55(4):189
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0f1b2dc9-aa01-48ea-92ff-d63faf01c576 · outbound
Uncovering and Shaping the Latent Representation of 3D Scene Topology in Vision-Language Models Oxford university press
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f6200967-202b-452d-8c8f-86f4c8d91f81 · outbound
Uncovering and Shaping the Latent Representation of 3D Scene Topology in Vision-Language Models Non-euclidean navigation.Journal of Experimental Biology, 222(Suppl_1):jeb187971
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation cd19fa2b-720d-4c83-96cd-69a6048d47bc · outbound
Uncovering and Shaping the Latent Representation of 3D Scene Topology in Vision-Language Models Structuring knowledge with cognitive maps and cognitive graphs.Trends in cognitive sciences, 25(1):37–54
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 27281b9e-500c-480a-bf0e-caf4a31cd9d7 · outbound
Uncovering and Shaping the Latent Representation of 3D Scene Topology in Vision-Language Models LLaVA-OneVision: Easy Visual Task Transfer
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5155a600-edc0-47e2-8798-bfc8cfff8a9d · outbound
Uncovering and Shaping the Latent Representation of 3D Scene Topology in Vision-Language Models Video-chatgpt: Towards detailed video understanding via large vision and language models
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ee7d748a-dd6c-4c90-a4d3-f1fe73c1ad7f · outbound
Uncovering and Shaping the Latent Representation of 3D Scene Topology in Vision-Language Models GSR-BENCH: A Benchmark for Grounded Spatial Reasoning Evaluation via Multimodal LLMs
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ef9c6ca6-02f4-4ae1-ab5f-df39ffd73df5 · outbound
Uncovering and Shaping the Latent Representation of 3D Scene Topology in Vision-Language Models SQA3D: Situated Question Answering in 3D Scenes
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8e743d46-bcb1-403b-86ee-2961d9043cd8 · outbound
Uncovering and Shaping the Latent Representation of 3D Scene Topology in Vision-Language Models Reasoning paths with reference objects elicit quantitative spatial reasoning in large vision-language models
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 927ebeab-a5bc-4728-99de-eb2cbe62b1dc · outbound
Uncovering and Shaping the Latent Representation of 3D Scene Topology in Vision-Language Models Spatial reasoning in multimodal large language models: A survey of tasks, benchmarks and methods
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6b8a87e7-0520-4788-87c1-62fa8f53e181 · outbound
Uncovering and Shaping the Latent Representation of 3D Scene Topology in Vision-Language Models Infinibench: Infinite benchmarking for visual spatial reasoning with customizable scene complexity
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8e86367b-a397-4a40-bd3a-17dffc49096b · outbound
Uncovering and Shaping the Latent Representation of 3D Scene Topology in Vision-Language Models Generic attention-model explainability for interpreting bi-modal and encoder-decoder transformers
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 85e2888c-989e-427a-b271-75f040e6cd36 · outbound
Uncovering and Shaping the Latent Representation of 3D Scene Topology in Vision-Language Models Why is spatial reasoning hard for vlms? an attention mechanism perspective on focus areas
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 614fd597-f426-4038-920b-02175952746f · outbound
Uncovering and Shaping the Latent Representation of 3D Scene Topology in Vision-Language Models Beyond semantics: Rediscovering spatial awareness in vision-language models
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c2cc8c22-7df3-492c-83e6-ca6d44185f43 · outbound
Uncovering and Shaping the Latent Representation of 3D Scene Topology in Vision-Language Models The Geometry of Categorical and Hierarchical Concepts in Large Language Models
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c6a61006-fe26-4a2a-82bb-a79f507e37c3 · outbound
Uncovering and Shaping the Latent Representation of 3D Scene Topology in Vision-Language Models Linear mechanisms for spa- tiotemporal reasoning in vision language models.arXiv preprint arXiv:2601.12626
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5d1eb770-c5a5-4726-b3d8-4ce8a463a431 · outbound
Uncovering and Shaping the Latent Representation of 3D Scene Topology in Vision-Language Models Visual symbolic mechanisms: Emergent sym- bol processing in vision language models
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 49f331ee-9b96-4d8e-a761-040cd6933770 · outbound
Uncovering and Shaping the Latent Representation of 3D Scene Topology in Vision-Language Models Analyzing the behavior of visual question answering models
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9bae1465-c803-42dc-adac-5b739f011eab · outbound
Uncovering and Shaping the Latent Representation of 3D Scene Topology in Vision-Language Models Making the v in vqa matter: Elevating the role of image understanding in visual question answering
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 037a3f59-e3e9-40d4-9a48-0538a26a2384 · outbound
Uncovering and Shaping the Latent Representation of 3D Scene Topology in Vision-Language Models Vision Transformers Need More Than Registers
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5b3b7ce2-3770-4241-add7-677b6fecd296 · outbound
Uncovering and Shaping the Latent Representation of 3D Scene Topology in Vision-Language Models Thinking in space: How multimodal large language models see, remember, and recall spaces
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a3c8bfe4-cf8e-4839-a798-c2f68178f098 · outbound
Uncovering and Shaping the Latent Representation of 3D Scene Topology in Vision-Language Models Mindcube: Spatial mental modeling from limited views.arXiv e-prints, pages arXiv–2506
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 87938ae6-6ca2-4747-b61f-b550c280af50 · outbound
Uncovering and Shaping the Latent Representation of 3D Scene Topology in Vision-Language Models arXiv preprint arXiv:2602.07082 (2026),https://arxiv.org/ abs/2602.070824
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 181f7c00-1c9a-4297-9606-1a175e9b9ef2 · outbound
Uncovering and Shaping the Latent Representation of 3D Scene Topology in Vision-Language Models Spatialrgpt: Grounded spatial reasoning in vision-language models
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 338621f3-c352-40ec-801e-fa934f5f5090 · outbound
Uncovering and Shaping the Latent Representation of 3D Scene Topology in Vision-Language Models Scaling spatial intelligence with multimodal foundation models
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ca4585f3-9bbd-455f-99f5-b88ee5681fbf · outbound
Uncovering and Shaping the Latent Representation of 3D Scene Topology in Vision-Language Models SpatialThinker: Reinforcing Scene Graph-Grounded Spatial Reasoning via Dense Rewards
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2d33efec-352e-454e-86c7-bb00797f260e · outbound
Uncovering and Shaping the Latent Representation of 3D Scene Topology in Vision-Language Models Reinforcing Spatial Reasoning in Vision-Language Models with Interwoven Thinking and Visual Drawing
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 37da16b4-2c4e-4f13-9efc-754e4fa9f5c7 · outbound
Uncovering and Shaping the Latent Representation of 3D Scene Topology in Vision-Language Models Causal abstractions of neural networks.Advances in neural information processing systems, 34:9574–9586
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b410ba3a-7c41-4aed-b402-320a88e5580e · outbound
Uncovering and Shaping the Latent Representation of 3D Scene Topology in Vision-Language Models Investigating gender bias in language models using causal mediation analysis.Advances in neural information processing systems, 33:12388–12401
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b743b19c-20f2-4c89-b492-e9dba5bb284c · outbound
Uncovering and Shaping the Latent Representation of 3D Scene Topology in Vision-Language Models Emergent linear representations in world models of self-supervised sequence models
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation cfd72f98-76e0-4803-b261-18f66a79e08b · outbound
Uncovering and Shaping the Latent Representation of 3D Scene Topology in Vision-Language Models Decomposing Representation Space into Interpretable Subspaces with Unsupervised Learning
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0c5c4716-1a45-4ccb-b2ce-10b46640e1ac · outbound
Uncovering and Shaping the Latent Representation of 3D Scene Topology in Vision-Language Models Does object binding naturally emerge in large pretrained vi- sion transformers?
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation cbe57a86-3ae8-4a3f-a2d5-0c3ca7acd9a5 · outbound
Uncovering and Shaping the Latent Representation of 3D Scene Topology in Vision-Language Models SpatialBench: Benchmarking Multimodal Large Language Models for Spatial Cognition
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 058cd24e-667c-4fd4-a396-ed5a7c8120bb · outbound
Uncovering and Shaping the Latent Representation of 3D Scene Topology in Vision-Language Models Spatialvid: A large-scale video dataset with spatial annotations.arXiv preprint arXiv:2509.09676, 2025a
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4ef4ad3a-7120-4fda-a182-e177201cabc0 · outbound
Uncovering and Shaping the Latent Representation of 3D Scene Topology in Vision-Language Models Qwen2.5-vl, January 2025
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0bbd56e6-483b-4a62-8228-81bf73abc4eb · outbound
Uncovering and Shaping the Latent Representation of 3D Scene Topology in Vision-Language Models InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6493fa68-41b8-4a0b-a46c-d4e7f916c00e · outbound
Uncovering and Shaping the Latent Representation of 3D Scene Topology in Vision-Language Models Linear Representations of Sentiment in Large Language Models
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7813ad13-0ad4-425f-bb5b-dd6db8a09fc0 · outbound
Uncovering and Shaping the Latent Representation of 3D Scene Topology in Vision-Language Models Linear spaces of meanings: compositional structures in vision-language models
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c4936031-72cf-45e6-bedf-49071b82a667 · outbound
Uncovering and Shaping the Latent Representation of 3D Scene Topology in Vision-Language Models Deciphering personalization: Towards fine-grained explainability in natural language for personalized image generation models
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 910488ed-2c93-4e9a-8e62-f61b379587e9 · outbound
Uncovering and Shaping the Latent Representation of 3D Scene Topology in Vision-Language Models The Linear Representation Hypothesis and the Geometry of Large Language Models
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9fe6844e-8149-443b-8446-9f62bf4d1938 · outbound
Uncovering and Shaping the Latent Representation of 3D Scene Topology in Vision-Language Models On the Origins of Linear Representations in Large Language Models
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation cad0295a-cb29-49e2-b2e8-09af52fc21fc · outbound
Uncovering and Shaping the Latent Representation of 3D Scene Topology in Vision-Language Models Line of Sight: On Linear Representations in VLLMs
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 203ca3fe-6312-480a-8f52-b33d225eff26 · outbound
Uncovering and Shaping the Latent Representation of 3D Scene Topology in Vision-Language Models Interpreting clip with sparse linear concept embeddings (splice).Advances in Neural Information Processing Systems, 37:84298–84328
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0d0fdd4e-6d1a-4a93-9455-6b94734e17c1 · outbound
Uncovering and Shaping the Latent Representation of 3D Scene Topology in Vision-Language Models Linear Spatial World Models Emerge in Large Language Models
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation fc1baed4-e6aa-4312-a630-77e07bc6c1e4 · outbound
Uncovering and Shaping the Latent Representation of 3D Scene Topology in Vision-Language Models Laplacian eigenmaps for dimensionality reduction and data representation.Neural computation, 15(6):1373–1396
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 689aa3bb-4778-4f13-9d7f-a1543783fed9 · outbound
Uncovering and Shaping the Latent Representation of 3D Scene Topology in Vision-Language Models ICLR: In-Context Learning of Representations
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation cabe0242-351e-46ee-8f59-8b5a6bfc806d · outbound
Uncovering and Shaping the Latent Representation of 3D Scene Topology in Vision-Language Models Transformers as Unrolled Inference in Probabilistic Laplacian Eigenmaps: An Interpretation and Potential Improvements
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5ec14619-68aa-49f3-8e0b-34c367b7f2d9 · outbound
Uncovering and Shaping the Latent Representation of 3D Scene Topology in Vision-Language Models Learning Continually by Spectral Regularization
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6d13ab07-fc8f-4fbc-8322-3c90687f59e2 · outbound
Uncovering and Shaping the Latent Representation of 3D Scene Topology in Vision-Language Models Principal spectral regularization makes momentum surpass adam for llm training
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 337f0e07-c135-47fd-a09f-8002e729bb5c · outbound
Uncovering and Shaping the Latent Representation of 3D Scene Topology in Vision-Language Models Representation Learning on Graphs: Methods and Applications
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 94272ea8-8158-4e44-879e-c6267b7de011 · outbound
Uncovering and Shaping the Latent Representation of 3D Scene Topology in Vision-Language Models Semi-Supervised Classification with Graph Convolutional Networks
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2d70da38-6d48-4dc0-9a45-ffe457edcad4 · outbound
Uncovering and Shaping the Latent Representation of 3D Scene Topology in Vision-Language Models cognitive-map
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5c985aba-b092-4bf8-8ed1-a8b7c7065b12 · outbound
Uncovering and Shaping the Latent Representation of 3D Scene Topology in Vision-Language Models Unresolved cited work
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5b09b197-e4c4-48ec-a358-ca4d05ba1cbe · outbound
Uncovering and Shaping the Latent Representation of 3D Scene Topology in Vision-Language Models Unresolved cited work
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 65e336db-7681-45b7-82c0-613d46b187b9 · outbound
Uncovering and Shaping the Latent Representation of 3D Scene Topology in Vision-Language Models Each row is rescaled to unit norm
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c8697e43-b4f4-4bc0-8f19-1de257aafae3 · outbound
Uncovering and Shaping the Latent Representation of 3D Scene Topology in Vision-Language Models Drop columns with diagonal |Rkk| below 10−6 of the maximum (rank-deficient class directions)
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 512bef4f-93ca-4709-b63c-60a3c64a58b3 · outbound
Uncovering and Shaping the Latent Representation of 3D Scene Topology in Vision-Language Models Algorithm summary.The end-to-end per-step computation is: 1.Forward pass, capturingH ℓ ∈R B×T×d at each Dirichlet-target layer via forward hooks
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 512158b3-f242-451b-a783-c838eca33480 · outbound
Uncovering and Shaping the Latent Representation of 3D Scene Topology in Vision-Language Models For the residMulti variant, steps 1–8 run independently at 5 layers and step 9 averages the resulting ratios
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b1f00412-5aeb-4dde-92e4-cf5cb283b46c · inbound
Decodable Is Not Grounded: A Vision-Ablation Arbiter for VLM Spatial Reasoning Uncovering and Shaping the Latent Representation of 3D Scene Topology in Vision-Language Models
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.