Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T15:02:27.579774Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 29 of 29 outbound references and 1 inbound Pith citation observation for arXiv:2505.16579.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T15:02:27.579774Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-05-09T21:32:08.503584Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-11T14:36:05.322735Z
29 of 29 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation d392d12d-df48-447d-b2c8-55afe1872500 · outbound
Bridging the Dynamic Perception Gap: Training-Free Draft Chain-of-Thought for Dynamic Multimodal Spatial Reasoning Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d369d101-b043-43a5-ac8e-35b385e3d5bc · outbound
Bridging the Dynamic Perception Gap: Training-Free Draft Chain-of-Thought for Dynamic Multimodal Spatial Reasoning Qwen2.5-VL Technical Report
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a768fc52-f4a4-45e4-8ead-41ade8a9bbb9 · outbound
Bridging the Dynamic Perception Gap: Training-Free Draft Chain-of-Thought for Dynamic Multimodal Spatial Reasoning M$^3$CoT: A Novel Benchmark for Multi-Domain Multi-step Multi-modal Chain-of-Thought
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 724ccaf1-5152-4020-b225-fbaa6e536ca7 · outbound
Bridging the Dynamic Perception Gap: Training-Free Draft Chain-of-Thought for Dynamic Multimodal Spatial Reasoning Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4e7f2367-3101-4041-9abb-96fe64d90401 · outbound
Bridging the Dynamic Perception Gap: Training-Free Draft Chain-of-Thought for Dynamic Multimodal Spatial Reasoning Interleaved-Modal Chain-of-Thought
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9f366451-7cae-4ceb-860e-2eb4ee718d4e · outbound
Bridging the Dynamic Perception Gap: Training-Free Draft Chain-of-Thought for Dynamic Multimodal Spatial Reasoning The Abduction of Sherlock Holmes: A Dataset for Visual Abductive Reasoning
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e7c8bc35-93ae-49e7-8dd3-60f36a5cfcb2 · outbound
Bridging the Dynamic Perception Gap: Training-Free Draft Chain-of-Thought for Dynamic Multimodal Spatial Reasoning MME-CoT: Benchmarking Chain-of-Thought in Large Multimodal Models for Reasoning Quality, Robustness, and Efficiency
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 904ec99a-5f78-4d84-8995-bd187d72ae45 · outbound
Bridging the Dynamic Perception Gap: Training-Free Draft Chain-of-Thought for Dynamic Multimodal Spatial Reasoning LLaVA-OneVision: Easy Visual Task Transfer
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2d134f9e-5c17-4b86-9d2c-5ff156aa1530 · outbound
Bridging the Dynamic Perception Gap: Training-Free Draft Chain-of-Thought for Dynamic Multimodal Spatial Reasoning Imagine while Reasoning in Space: Multimodal Visualization-of-Thought
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 760a5f88-7882-4a49-ba54-e8a8f01cc2d8 · outbound
Bridging the Dynamic Perception Gap: Training-Free Draft Chain-of-Thought for Dynamic Multimodal Spatial Reasoning Unresolved cited work
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c39c6697-9f26-4516-900b-2db79dd05e48 · outbound
Bridging the Dynamic Perception Gap: Training-Free Draft Chain-of-Thought for Dynamic Multimodal Spatial Reasoning Med-PMC: Medical Personalized Multi-modal Consultation with a Proactive Ask-First-Observe-Next Paradigm
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 65a724fb-6abd-4279-8576-c22f418abf0d · outbound
Bridging the Dynamic Perception Gap: Training-Free Draft Chain-of-Thought for Dynamic Multimodal Spatial Reasoning NVILA: Efficient Frontier Visual Language Models
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 20fc7cfd-c3b1-4448-9b4a-6f89d7d9f9c1 · outbound
Bridging the Dynamic Perception Gap: Training-Free Draft Chain-of-Thought for Dynamic Multimodal Spatial Reasoning MathVista: Evaluating Mathematical Reasoning of Foundation Models in Visual Contexts
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 122c7166-4311-4d08-a4a3-f36094e61cd0 · outbound
Bridging the Dynamic Perception Gap: Training-Free Draft Chain-of-Thought for Dynamic Multimodal Spatial Reasoning Compositional Chain-of-Thought Prompting for Large Multimodal Models
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2f609678-7fec-42bd-9efb-9440d1536144 · outbound
Bridging the Dynamic Perception Gap: Training-Free Draft Chain-of-Thought for Dynamic Multimodal Spatial Reasoning Unresolved cited work
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 5c8956ee-8a8d-4793-8f62-05ccc2f48825 · outbound
Bridging the Dynamic Perception Gap: Training-Free Draft Chain-of-Thought for Dynamic Multimodal Spatial Reasoning DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3a676439-2414-4e49-95cd-ea8f163b8edd · outbound
Bridging the Dynamic Perception Gap: Training-Free Draft Chain-of-Thought for Dynamic Multimodal Spatial Reasoning Unresolved cited work
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation b695ae8d-aa12-48c9-a326-f4abd40500b4 · outbound
Bridging the Dynamic Perception Gap: Training-Free Draft Chain-of-Thought for Dynamic Multimodal Spatial Reasoning LLaVA-CoT: Let Vision Language Models Reason Step-by-Step
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ce0a4132-05fa-4f6b-81e4-15ea8a5f2294 · outbound
Bridging the Dynamic Perception Gap: Training-Free Draft Chain-of-Thought for Dynamic Multimodal Spatial Reasoning Qwen2.5-Omni Technical Report
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 259f02f7-c6b6-437b-b6b3-bbfaca65695f · outbound
Bridging the Dynamic Perception Gap: Training-Free Draft Chain-of-Thought for Dynamic Multimodal Spatial Reasoning The Dawn of LMMs: Preliminary Explorations with GPT-4V(ision)
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 18c2e927-af90-4a3f-8cb7-9d6565d0d76e · outbound
Bridging the Dynamic Perception Gap: Training-Free Draft Chain-of-Thought for Dynamic Multimodal Spatial Reasoning MM-REACT: Prompting ChatGPT for Multimodal Reasoning and Action
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 35bdbbd9-123a-4424-8122-d3144942bcae · outbound
Bridging the Dynamic Perception Gap: Training-Free Draft Chain-of-Thought for Dynamic Multimodal Spatial Reasoning Tree of Thoughts: Deliberate Problem Solving with Large Language Models
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 565bbe37-6430-40e2-af36-f08aad36afa9 · outbound
Bridging the Dynamic Perception Gap: Training-Free Draft Chain-of-Thought for Dynamic Multimodal Spatial Reasoning MiniCPM-V: A GPT-4V Level MLLM on Your Phone
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 78815e43-b6a4-4974-be74-be751308dd65 · outbound
Bridging the Dynamic Perception Gap: Training-Free Draft Chain-of-Thought for Dynamic Multimodal Spatial Reasoning Unresolved cited work
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 215337b5-4fd6-4475-95f7-3aba6a64c905 · outbound
Bridging the Dynamic Perception Gap: Training-Free Draft Chain-of-Thought for Dynamic Multimodal Spatial Reasoning VideoLLaMA 3: Frontier Multimodal Foundation Models for Image and Video Understanding
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4887c9c4-572d-4dee-8cd3-c88ea7972825 · outbound
Bridging the Dynamic Perception Gap: Training-Free Draft Chain-of-Thought for Dynamic Multimodal Spatial Reasoning DDCoT: Duty-Distinct Chain-of-Thought Prompting for Multimodal Reasoning in Language Models
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 75822fd1-1674-4244-a18f-b958e6bb4628 · outbound
Bridging the Dynamic Perception Gap: Training-Free Draft Chain-of-Thought for Dynamic Multimodal Spatial Reasoning InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f22aed27-8280-444e-becf-9813a3e5a6f0 · outbound
Bridging the Dynamic Perception Gap: Training-Free Draft Chain-of-Thought for Dynamic Multimodal Spatial Reasoning online" 'onlinestring :=
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 548b94c8-3f8f-4a2e-8a38-ceb2116087ed · outbound
Bridging the Dynamic Perception Gap: Training-Free Draft Chain-of-Thought for Dynamic Multimodal Spatial Reasoning write newline
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a3be3925-d3a2-4896-8e84-2d466f961232 · inbound
SketchVLM: Vision language models can annotate images to explain thoughts and guide users Bridging the Dynamic Perception Gap: Training-Free Draft Chain-of-Thought for Dynamic Multimodal Spatial Reasoning
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.