Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 22 inbound Pith citation observations for arXiv:2501.03230.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T11:18:57.344627Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-03T08:57:47.799001Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 2347e678-cf70-4e00-b626-809cef25365c · inbound
Uncertainty-o: One Model-agnostic Framework for Unveiling Uncertainty in Large Multimodal Models Video-of-Thought: Step-by-Step Video Reasoning from Perception to Cognition
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9675d94b-f65d-4fbb-a217-82b0059f9647 · inbound
VRBench: A Benchmark for Multi-Step Reasoning in Long Narrative Videos Video-of-Thought: Step-by-Step Video Reasoning from Perception to Cognition
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 90ecb6a8-ce9f-4510-9908-270fc34ff20f · inbound
VFaith: Do Large Multimodal Models Really Reason on Seen Images Rather than Previous Memories? Video-of-Thought: Step-by-Step Video Reasoning from Perception to Cognition
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 480c83b2-d71e-49d9-a101-d9b45129049b · inbound
Argus Inspection: Do Multimodal Large Language Models Possess the Eye of Panoptes? Video-of-Thought: Step-by-Step Video Reasoning from Perception to Cognition
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2ce8fd4d-5cef-4121-8368-d09de60cdb0a · inbound
DAVID-XR1: Detecting AI-Generated Videos with Explainable Reasoning Video-of-Thought: Step-by-Step Video Reasoning from Perception to Cognition
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 01f3cec6-94bc-4467-b9a8-00f90c6b63fb · inbound
CausalStep: A Benchmark for Explicit Stepwise Causal Reasoning in Videos Video-of-Thought: Step-by-Step Video Reasoning from Perception to Cognition
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4359e34a-b13d-4732-aa9a-1f327d90bb74 · inbound
Empowering Multimodal LLMs with External Tools: A Comprehensive Survey Video-of-Thought: Step-by-Step Video Reasoning from Perception to Cognition
Reference 155
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f57066a7-b0b7-4587-9c3a-b9919b65d70d · inbound
ProPy: Building Interactive Prompt Pyramids upon CLIP for Partially Relevant Video Retrieval Video-of-Thought: Step-by-Step Video Reasoning from Perception to Cognition
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 06ccf859-2cfb-4de3-bcfe-fe36e36c6fe1 · inbound
AdsQA: Towards Advertisement Video Understanding Video-of-Thought: Step-by-Step Video Reasoning from Perception to Cognition
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0cab05a2-37db-4e22-9ab5-6e1e9e8df3d7 · inbound
LaRe: Latent Refocusing for Multimodal Reasoning Video-of-Thought: Step-by-Step Video Reasoning from Perception to Cognition
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ffab02f4-3f31-4207-9b4f-1da7cea574b9 · inbound
SCP: Spatial Causal Prediction in Video Video-of-Thought: Step-by-Step Video Reasoning from Perception to Cognition
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0fdad807-aa9f-4112-b2a9-f42f8e0db055 · inbound
EndoCoT: Scaling Endogenous Chain-of-Thought Reasoning in Diffusion Models Video-of-Thought: Step-by-Step Video Reasoning from Perception to Cognition
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0d6f5247-243e-4e29-b227-e69e78685e10 · inbound
Reasoning Resides in Layers: Restoring Temporal Reasoning in Video-Language Models with Layer-Selective Merging Video-of-Thought: Step-by-Step Video Reasoning from Perception to Cognition
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6d566b1f-c33e-478e-be43-1413a3faa1dd · inbound
Chain-of-Glimpse: Search-Guided Progressive Object-Grounded Reasoning for Video Understanding Video-of-Thought: Step-by-Step Video Reasoning from Perception to Cognition
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e74bd42d-a812-4456-a4f0-cd6f30bd7042 · inbound
Chain-of-Glimpse: Search-Guided Progressive Object-Grounded Reasoning for Video Understanding Video-of-Thought: Step-by-Step Video Reasoning from Perception to Cognition
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ab8ee26b-d76d-4004-a8de-38e451d00b96 · inbound
When Text Hijacks Vision: Benchmarking and Mitigating Text Overlay-Induced Hallucination in Vision Language Models Video-of-Thought: Step-by-Step Video Reasoning from Perception to Cognition
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation aecfa82f-bb1c-4e17-b1ea-9835a388b5e2 · inbound
Act2See: Emergent Active Visual Perception for Video Reasoning Video-of-Thought: Step-by-Step Video Reasoning from Perception to Cognition
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 09ead882-9a5b-48be-b467-633f37849428 · inbound
ViSRA: A Video-based Spatial Reasoning Agent for Multi-modal Large Language Models Video-of-Thought: Step-by-Step Video Reasoning from Perception to Cognition
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation bbb98392-7b8e-4250-bac0-197f37dd3781 · inbound
AVTrack: Audio-Visual Tracking in Human-centric Complex Scenes Video-of-Thought: Step-by-Step Video Reasoning from Perception to Cognition
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation db96b895-cc57-46af-b777-73f535fa5fc5 · inbound
Watch, Remember, Reason: Human-View Video Understanding with MLLMs Video-of-Thought: Step-by-Step Video Reasoning from Perception to Cognition
Reference 178
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d3fa565a-c03c-4bc6-82f0-7c6b65c5f10a · inbound
Counterfactual Reasoning for Fine-Grained Evidence Disentanglement in VideoQA Video-of-Thought: Step-by-Step Video Reasoning from Perception to Cognition
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation cce4ce6e-70f5-48c3-bd5d-1e146ed7bfbf · inbound
Plan-and-Verify Video Reward Reasoning with Spatio-Temporal Scene Graph Grounding Video-of-Thought: Step-by-Step Video Reasoning from Perception to Cognition
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.