Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 34 inbound Pith citation observations for arXiv:2503.11495.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-06T23:13:11.605015Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T03:29:31.466663Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 70e69085-bb97-406a-9398-6610082473d3 · inbound
MUSEG: Reinforcing Video Temporal Understanding via Timestamp-Aware Multi-Segment Grounding V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e4c8650b-988e-4782-8feb-b1e7b9e6369e · inbound
Video-Holmes: Can MLLM Think Like Holmes for Complex Video Reasoning? V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0c4697dc-4d72-452b-95f8-863bae8de435 · inbound
Video-XL-2: Towards Very Long-Video Understanding Through Task-Aware KV Sparsification V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 14647125-7f8d-47c6-b29e-99a62ffd6386 · inbound
Position: Reasoning After Perception Means Reasoning Without Vision V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ff4940c9-173f-439d-af2e-64e36c8d9b52 · inbound
Video Reasoning without Training V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0dab8555-7807-4152-81bb-771f59931837 · inbound
SPHINX: A Synthetic Environment for Visual Perception and Reasoning V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 49fdc14f-3a78-4c35-8d12-ed88dc970b80 · inbound
ToG-Bench: Task-Oriented Spatio-Temporal Grounding in Egocentric Videos V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 05787ab9-47af-4202-a225-2e5012a2b0f3 · inbound
Detector-Empowered Video Large Language Model for Efficient Spatio-Temporal Grounding V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 34752852-56d0-4893-b0d6-2cc488e0785a · inbound
$M^3-Verse$: A "Spot the Difference" Challenge for Large Multimodal Models V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4434c415-058f-4bc4-b1d9-d3381d95eb76 · inbound
GraphThinker: Reinforcing Temporally Grounded Video Reasoning with Event Graph Thinking V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 01b27391-7998-452c-82c8-15a0639236b9 · inbound
Reasoning over Video: Evaluating How MLLMs Extract, Integrate, and Reconstruct Spatiotemporal Evidence V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 39ea0772-cfc9-4060-adbe-f6b28a0d1bf0 · inbound
LanteRn: Latent Visual Structured Reasoning V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d296c926-fb7b-4b53-bc23-7760c7688fd7 · inbound
Thinking Diffusion: Penalize and Guide Visual-Grounded Reasoning in Diffusion Multimodal Language Models V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5fe25213-2759-40e5-916b-3de23524f770 · inbound
PinpointQA: A Dataset and Benchmark for Small Object-Centric Spatial Understanding in Indoor Videos V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation fdd9b21a-10ae-4891-874d-3e9774f2f93b · inbound
PinpointQA: A Dataset and Benchmark for Small Object-Centric Spatial Understanding in Indoor Videos V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation bfed0c39-e49c-4fa0-8d6a-bc365b8e3c52 · inbound
Grounding Video Reasoning in Physical Signals V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4f64092f-2de4-4f72-848b-535b6c76bdb1 · inbound
Lost in the Vibrations: Vision Language Models Fail the Dynamic Gauges Test V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 580cd081-71b7-489a-8ca3-6df290398c05 · inbound
State Beyond Appearance: Diagnosing and Improving State Consistency in Dial-Based Measurement Reading V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 40c04b4a-def2-4a06-94e8-64e9fb4882ac · inbound
VISD: Enhancing Video Reasoning via Structured Self-Distillation V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3139deae-cdd6-464c-9dac-4bae4faeca9b · inbound
VISD: Enhancing Video Reasoning via Structured Self-Distillation V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e03cac00-ffab-4db3-8d13-f374ff4d5d1c · inbound
VISD: Enhancing Video Reasoning via Structured Self-Distillation V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4d0633f3-9ddf-4e90-90e3-aadf43a69681 · inbound
VISD: Enhancing Video Reasoning via Structured Self-Distillation V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 575950d7-67f5-4551-84ea-80792507ff47 · inbound
Tracking the Truth: Object-Centric Spatio-Temporal Monitoring for Video Large Language Models V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 99feebf8-132f-4a55-8ed9-5c6316bba8ce · inbound
STAR: Failure-Aware Markovian Routing for Multi-Agent Spatiotemporal Reasoning V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 60f3c841-2f41-4531-bd2f-6af2879bfb36 · inbound
STAR: Failure-Aware Markovian Routing for Multi-Agent Spatiotemporal Reasoning V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation bdfd50cc-c3a1-45ab-a806-613502762a66 · inbound
STAR: Failure-Aware Markovian Routing for Multi-Agent Spatiotemporal Reasoning V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 43114469-3aed-4f9b-93cc-492c154b77da · inbound
Leveraging Latent Visual Reasoning in Silence V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b161ec6d-aff5-4fb8-a9d6-aba3e7bfb6e1 · inbound
EgoCoT-Bench: Benchmarking Grounded and Verifiable Operation-Centric Chain of Thought Reasoning for MLLMs V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation acc17bc1-d23a-4fb0-bffd-c0c25bd9709d · inbound
Maestro: Reinforcement Learning to Orchestrate Hierarchical Model-Skill Ensembles V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4118160a-482e-4743-9a85-123aa026a9e0 · inbound
X-Stream: Exploring MLLMs as Multiplexers for Multi-Stream Understanding V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f42b710a-b94a-43a5-a40c-43ae22d8d8c4 · inbound
X-Stream: Exploring MLLMs as Multiplexers for Multi-Stream Understanding V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3d6a04cc-ad82-4abd-948a-7b359b204343 · inbound
CARE: Competence-Aware Reward Shaping for Adaptive Reasoning Length in Video-MLLMs V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8187d767-d0f0-47db-a4b8-d683cccb1b57 · inbound
Video-MME-Logical: A Controlled Diagnostic Benchmark for Video Temporal-Logical Reasoning V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 30754d7d-7d56-47de-8646-0b001dca03f1 · inbound
AnyGroundBench: A Specialized-Domain Benchmark for Video Grounding in Vision-Language Models V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.