Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 18 inbound Pith citation observations for arXiv:2406.11303.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T15:34:43.788047Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-02T12:26:57.184700Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation df47b175-ed30-402b-9ad1-4104542abff3 · inbound
MotionBench: Benchmarking and Improving Fine-grained Video Motion Understanding for Vision Language Models VideoVista: A Versatile Benchmark for Video Understanding and Reasoning
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 29aaa453-09ca-4625-84b3-d9457014c874 · inbound
VideoEval-Pro: Robust and Realistic Long Video Understanding Evaluation VideoVista: A Versatile Benchmark for Video Understanding and Reasoning
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fc66876d-c16b-4d4b-977f-7b397612597f · inbound
TUNA: Comprehensive Fine-grained Temporal Understanding Evaluation on Dense Dynamic Videos VideoVista: A Versatile Benchmark for Video Understanding and Reasoning
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7514bf63-6fd6-4d1f-91bb-5fac0532564c · inbound
VCapsBench: A Large-scale Fine-grained Benchmark for Video Caption Quality Evaluation VideoVista: A Versatile Benchmark for Video Understanding and Reasoning
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f4e62474-1e58-4261-b56c-11dbe3e6481f · inbound
VF-Eval: Evaluating Multimodal LLMs for Generating Feedback on AIGC Videos VideoVista: A Versatile Benchmark for Video Understanding and Reasoning
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6f5a593c-f7dc-4764-85f5-da4346e18514 · inbound
ScaleLong: A Multi-Timescale Benchmark for Long Video Understanding VideoVista: A Versatile Benchmark for Video Understanding and Reasoning
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 408e7f61-f69b-46cd-97ec-da0311486d57 · inbound
VUDG: A Dataset for Video Understanding Domain Generalization VideoVista: A Versatile Benchmark for Video Understanding and Reasoning
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f72b07e1-89ec-4e63-a23d-54c6a082e19c · inbound
SIV-Bench: A Video Benchmark for Social Interaction Understanding and Reasoning VideoVista: A Versatile Benchmark for Video Understanding and Reasoning
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 72944cf2-74c2-4dd2-91fa-073fcf1672fa · inbound
VRBench: A Benchmark for Multi-Step Reasoning in Long Narrative Videos VideoVista: A Versatile Benchmark for Video Understanding and Reasoning
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fd3247c8-5278-4a34-86de-a9e8723a234c · inbound
SmartHome-Bench: A Comprehensive Benchmark for Video Anomaly Detection in Smart Homes Using Multi-Modal Large Language Models VideoVista: A Versatile Benchmark for Video Understanding and Reasoning
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7d166b95-e89a-4822-9a73-8309863d4e2a · inbound
VideoForest: Person-Anchored Hierarchical Reasoning for Cross-Video Question Answering VideoVista: A Versatile Benchmark for Video Understanding and Reasoning
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4880c9fe-dc2e-4103-adad-2d91bd29d674 · inbound
NeMo: Needle in a Montage for Video-Language Understanding VideoVista: A Versatile Benchmark for Video Understanding and Reasoning
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4f18c017-fa11-49b1-9ff5-5a40e6d0cc64 · inbound
Seeing the Scene Matters: Revealing Forgetting in Video Understanding Models with a Scene-Aware Long-Video Benchmark VideoVista: A Versatile Benchmark for Video Understanding and Reasoning
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c02b375f-471e-495f-ba2d-dd8d1b12bdb0 · inbound
Minerva-Ego: Spatiotemporal Hints for Egocentric Video Understanding VideoVista: A Versatile Benchmark for Video Understanding and Reasoning
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b7b3ffe2-8116-4a87-b31b-d0d2b8287e22 · inbound
VidMsg: A Benchmark for Implicit Message Inference in Short Videos VideoVista: A Versatile Benchmark for Video Understanding and Reasoning
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation fa3a6b1c-d6de-40ef-bfcd-2f77df218b05 · inbound
StoryVideoQA: Scaling Deep Video Understanding with a Large-Scale, Multi-Genre and Auto-Generated Dataset VideoVista: A Versatile Benchmark for Video Understanding and Reasoning
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7ad99eda-7280-48fb-a4c6-8582ee1d7c21 · inbound
Animation2Code: Evaluating Temporal Visual Reasoning in Video-to-Code Generation VideoVista: A Versatile Benchmark for Video Understanding and Reasoning
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e059e240-3680-4a48-8baa-dbd830e98b72 · inbound
The Low Frequency Trap: Video Language Models Fail at Simple Event Bookkeeping VideoVista: A Versatile Benchmark for Video Understanding and Reasoning
Reference 73
Source-reported events for the cited work
Unavailable: canonical work link unavailable.