Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T05:51:07.923531Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 31 of 31 outbound references and 1 inbound Pith citation observation for arXiv:2506.06928.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T05:51:07.923531Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-02T05:40:00.723218Z
A source-named dated measurement, never combined with another source.
Source: cited_works
31 of 31 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 65c748a2-cd0b-461a-97a9-c9cd92477f4f · outbound
How Important are Videos for Training Video LLMs? Qwen2.5-VL Technical Report
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 285cf941-cb86-40de-8a0d-f2cc99af891e · outbound
How Important are Videos for Training Video LLMs? ShareGPT4Video: Improving Video Under- standing and Generation with Better Captions
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ea581482-7c3a-42a9-a25d-ee24df90f017 · outbound
How Important are Videos for Training Video LLMs? VideoLLaMA 2: Advancing Spatial-Temporal Modeling and Audio Understanding in Video-LLMs
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 03cff194-f5a0-4074-90c0-a2701db95f6b · outbound
How Important are Videos for Training Video LLMs? Lost in Time: A New Temporal Benchmark for VideoLLMs
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4f6aa8c0-6d3b-4c63-9c20-f096c821bc1d · outbound
How Important are Videos for Training Video LLMs? Video-MME: The First-Ever Comprehensive Evaluation Benchmark of Multi-modal LLMs in Video Analysis
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1c2a4e8b-65ba-4337-a695-abd2ff200285 · outbound
How Important are Videos for Training Video LLMs? The Llama 3 Herd of Models
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bf156721-0a82-4991-b65e-8aa7f00a3508 · outbound
How Important are Videos for Training Video LLMs? MA-LMM: Memory-Augmented Large Multimodal Model for Long-Term Video Understanding
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 063e0f5c-e995-4454-a493-83357a8972d4 · outbound
How Important are Videos for Training Video LLMs? LoRA: Low- Rank Adaptation of Large Language Models
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ed31497f-5f07-422d-869b-cc8a4a3fd827 · outbound
How Important are Videos for Training Video LLMs? TGIF-QA: Toward Spatio-Temporal Reason- ing in Visual Question Answering
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 09ad63b2-12d0-452e-b6c1-6bcae633622a · outbound
How Important are Videos for Training Video LLMs? CLEVR: A Diagnostic Dataset for Compositional Language and Elementary Visual Reasoning
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 053fb2fb-be20-41d5-ad0e-0dc963146e8b · outbound
How Important are Videos for Training Video LLMs? JUWELS Cluster and Booster: Exascale Pathfinder with Modular Supercomputing Architecture at Juelich Supercomputing Centre
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a8b5ff99-15ec-41f0-8177-5e10f582f342 · outbound
How Important are Videos for Training Video LLMs? LLaVA-OneVision: Easy Visual Task Transfer
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 39381d6c-50ee-468e-a4a9-55417ca9e93c · outbound
How Important are Videos for Training Video LLMs? MVBench: A Comprehensive Multi-modal Video Under- standing Benchmark
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ad71e2fe-5895-4c17-814c-8f8992cfa555 · outbound
How Important are Videos for Training Video LLMs? Temporal Preference Optimization for Long-Form Video Understanding
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6e1e2824-7908-4c1d-9206-0cf2f47da698 · outbound
How Important are Videos for Training Video LLMs? LLaMA-VID: An Image is Worth 2 Tokens in Large Language Models
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9f8edbfe-10aa-4d2c-ae75-c677098d699e · outbound
How Important are Videos for Training Video LLMs? Video-LLaV A: Learning United Visual Representation by Alignment Before Projection
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5b7ce0b0-5171-461d-a9b3-8f1314059ba4 · outbound
How Important are Videos for Training Video LLMs? Microsoft COCO: Common Objects in Context
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9ebd11f4-cb7d-4fc7-9870-6582a89258f8 · outbound
How Important are Videos for Training Video LLMs? Oryx MLLM: On-Demand Spatial-Temporal Understanding at Arbitrary Resolution
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bf28ca4d-4cfe-4ed5-8dd5-db851bd7a74d · outbound
How Important are Videos for Training Video LLMs? Video-ChatGPT: Towards Detailed Video Under- standing via Large Vision and Language Models
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d9d66231-4270-478c-bf57-9fa8221d2497 · outbound
How Important are Videos for Training Video LLMs? Unresolved cited work
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f86bf564-bf9b-4076-b424-f45354ce05a1 · outbound
How Important are Videos for Training Video LLMs? Learning Transferable Visual Models From Natural Language Super- vision
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6dce3466-df81-459c-a18d-87100f0d2f40 · outbound
How Important are Videos for Training Video LLMs? LongVU: Spatiotemporal Adaptive Compression for Long Video-Language Understanding
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0cb707a5-6917-4e9f-b00d-1c255f096643 · outbound
How Important are Videos for Training Video LLMs? Tarsier: Recipes for Training and Evaluating Large Video Description Models
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c3ee944a-08d6-479e-a3b3-22b187f7925e · outbound
How Important are Videos for Training Video LLMs? LongVideoBench: A Benchmark for Long-context Inter- leaved Video-Language Understanding
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4e0dc231-9de6-4f15-8f1e-0112554d8688 · outbound
How Important are Videos for Training Video LLMs? MSR-VTT: A Large Video Description Dataset for Bridging Video and Language
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 897dc92b-73ff-4532-9744-3ca2e42c1f23 · outbound
How Important are Videos for Training Video LLMs? PLLaVA : Parameter-free LLaVA Extension from Images to Videos for Video Dense Captioning
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 87ec161b-a15b-4e8c-afd8-4be53ef31414 · outbound
How Important are Videos for Training Video LLMs? CLEVRER: CoLlision Events for Video REpresentation and Reasoning
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ff108d37-dfb5-4f80-8fa6-b5a81e09e326 · outbound
How Important are Videos for Training Video LLMs? ActivityNet-QA: A Dataset for Understanding Complex Web Videos via Question Answer- ing
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c96d5b97-b387-450b-a5f1-13dc333d2a9f · outbound
How Important are Videos for Training Video LLMs? Tarsier2: Advancing Large Vision-Language Models from Detailed Video Description to Comprehensive Video Understanding
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5faf4ef4-d8e6-4340-b5db-458784a5186f · outbound
How Important are Videos for Training Video LLMs? Video-LLaMA: An Instruction-tuned Audio-Visual Language Model for Video Understanding
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b6f15c9d-18d7-4e73-9bcc-4a2501bf2bc6 · outbound
How Important are Videos for Training Video LLMs? LLaVA-Video: Video Instruction Tuning With Synthetic Data
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8c7647d5-6f3e-4110-8c64-34e90ee76aa9 · inbound
Accuracy Without Grounding: Diagnosing Visual Dependency Dissociation in Video LLM Benchmarks How Important are Videos for Training Video LLMs?
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.