Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:27:45.034743Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 47 of 47 outbound references and 1 inbound Pith citation observation for arXiv:2505.18855.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:27:45.034743Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-05-18T18:35:01.328250Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-18T18:36:44.198389Z
47 of 47 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 07b74f4f-c048-42b1-beeb-d744185889e8 · outbound
Inference Compute-Optimal Video Vision Language Models online" 'onlinestring :=
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0445e824-2204-4648-aa84-a175c14c103d · outbound
Inference Compute-Optimal Video Vision Language Models write newline
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c2dbc6aa-012a-43e1-bbfe-232752a0c256 · outbound
Inference Compute-Optimal Video Vision Language Models Unresolved cited work
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 8ec7195f-3fe4-4ca6-9a2a-4f3095b87221 · outbound
Inference Compute-Optimal Video Vision Language Models Unresolved cited work
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 62fbf554-3b05-4a34-9ad5-b4f368a48020 · outbound
Inference Compute-Optimal Video Vision Language Models Unresolved cited work
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fbb83004-b817-40b0-b693-f5540c72c7ca · outbound
Inference Compute-Optimal Video Vision Language Models Large Language Monkeys: Scaling Inference Compute with Repeated Sampling
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1e103d7d-ea50-4390-81fe-b30ffb34b2aa · outbound
Inference Compute-Optimal Video Vision Language Models Unresolved cited work
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c03d8fc7-7216-4192-a70f-aa2dbca8e89b · outbound
Inference Compute-Optimal Video Vision Language Models Unresolved cited work
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cfe0863a-a698-41de-acaa-5f12bc415f40 · outbound
Inference Compute-Optimal Video Vision Language Models Unresolved cited work
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 225eb203-084c-4bf1-9632-343f19633839 · outbound
Inference Compute-Optimal Video Vision Language Models Unresolved cited work
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation dae1feb4-8d23-49ed-b72d-8a26f71dc8f0 · outbound
Inference Compute-Optimal Video Vision Language Models Video-MME: The First-Ever Comprehensive Evaluation Benchmark of Multi-modal LLMs in Video Analysis
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 82eb4ea5-b4b3-4ec3-b6cd-ee3f20466cc4 · outbound
Inference Compute-Optimal Video Vision Language Models Language models scale reliably with over-training and on downstream tasks
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 64e41c9e-d7c3-49a0-b4c3-56361f1fc986 · outbound
Inference Compute-Optimal Video Vision Language Models Unresolved cited work
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 17972598-b0f3-4be8-87d7-d42a2626423b · outbound
Inference Compute-Optimal Video Vision Language Models The Llama 3 Herd of Models
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3059b274-9953-4fa7-b4b1-edfdf933cd30 · outbound
Inference Compute-Optimal Video Vision Language Models Scaling Laws for Autoregressive Generative Modeling
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d7ea4732-2031-48f9-80b1-ee4f71406ba6 · outbound
Inference Compute-Optimal Video Vision Language Models Scaling Laws for Transfer
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 75dfb8b4-0c70-4059-b403-b5863bfddc5c · outbound
Inference Compute-Optimal Video Vision Language Models Rae, and Laurent Sifre
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 832018fe-7335-4025-a10c-140e8f56772b · outbound
Inference Compute-Optimal Video Vision Language Models Scaling Laws for Neural Language Models
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 700f3115-d33a-482f-8868-b24cd634b6b4 · outbound
Inference Compute-Optimal Video Vision Language Models The Kinetics Human Action Video Dataset
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 20e5dee2-3837-4158-82c3-7d59c0ea90f4 · outbound
Inference Compute-Optimal Video Vision Language Models Unresolved cited work
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8baaa746-2b83-48fd-8bb6-2757cfeaa36b · outbound
Inference Compute-Optimal Video Vision Language Models Unresolved cited work
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1b39e2d7-5d94-43cc-8e60-8b800fd2fd05 · outbound
Inference Compute-Optimal Video Vision Language Models Unresolved cited work
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ae500613-91fa-49a2-84f4-013c76e1b21d · outbound
Inference Compute-Optimal Video Vision Language Models Unresolved cited work
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 012bea90-155a-46ab-b8f5-374c85bdc1b7 · outbound
Inference Compute-Optimal Video Vision Language Models Semedo, and J
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation cd519b10-c0cf-4a68-b1d6-db4ab885c164 · outbound
Inference Compute-Optimal Video Vision Language Models Unresolved cited work
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 5d35147b-7e4c-4b6c-86d3-d1f956b47739 · outbound
Inference Compute-Optimal Video Vision Language Models Unresolved cited work
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation da8d7d57-e45e-42a6-988c-7b8dffdf5fa9 · outbound
Inference Compute-Optimal Video Vision Language Models Unresolved cited work
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 64ba8ce0-7144-42a4-a354-00183e87280d · outbound
Inference Compute-Optimal Video Vision Language Models VideoGPT+: Integrating Image and Video Encoders for Enhanced Video Understanding
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 23cde5af-c029-4f38-bb79-fabeee9c418b · outbound
Inference Compute-Optimal Video Vision Language Models Unresolved cited work
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 00a701f6-41b5-4a06-84e4-d6a0ca130aa0 · outbound
Inference Compute-Optimal Video Vision Language Models How predictable is language model benchmark performance?
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fb68088f-4013-445f-8c70-7f4241c908cd · outbound
Inference Compute-Optimal Video Vision Language Models Unresolved cited work
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 4f092b22-d69b-49fd-9930-45e28eacb26c · outbound
Inference Compute-Optimal Video Vision Language Models Unresolved cited work
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 52dfa61b-d7cf-4989-a948-faa6421cf727 · outbound
Inference Compute-Optimal Video Vision Language Models Unresolved cited work
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7cd98684-6596-402f-89b1-81ad787cb056 · outbound
Inference Compute-Optimal Video Vision Language Models Unresolved cited work
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation da2b4e41-c821-4489-9c49-4f823dd0ab13 · outbound
Inference Compute-Optimal Video Vision Language Models Unresolved cited work
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ca258934-4862-4c09-abae-844fffef3dd6 · outbound
Inference Compute-Optimal Video Vision Language Models Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9e248fc2-dd7a-48e3-b654-ab2c160384c4 · outbound
Inference Compute-Optimal Video Vision Language Models Dai, and Quoc V
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 8bbc438b-6c32-454a-abd8-7090e9496fef · outbound
Inference Compute-Optimal Video Vision Language Models Unresolved cited work
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 11c8e44c-a114-40d8-8116-a0eb7f7e363a · outbound
Inference Compute-Optimal Video Vision Language Models Unresolved cited work
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation aa0e8821-e279-4923-b1cf-f6b99f9c6baa · outbound
Inference Compute-Optimal Video Vision Language Models Unresolved cited work
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation addd37c0-2468-411a-954a-0e2edbbe2ff7 · outbound
Inference Compute-Optimal Video Vision Language Models Tenenbaum
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 6e660bee-2e3e-496c-ab59-c8d256ac7f3b · outbound
Inference Compute-Optimal Video Vision Language Models Unresolved cited work
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cf439be7-efd2-4428-9269-084f4475005d · outbound
Inference Compute-Optimal Video Vision Language Models Unresolved cited work
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7638903b-bae8-48b6-845f-76c5fed51d4a · outbound
Inference Compute-Optimal Video Vision Language Models Unresolved cited work
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 665b20c0-642a-4749-adce-bb09dbca8641 · outbound
Inference Compute-Optimal Video Vision Language Models LMMs-Eval: Reality Check on the Evaluation of Large Multimodal Models
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation efd54e7f-d6f5-4461-9d60-0546de327bce · outbound
Inference Compute-Optimal Video Vision Language Models Unresolved cited work
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 08d2a24f-71fa-421f-afa5-f205f726ef01 · outbound
Inference Compute-Optimal Video Vision Language Models LLaVA-Video: Video Instruction Tuning With Synthetic Data
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 83da37bd-b530-4ae1-a133-ab1ae060601e · inbound
Video Parallel Scaling: Aggregating Diverse Frame Subsets for VideoLLMs Inference Compute-Optimal Video Vision Language Models
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.