Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T05:24:17.842065Z
Paper Citation Record · LEDGER
As of 13 August 2026, this Paper Citation Record lists 47 of 47 outbound references and 2 inbound Pith citation observations for arXiv:2412.17637.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T05:24:17.842065Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-05-10T19:10:07.699225Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-10T23:25:51.674329Z
47 of 47 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 8b6535f5-61bb-4049-acd1-2fe75a8907d3 · outbound
SCBench: A Sports Commentary Benchmark for Video LLMs write newline
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e55d5ec1-caed-4f9d-bfec-fc3518af17e9 · outbound
SCBench: A Sports Commentary Benchmark for Video LLMs Spice: Semantic propositional image caption evaluation, 2016
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 92643bb3-53a7-4370-82e1-8eb769cc1569 · outbound
SCBench: A Sports Commentary Benchmark for Video LLMs METEOR : An automatic metric for MT evaluation with improved correlation with human judgments
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 81aa4ebc-e894-4060-853e-1f87748c46ba · outbound
SCBench: A Sports Commentary Benchmark for Video LLMs P2anet: A dataset and benchmark for dense action detection from table tennis match broadcasting videos, 2024
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation d38a36f9-cdc7-4668-81bc-d2bff5b5d90e · outbound
SCBench: A Sports Commentary Benchmark for Video LLMs Unresolved cited work
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 01fb1590-e1ab-4220-ae56-24e768fe32bf · outbound
SCBench: A Sports Commentary Benchmark for Video LLMs Autoeval-video: An automatic benchmark for assessing large vision language models in open-ended video question answering, 2024 a
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 4885a10a-2b21-4003-bcdf-004717ab064e · outbound
SCBench: A Sports Commentary Benchmark for Video LLMs Internvl: Scaling up vision foundation models and aligning for generic visual-linguistic tasks, 2024 b
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 49590249-4420-459c-86fb-112a787c439a · outbound
SCBench: A Sports Commentary Benchmark for Video LLMs Sports re-id: Improving re-identification of players in broadcast videos of team sports, 2022
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation a09997db-526f-4884-9426-88cfca800994 · outbound
SCBench: A Sports Commentary Benchmark for Video LLMs Seikavandi, Jacob V
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation bd052060-cac3-4270-9fd0-2fc4fa6f6edc · outbound
SCBench: A Sports Commentary Benchmark for Video LLMs Mmbench-video: A long-form multi-shot benchmark for holistic video understanding, 2024
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation b6b813a9-e419-4a29-be55-548f911c8fe4 · outbound
SCBench: A Sports Commentary Benchmark for Video LLMs Unresolved cited work
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 201e4a8c-2feb-4498-8750-0a5980363c45 · outbound
SCBench: A Sports Commentary Benchmark for Video LLMs Chatpose: Chatting about 3d human pose
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation ee94795f-eb39-4741-ba3c-83ad9bbe4e93 · outbound
SCBench: A Sports Commentary Benchmark for Video LLMs Video-mme: The first-ever comprehensive evaluation benchmark of multi-modal llms in video analysis, 2024
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f753cd98-bff0-460d-ac1f-22c9961d08f5 · outbound
SCBench: A Sports Commentary Benchmark for Video LLMs Mini-internvl: A flexible-transfer pocket multimodal model with 5
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 5dab9f72-2add-4922-81fc-e92616989e7d · outbound
SCBench: A Sports Commentary Benchmark for Video LLMs Tgif-qa: Toward spatio-temporal reasoning in visual question answering, 2017
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation f367ceff-082e-4348-8159-16ea306ededa · outbound
SCBench: A Sports Commentary Benchmark for Video LLMs Chat-univi: Unified visual representation empowers large language models with image and video understanding, 2024
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 7fa9b383-63be-4c0d-96d7-e83b3e6a0fb5 · outbound
SCBench: A Sports Commentary Benchmark for Video LLMs Retrieval-augmented generation for knowledge-intensive nlp tasks, 2021
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d0c9d076-ab13-4cf7-ba77-886c8f7d102d · outbound
SCBench: A Sports Commentary Benchmark for Video LLMs Mvbench: A comprehensive multi-modal video understanding benchmark, 2024
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 63c2705e-2a9b-4b8f-94ca-53e371c0145a · outbound
SCBench: A Sports Commentary Benchmark for Video LLMs Video-llava: Learning united visual representation by alignment before projection, 2023
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a8cb6f5d-a8c1-4014-abe8-39cf5a8b1c5f · outbound
SCBench: A Sports Commentary Benchmark for Video LLMs ROUGE : A package for automatic evaluation of summaries
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2dde1edb-3b92-4512-b7ae-b2c24263de32 · outbound
SCBench: A Sports Commentary Benchmark for Video LLMs Vila: On pre-training for visual language models, 2024
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 272a74c3-c035-45fa-b6d7-4b5bcc9a23e1 · outbound
SCBench: A Sports Commentary Benchmark for Video LLMs Llava-next: Improved reasoning, ocr, and world knowledge, 2024 a
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 5f312c6c-c424-4e8b-9f60-e60a9de35532 · outbound
SCBench: A Sports Commentary Benchmark for Video LLMs Kangaroo: A powerful video-language model supporting long-context video input, 2024 b
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation d441747c-01e2-428a-a79d-1b96fcec39e9 · outbound
SCBench: A Sports Commentary Benchmark for Video LLMs Fineaction: A fine-grained video dataset for temporal action localization
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation a5b2fb07-83a1-4be2-8113-58add5c6f016 · outbound
SCBench: A Sports Commentary Benchmark for Video LLMs Tempcompass: Do video llms really understand videos?, 2024 c
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 495a771b-3799-45de-8bc8-c6f62180a8e1 · outbound
SCBench: A Sports Commentary Benchmark for Video LLMs Cross-block fine-grained semantic cascade for skeleton-based sports action recognition, 2024 d
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 1bf485d7-49e8-4bf9-87c5-ea611345cbe1 · outbound
SCBench: A Sports Commentary Benchmark for Video LLMs Egoschema: A diagnostic benchmark for very long-form video language understanding, 2023
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 914eb97b-14ed-45ae-b0f7-c9c1f81aa64d · outbound
SCBench: A Sports Commentary Benchmark for Video LLMs Howto100m: Learning a text-video embedding by watching hundred million narrated video clips
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 790bf743-95bc-4de3-be39-e8ff36a81b80 · outbound
SCBench: A Sports Commentary Benchmark for Video LLMs Video-Bench: A Comprehensive Benchmark and Toolkit for Evaluating Video-based Large Language Models
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 00e92d4e-b813-4d32-9c87-1de79dd35e53 · outbound
SCBench: A Sports Commentary Benchmark for Video LLMs Video-bench: A comprehensive benchmark and toolkit for evaluating video-based large language models, 2023 b
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 318f1244-9d23-4391-96f8-c4276da2d98e · outbound
SCBench: A Sports Commentary Benchmark for Video LLMs B leu: a method for automatic evaluation of machine translation
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 1e62d750-37c1-48f2-a52e-ff97f2b59ddd · outbound
SCBench: A Sports Commentary Benchmark for Video LLMs Perception test: A diagnostic benchmark for multimodal video models, 2023
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 7ec357c2-f255-4e71-b76b-c16c38beb508 · outbound
SCBench: A Sports Commentary Benchmark for Video LLMs A survey of video datasets for grounded event understanding
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 194d3318-23d9-447d-b2c2-4ee768d1fea6 · outbound
SCBench: A Sports Commentary Benchmark for Video LLMs Finegym: A hierarchical video dataset for fine-grained action understanding, 2020
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1ad71774-beb3-43d1-9c46-8f1b44452bf0 · outbound
SCBench: A Sports Commentary Benchmark for Video LLMs Visual cot: Advancing multi-modal language models with a comprehensive dataset and benchmark for chain-of-thought reasoning, 2024
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 22ceb1ea-6e79-4b8f-a840-4c7cec682770 · outbound
SCBench: A Sports Commentary Benchmark for Video LLMs Playertv: Advanced player tracking and identification for automatic soccer highlight clips, 2024
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 0e4d8d06-c469-4c5d-b539-8f68e0a4bde8 · outbound
SCBench: A Sports Commentary Benchmark for Video LLMs Internvl2: Better than the best—expanding performance boundaries of open-source multimodal models with the progressive scaling strategy
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 712f6b79-4947-472c-a454-94d31e375393 · outbound
SCBench: A Sports Commentary Benchmark for Video LLMs CIDEr: Consensus-based Image Description Evaluation
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 48281dc9-9464-4065-9549-cac8cc4408b2 · outbound
SCBench: A Sports Commentary Benchmark for Video LLMs Lvbench: An extreme long video understanding benchmark, 2024
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9138a27a-d2d9-4035-8d99-52e1913656b4 · outbound
SCBench: A Sports Commentary Benchmark for Video LLMs Chain-of-thought prompting elicits reasoning in large language models, 2023
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2b9fd42c-5dab-4c68-852f-ececbfabeb6a · outbound
SCBench: A Sports Commentary Benchmark for Video LLMs Sportshhi: A dataset for human-human interaction detection in sports videos, 2024
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 191afc41-22d3-4cb1-8203-7215a9c1bc65 · outbound
SCBench: A Sports Commentary Benchmark for Video LLMs Next-qa:next phase of question-answering to explaining temporal actions
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 5a12612e-03b0-4cd0-92ba-11cf5040ecb1 · outbound
SCBench: A Sports Commentary Benchmark for Video LLMs Video question answering via gradually refined attention over appearance and motion
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 6ece9f76-4af6-42f9-8fe9-ec73996239ee · outbound
SCBench: A Sports Commentary Benchmark for Video LLMs Youku-mplug: A 10 million large-scale chinese video-language dataset for pre-training and benchmarks, 2023
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 2f72312a-1c0c-4b73-93f4-b98ac8a319bb · outbound
SCBench: A Sports Commentary Benchmark for Video LLMs Activitynet-qa: A dataset for understanding complex web videos via question answering, 2019
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ec67df3c-0e60-44f4-b323-567b1480ea59 · outbound
SCBench: A Sports Commentary Benchmark for Video LLMs Long context transfer from language to vision, 2024
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f966328a-0266-435b-8b64-9392d51e1834 · outbound
SCBench: A Sports Commentary Benchmark for Video LLMs A Survey on Generative AI and LLM for Video Generation, Understanding, and Streaming
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9412f7ea-5c9d-4aa8-b13c-2020277ff3bd · inbound
BoxComm: Benchmarking Category-Aware Commentary Generation and Narration Rhythm in Boxing SCBench: A Sports Commentary Benchmark for Video LLMs
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 5ba8e445-dd51-4bbc-964e-2b2d7c78c243 · inbound
RefereeBench: Are Video MLLMs Ready to be Multi-Sport Referees SCBench: A Sports Commentary Benchmark for Video LLMs
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.