Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T13:12:41.078055Z
Paper Citation Record · LEDGER
As of 10 August 2026, this Paper Citation Record lists 19 of 19 outbound references and 19 inbound Pith citation observations for arXiv:2507.20939.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T13:12:41.078055Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-06T00:03:28.047281Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T03:49:30.275005Z
19 of 19 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 09faad0a-b366-49d8-9cb6-2caf2d9a1b54 · outbound
ARC-Hunyuan-Video-7B: Structured Video Comprehension of Real-World Shorts Qwen2.5-VL Technical Report
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7548790d-90d4-4c51-8750-4ca490328737 · outbound
ARC-Hunyuan-Video-7B: Structured Video Comprehension of Real-World Shorts SEED-Bench: Benchmarking Multimodal LLMs with Generative Comprehension
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9b6c2843-4cbc-4135-a943-43880b395e63 · outbound
ARC-Hunyuan-Video-7B: Structured Video Comprehension of Real-World Shorts EgoPlan-Bench2: A Benchmark for Multimodal Large Language Model Planning in Real-World Scenarios
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b4a0a645-4785-4ba4-a8bd-573947db35c1 · outbound
ARC-Hunyuan-Video-7B: Structured Video Comprehension of Real-World Shorts Audio-Visual LLM for Video Understanding
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 121a0ef3-6a52-41a4-83b9-bcc9bd4d0e63 · outbound
ARC-Hunyuan-Video-7B: Structured Video Comprehension of Real-World Shorts video-SALMONN: Speech-Enhanced Audio-Visual Large Language Models
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f9929eab-f34d-4ddf-99e4-742369172e99 · outbound
ARC-Hunyuan-Video-7B: Structured Video Comprehension of Real-World Shorts video-SALMONN-o1: Reasoning-enhanced Audio-visual Large Language Model
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cd8c78b8-5b4e-411d-a750-cbaf2e8ec76f · outbound
ARC-Hunyuan-Video-7B: Structured Video Comprehension of Real-World Shorts video- salmonn 2: Captioning-enhanced audio-visual large language models.arXiv preprint arXiv:2506.15220,
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 36e2cfcd-788a-4d8c-85b5-7d2114fb4bea · outbound
ARC-Hunyuan-Video-7B: Structured Video Comprehension of Real-World Shorts Kwai Keye-VL Technical Report
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d90faea1-3e4d-4f84-9b8d-1851eb24d599 · outbound
ARC-Hunyuan-Video-7B: Structured Video Comprehension of Real-World Shorts Video-LLaMA: An Instruction-tuned Audio-Visual Language Model for Video Understanding
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7e81c92a-5039-4fec-b214-1c1dc739921b · outbound
ARC-Hunyuan-Video-7B: Structured Video Comprehension of Real-World Shorts Self-supervised product title rewrite for product listing ads
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 0330b3b7-885f-4f0f-b3ed-c34643bedbe1 · outbound
ARC-Hunyuan-Video-7B: Structured Video Comprehension of Real-World Shorts Pre-trained language model based ranking in baidu search
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 19e9010d-dacf-4543-9715-976def34db69 · outbound
ARC-Hunyuan-Video-7B: Structured Video Comprehension of Real-World Shorts EgoPlan-Bench: Benchmarking Multimodal Large Language Models for Human-Level Planning
Reference 2015
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 83b99e30-bf69-4500-bbe8-019717c45630 · outbound
ARC-Hunyuan-Video-7B: Structured Video Comprehension of Real-World Shorts DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 2017
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8adc1884-5841-4d92-9623-872b8dd0785b · outbound
ARC-Hunyuan-Video-7B: Structured Video Comprehension of Real-World Shorts CLIP2Video: Mastering Video-Text Retrieval via Image CLIP
Reference 2020
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4bb1c417-c44c-49eb-b3bd-7193a9df94e0 · outbound
ARC-Hunyuan-Video-7B: Structured Video Comprehension of Real-World Shorts VCR-Bench: A Comprehensive Evaluation Framework for Video Chain-of-Thought Reasoning
Reference 2021
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 929118e3-4139-417a-8134-569457a97108 · outbound
ARC-Hunyuan-Video-7B: Structured Video Comprehension of Real-World Shorts Qwen2.5-Omni Technical Report
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ef3698a5-928b-43b8-b734-8e05bca71011 · outbound
ARC-Hunyuan-Video-7B: Structured Video Comprehension of Real-World Shorts Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 17226eaf-2214-427b-b571-910ed26a187a · outbound
ARC-Hunyuan-Video-7B: Structured Video Comprehension of Real-World Shorts Video-Holmes: Can MLLM Think Like Holmes for Complex Video Reasoning?
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7da6f7e7-2ed8-4976-b4ca-a5e4d7ee1be7 · outbound
ARC-Hunyuan-Video-7B: Structured Video Comprehension of Real-World Shorts An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d82c0732-c3fe-44a9-bd50-e5bb579ddea6 · inbound
Empowering Nanoscale Connectivity through Molecular Communication: A Case Study of Virus Infection ARC-Hunyuan-Video-7B: Structured Video Comprehension of Real-World Shorts
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 016ec883-792f-4898-8ee7-42fd8c75a35d · inbound
OmniZip: Audio-Guided Dynamic Token Compression for Fast Omnimodal Large Language Models ARC-Hunyuan-Video-7B: Structured Video Comprehension of Real-World Shorts
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 431cd667-1fe8-4634-a1e5-be3605422d64 · inbound
AdaTooler-V: Adaptive Tool-Use for Images and Videos ARC-Hunyuan-Video-7B: Structured Video Comprehension of Real-World Shorts
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 895f2136-3282-4deb-91f7-27e2bff74873 · inbound
Streaming Video Instruction Tuning ARC-Hunyuan-Video-7B: Structured Video Comprehension of Real-World Shorts
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 30240866-491b-453d-8a29-00fb300b939d · inbound
Video-OPD: Efficient Post-Training of Multimodal Large Language Models for Temporal Video Grounding via On-Policy Distillation ARC-Hunyuan-Video-7B: Structured Video Comprehension of Real-World Shorts
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 0d5a81a4-f4a7-4ce7-9d4f-f670f02cb94f · inbound
Video-OPD: Efficient Post-Training of Multimodal Large Language Models for Temporal Video Grounding via On-Policy Distillation ARC-Hunyuan-Video-7B: Structured Video Comprehension of Real-World Shorts
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2f3689b5-1689-4621-8177-8e9a4589739a · inbound
OmniScript: Towards Audio-Visual Script Generation for Long-Form Cinematic Video ARC-Hunyuan-Video-7B: Structured Video Comprehension of Real-World Shorts
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 110d5eda-42cd-4a0c-a706-ced4bcebc4c6 · inbound
StoryTR: Narrative-Centric Video Temporal Retrieval with Theory of Mind Reasoning ARC-Hunyuan-Video-7B: Structured Video Comprehension of Real-World Shorts
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 87b80d37-3a54-4c85-9892-6f7cf8472eb7 · inbound
OmniRefine: Alignment-Aware Cooperative Compression for Efficient Omnimodal Large Language Models ARC-Hunyuan-Video-7B: Structured Video Comprehension of Real-World Shorts
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation a522d834-08cb-4dae-914a-d4ee04358f2d · inbound
Stage-adaptive Token Selection for Efficient Omni-modal LLMs ARC-Hunyuan-Video-7B: Structured Video Comprehension of Real-World Shorts
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 0f3610da-8dd6-4be3-aadc-32053dd315be · inbound
O-MARC: Omni Memory-Augmented Compression Distillation for Efficient Video Understanding ARC-Hunyuan-Video-7B: Structured Video Comprehension of Real-World Shorts
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation f4ef4a4c-eac4-4efb-a123-54a804582313 · inbound
CogniRoute: Learning to Route Social Evidence in Omni-Modal Models ARC-Hunyuan-Video-7B: Structured Video Comprehension of Real-World Shorts
Reference 107
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation f5d9aaa1-58f1-450d-aa96-2d51e362a60a · inbound
Learning to Deny: Action Denial in Multimodal Large Language Models ARC-Hunyuan-Video-7B: Structured Video Comprehension of Real-World Shorts
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation b1162d31-7bdc-41ae-b96f-3fdc4ef55fc5 · inbound
Temporal and Cross-Modal Alignment for Enhanced Audiovisual Video Captioning ARC-Hunyuan-Video-7B: Structured Video Comprehension of Real-World Shorts
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 182b0ee0-427f-43da-aee5-3e1c50ac9f1e · inbound
Parallelized Autoregressive Decoding for Omni-Modal Dense Video Captioning ARC-Hunyuan-Video-7B: Structured Video Comprehension of Real-World Shorts
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e773c642-522a-4375-8026-1b9b52077a05 · inbound
PercepCap: Video Captioner with Structured Spatio-Temporal Perception ARC-Hunyuan-Video-7B: Structured Video Comprehension of Real-World Shorts
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6d25759c-5f23-49ae-9276-95bc2c367031 · inbound
RefCaptioner: Multi-Reference Image-Grounded Video Captioning ARC-Hunyuan-Video-7B: Structured Video Comprehension of Real-World Shorts
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 70c59f4e-51a2-44fa-bca1-f081081e2100 · inbound
Allocation Before Ranking: Decoupled Token Compression for OmniLLMs ARC-Hunyuan-Video-7B: Structured Video Comprehension of Real-World Shorts
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation df338285-baa2-473b-a284-b677b5a98a8e · inbound
OmniPack: Unified Token Compression for Efficient Omni-modal Large Language Models ARC-Hunyuan-Video-7B: Structured Video Comprehension of Real-World Shorts
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.