Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 30 inbound Pith citation observations for arXiv:2409.14485.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T15:34:43.251039Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T06:39:37.682244Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 60d92558-0dc5-40b6-b968-b9685ac8c8b2 · inbound
MLVU: Benchmarking Multi-task Long Video Understanding Video-XL: Extra-Long Vision Language Model for Hour-Scale Video Understanding
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1cf91e4a-058b-46e0-825a-d4ae68ca3742 · inbound
NVILA: Efficient Frontier Visual Language Models Video-XL: Extra-Long Vision Language Model for Hour-Scale Video Understanding
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 731ff6fb-eb7b-48c7-b925-bcf2ff8d0414 · inbound
VideoChat-Flash: Hierarchical Compression for Long-Context Video Modeling Video-XL: Extra-Long Vision Language Model for Hour-Scale Video Understanding
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8467a30b-0658-47f3-8e17-081184e9bdde · inbound
InternVideo2.5: Empowering Video MLLMs with Long and Rich Context Modeling Video-XL: Extra-Long Vision Language Model for Hour-Scale Video Understanding
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5523b6de-9a63-4e83-8221-c14dd59cc6bc · inbound
VideoEval-Pro: Robust and Realistic Long Video Understanding Evaluation Video-XL: Extra-Long Vision Language Model for Hour-Scale Video Understanding
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8d9c289e-2d5a-4082-a8ec-fcc72a4a1192 · inbound
ViaRL: Adaptive Temporal Grounding via Visual Iterated Amplification Reinforcement Learning Video-XL: Extra-Long Vision Language Model for Hour-Scale Video Understanding
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f19b5cbc-5429-425a-8010-6d84e9d4ed61 · inbound
Threading Keyframe with Narratives: MLLMs as Strong Long Video Comprehenders Video-XL: Extra-Long Vision Language Model for Hour-Scale Video Understanding
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 358355fd-9998-444c-aae7-b87c1418de3a · inbound
FlexSelect: Flexible Token Selection for Efficient Long Video Understanding Video-XL: Extra-Long Vision Language Model for Hour-Scale Video Understanding
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c8017646-7848-455d-855d-be4096fd1aca · inbound
Vid-SME: Membership Inference Attacks against Large Video Understanding Models Video-XL: Extra-Long Vision Language Model for Hour-Scale Video Understanding
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e951b058-651e-41cb-82f6-f9ca968c09dd · inbound
CyberV: Cybernetics for Test-time Scaling in Video Understanding Video-XL: Extra-Long Vision Language Model for Hour-Scale Video Understanding
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b3b21a56-e1ad-4778-a6ce-c322c3a42042 · inbound
Task-Aware KV Compression For Cost-Effective Long Video Understanding Video-XL: Extra-Long Vision Language Model for Hour-Scale Video Understanding
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 51b1b8f9-be18-4d35-bc41-a649916768d0 · inbound
Q-Frame: Query-aware Frame Selection and Multi-Resolution Adaptation for Video-LLMs Video-XL: Extra-Long Vision Language Model for Hour-Scale Video Understanding
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bee69751-5848-4f41-9434-e3ee6d7c25b8 · inbound
MANTA: Cross-Modal Semantic Alignment and Information-Theoretic Optimization for Long-form Multimodal Understanding Video-XL: Extra-Long Vision Language Model for Hour-Scale Video Understanding
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 402060c6-11ee-452e-afb6-8d77ad4cb01e · inbound
LongAnimation: Long Animation Generation with Dynamic Global-Local Memory Video-XL: Extra-Long Vision Language Model for Hour-Scale Video Understanding
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 11789708-fe6c-41c0-b5a3-1c112e649775 · inbound
AuroraLong: Bringing RNNs Back to Efficient Open-Ended Video Understanding Video-XL: Extra-Long Vision Language Model for Hour-Scale Video Understanding
Reference 76
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a03ecc67-4006-4690-9bf6-bdae02f2ad6f · inbound
Iterative Zoom-In: Temporal Interval Exploration for Long Video Understanding Video-XL: Extra-Long Vision Language Model for Hour-Scale Video Understanding
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c2df738d-2eb6-4bef-b7fd-a788422984d2 · inbound
Infinite Video Understanding Video-XL: Extra-Long Vision Language Model for Hour-Scale Video Understanding
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e6347225-408e-4c8d-a637-f6576a7f0245 · inbound
ExpStar: Towards Automatic Commentary Generation for Multi-discipline Scientific Experiments Video-XL: Extra-Long Vision Language Model for Hour-Scale Video Understanding
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dab9472d-2449-4130-acf1-d9b6d776d216 · inbound
VRU-Accident: A Vision-Language Benchmark for Video Question Answering and Dense Captioning for Accident Scene Understanding Video-XL: Extra-Long Vision Language Model for Hour-Scale Video Understanding
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e27b375b-5317-4540-b42d-d6fe98846f67 · inbound
VAGU & GtS: LLM-Based Benchmark and Framework for Joint Video Anomaly Grounding and Understanding Video-XL: Extra-Long Vision Language Model for Hour-Scale Video Understanding
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b7efc35b-e33d-4610-9fea-8e2297ab618c · inbound
MESH -- Understanding Videos Like Human: Measuring Hallucinations in Large Video Models Video-XL: Extra-Long Vision Language Model for Hour-Scale Video Understanding
Reference 68
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b0b2081b-5f53-4d21-8b42-08d2cc83303b · inbound
Active Video Perception: Iterative Evidence Seeking for Agentic Long Video Understanding Video-XL: Extra-Long Vision Language Model for Hour-Scale Video Understanding
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation be7f6567-0ae3-4f92-81cc-1cc6f5346e46 · inbound
Seeing the Scene Matters: Revealing Forgetting in Video Understanding Models with a Scene-Aware Long-Video Benchmark Video-XL: Extra-Long Vision Language Model for Hour-Scale Video Understanding
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 32f16455-d6c1-4bea-9e65-a38aff7fcc5c · inbound
POINTS-Long: Adaptive Dual-Mode Visual Reasoning in MLLMs Video-XL: Extra-Long Vision Language Model for Hour-Scale Video Understanding
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation eb07d3ab-fd95-49e8-babd-079361566bca · inbound
One Token per Highly Selective Frame: Towards Extreme Compression for Long Video Understanding Video-XL: Extra-Long Vision Language Model for Hour-Scale Video Understanding
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 396f883b-4be0-4ea3-8768-009ca23e2b43 · inbound
GOPAgen: Motion-Aware and Efficient Agentic Long-Video Understanding with Structural Memory and Hierarchical Reasoning Video-XL: Extra-Long Vision Language Model for Hour-Scale Video Understanding
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6dc7a3a2-5fda-4960-a6a6-90e408212562 · inbound
UniReason-Med: A Shared Grounded Reasoning Interface for 2D-to-3D Transfer in Medical VQA Video-XL: Extra-Long Vision Language Model for Hour-Scale Video Understanding
Reference 111
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a4d5c3f9-8265-4d0c-8731-2cf740e7b81c · inbound
InternVideo3: Agentify Foundation Models with Multimodal Contextual Reasoning Video-XL: Extra-Long Vision Language Model for Hour-Scale Video Understanding
Reference 253
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 73de4bda-be24-4798-8bc7-9b58a48d0114 · inbound
HPP: Hierarchical Programmatic Probing for Long Video Understanding by Decoupling Perception and Reasoning Video-XL: Extra-Long Vision Language Model for Hour-Scale Video Understanding
Reference 139
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7dc63e75-60e7-488f-827d-82080ef78b28 · inbound
CRAFT: Compression via Recursive Adaptive Fusion of Video Tokens for Vision-Language Models Video-XL: Extra-Long Vision Language Model for Hour-Scale Video Understanding
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.