Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-04T19:28:35.241017Z
Paper Citation Record · LEDGER
As of 21 August 2026, this Paper Citation Record lists 46 of 46 outbound references and 1 inbound Pith citation observation for arXiv:2509.09263.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-04T19:28:35.241017Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-07-13T21:37:55.887477Z
A source-named dated measurement, never combined with another source.
Source: cited_works
46 of 46 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 6cea1c6b-09c3-4af2-9218-279d935743ce · outbound
DATE: Dynamic Absolute Time Enhancement for Long Video Understanding Flamingo: a visual language model for few-shot learning,
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6c75c933-e1ba-4588-86d6-462aacd90fcb · outbound
DATE: Dynamic Absolute Time Enhancement for Long Video Understanding VideoLLaMA 2: Advancing Spatial-Temporal Modeling and Audio Understanding in Video-LLMs
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5dd4e9d0-4a99-4d0f-98db-59072db3e3dd · outbound
DATE: Dynamic Absolute Time Enhancement for Long Video Understanding Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e166203b-70e5-4cfd-8f51-2b98926b70e7 · outbound
DATE: Dynamic Absolute Time Enhancement for Long Video Understanding Qwen2.5-VL Technical Report
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e97afe11-87d4-490f-8de7-50293f2847aa · outbound
DATE: Dynamic Absolute Time Enhancement for Long Video Understanding Roformer: Enhanced transformer with rotary position embedding,
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cf1d15db-9e10-4c41-918f-dccbe1e2c010 · outbound
DATE: Dynamic Absolute Time Enhancement for Long Video Understanding Adaptive Keyframe Sampling for Long Video Understanding
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 29e0a2e4-663f-4f7d-943e-8ffade73c217 · outbound
DATE: Dynamic Absolute Time Enhancement for Long Video Understanding Learning transferable visual models from natural language supervision,
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6634ecc5-1a9a-4aef-b5a1-b43d5f433feb · outbound
DATE: Dynamic Absolute Time Enhancement for Long Video Understanding GPT-4 Technical Report
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cc6ebdbc-ffea-4b5d-89ec-aad8bc74d4bc · outbound
DATE: Dynamic Absolute Time Enhancement for Long Video Understanding Language models are few-shot learners,
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b0bfc4e1-4f82-4885-950b-01b2402ad1bb · outbound
DATE: Dynamic Absolute Time Enhancement for Long Video Understanding Vicuna: An open-source chatbot impressing gpt-4 with 90%* chatgpt quality,
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ca93cbda-637c-41bd-8fae-f84d434f64dc · outbound
DATE: Dynamic Absolute Time Enhancement for Long Video Understanding Palm: Scaling language modeling with pathways,
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0e27cd1c-cb58-48de-9798-7d54c9a12564 · outbound
DATE: Dynamic Absolute Time Enhancement for Long Video Understanding Scaling instruction-finetuned language models,
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7c3db572-5161-4eea-9d24-ce54b66efbae · outbound
DATE: Dynamic Absolute Time Enhancement for Long Video Understanding The Llama 3 Herd of Models
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cce955e1-8ccc-49e4-9d1a-fafbafd35fc2 · outbound
DATE: Dynamic Absolute Time Enhancement for Long Video Understanding LLaMA: Open and Efficient Foundation Language Models
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5ce0c97a-7680-43e3-a505-1d8127a1d2fd · outbound
DATE: Dynamic Absolute Time Enhancement for Long Video Understanding Llama 2: Open Foundation and Fine-Tuned Chat Models
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 46c9fea0-ba0c-4c96-9e4f-5f447ab6a67b · outbound
DATE: Dynamic Absolute Time Enhancement for Long Video Understanding Chatgpt: A comprehensive review on background, applications, key challenges, bias, ethics, limitations and future scope,
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 857f07fd-9954-45a6-ad11-5ab223d512e0 · outbound
DATE: Dynamic Absolute Time Enhancement for Long Video Understanding Fewer Tokens and Fewer Videos: Extending Video Understanding Abilities in Large Vision-Language Models
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 55a0459e-e30a-47eb-93d2-2ea7c978ded2 · outbound
DATE: Dynamic Absolute Time Enhancement for Long Video Understanding Lisa: Reasoning segmentation via large language model,
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5bfac942-3429-4c47-9b39-dfd7e4d5055a · outbound
DATE: Dynamic Absolute Time Enhancement for Long Video Understanding Visual instruction tuning,
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4c588222-a42d-4351-b04c-d24c9f3e7c07 · outbound
DATE: Dynamic Absolute Time Enhancement for Long Video Understanding Video-ChatGPT: Towards Detailed Video Understanding via Large Vision and Language Models
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 93b70808-69bd-43bb-b880-3955f547ac61 · outbound
DATE: Dynamic Absolute Time Enhancement for Long Video Understanding Sharegpt4video: Improving video understanding and generation with better captions,
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 147cd8f3-db37-4d87-995b-fb11bb041f78 · outbound
DATE: Dynamic Absolute Time Enhancement for Long Video Understanding Dibs: Enhancing dense video captioning with unlabeled videos via pseudo boundary enrichment and online refinement,
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 476d495c-d00b-44d6-919c-42593243e6ca · outbound
DATE: Dynamic Absolute Time Enhancement for Long Video Understanding Morevqa: Exploring modular reasoning models for video question answering,
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7efcc816-3e85-4a3e-9cc1-8dabf50b7804 · outbound
DATE: Dynamic Absolute Time Enhancement for Long Video Understanding Momentor: Advancing Video Large Language Model with Fine-Grained Temporal Reasoning
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bb5557a8-9227-45ae-abb6-bc505cfc1a82 · outbound
DATE: Dynamic Absolute Time Enhancement for Long Video Understanding Negative sample matters: A renaissance of metric learning for temporal grounding,
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b59b4d2e-d85d-439c-89f8-bb58049600ca · outbound
DATE: Dynamic Absolute Time Enhancement for Long Video Understanding Video-LLaVA: Learning United Visual Representation by Alignment Before Projection
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2a6b7a38-761c-4dc1-80c4-79e67f6e1854 · outbound
DATE: Dynamic Absolute Time Enhancement for Long Video Understanding LLaVA-Video: Video Instruction Tuning With Synthetic Data
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fdb33a5f-58f8-4185-8734-36fbeeb8005a · outbound
DATE: Dynamic Absolute Time Enhancement for Long Video Understanding Video-LLaMA: An Instruction-tuned Audio-Visual Language Model for Video Understanding
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 98bc26e1-b95b-4169-a7de-4e94406e3e86 · outbound
DATE: Dynamic Absolute Time Enhancement for Long Video Understanding VideoLLaMA 3: Frontier Multimodal Foundation Models for Image and Video Understanding
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b36bf2d3-60fd-49bb-82f2-07bc32b5146b · outbound
DATE: Dynamic Absolute Time Enhancement for Long Video Understanding Video-MME: The First-Ever Comprehensive Evaluation Benchmark of Multi-modal LLMs in Video Analysis
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6b128814-feca-4e5a-bb51-019d44bc8eb8 · outbound
DATE: Dynamic Absolute Time Enhancement for Long Video Understanding Longvideobench: A benchmark for long-context interleaved video- language understanding,
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 69d74be3-8a17-49e7-b34e-6e81746dc78f · outbound
DATE: Dynamic Absolute Time Enhancement for Long Video Understanding LVBench: An Extreme Long Video Understanding Benchmark
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation feb1c0be-5028-459b-94a2-88cc60b01236 · outbound
DATE: Dynamic Absolute Time Enhancement for Long Video Understanding Interpolating Video-LLMs: Toward Longer-sequence LMMs in a Training-free Manner
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a75667df-189b-40d5-a2e7-ab369304ba27 · outbound
DATE: Dynamic Absolute Time Enhancement for Long Video Understanding Long Context Transfer from Language to Vision
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 22e1696a-3703-4675-9506-ef2c6cae2c87 · outbound
DATE: Dynamic Absolute Time Enhancement for Long Video Understanding Visual Context Window Extension: A New Perspective for Long Video Understanding
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation efb52306-5104-48a7-aef7-4735da755979 · outbound
DATE: Dynamic Absolute Time Enhancement for Long Video Understanding LongVILA: Scaling Long-Context Visual Language Models for Long Videos
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c2c22972-8ae0-4bfd-8dab-8065d9bb208b · outbound
DATE: Dynamic Absolute Time Enhancement for Long Video Understanding AdaReTaKe: Adaptive Redundancy Reduction to Perceive Longer for Video-language Understanding
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bdb74b2a-7170-4308-8f2e-da0a092a7ddb · outbound
DATE: Dynamic Absolute Time Enhancement for Long Video Understanding Enhancing Long Video Understanding via Hierarchical Event-Based Memory
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a2c89837-b407-4cab-898d-d6e0d3c2c885 · outbound
DATE: Dynamic Absolute Time Enhancement for Long Video Understanding ZipVL: Efficient Large Vision-Language Models with Dynamic Token Sparsification
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c73f5009-f744-476c-ad1e-6f9c6992c300 · outbound
DATE: Dynamic Absolute Time Enhancement for Long Video Understanding Ma-lmm: Memory- augmented large multimodal model for long-term video understanding,
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3b556654-4211-4f6f-b6dd-fc0ec69fb8fd · outbound
DATE: Dynamic Absolute Time Enhancement for Long Video Understanding TimeMarker: A Versatile Video-LLM for Long and Short Video Understanding with Superior Temporal Localization Ability
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fd14bf21-e9ae-4f60-9646-65fb995f1861 · outbound
DATE: Dynamic Absolute Time Enhancement for Long Video Understanding Timechat: A time-sensitive multimodal large language model for long video understanding,
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0ec38891-e23f-4f47-a80a-d702265396f9 · outbound
DATE: Dynamic Absolute Time Enhancement for Long Video Understanding Qwen2.5-Omni Technical Report
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9800e012-7f9c-4229-b3cf-744921626fca · outbound
DATE: Dynamic Absolute Time Enhancement for Long Video Understanding LLaVA-OneVision: Easy Visual Task Transfer
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 65c7535d-2b92-4fda-8351-eed921ad8a2a · outbound
DATE: Dynamic Absolute Time Enhancement for Long Video Understanding BOLT: Boost Large Vision-Language Model Without Training for Long-form Video Understanding
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 04b4fa8c-bc72-4680-8df4-4e81a2d28cd1 · outbound
DATE: Dynamic Absolute Time Enhancement for Long Video Understanding DeepSeek-V3 Technical Report
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5c04fc4a-164b-495d-8250-15706bd20a54 · inbound
CoVR-R:Reason-Aware Composed Video Retrieval DATE: Dynamic Absolute Time Enhancement for Long Video Understanding
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.