Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-06-29T17:50:00.740770Z
Paper Citation Record · LEDGER
As of 10 August 2026, this Paper Citation Record lists 57 of 57 outbound references and 0 inbound Pith citation observations for arXiv:2605.26680.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-06-29T17:50:00.740770Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
57 of 57 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation cd60909a-2a0e-43f3-8e2c-d3ed7f1c64cf · outbound
DynFrame: Adaptive Reasoning-Driven Multimodal Framework with Dynamic Frame Augmentation for Complex Video Understanding Qwen3-VL Technical Report
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation a2dca62f-22ee-4a28-bac8-b4589b824905 · outbound
DynFrame: Adaptive Reasoning-Driven Multimodal Framework with Dynamic Frame Augmentation for Complex Video Understanding Qwen2.5-VL Technical Report
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation e2916a80-bca9-4063-b06e-8d2eedb67c33 · outbound
DynFrame: Adaptive Reasoning-Driven Multimodal Framework with Dynamic Frame Augmentation for Complex Video Understanding ReXTime: A Benchmark Suite for Reasoning-Across-Time in Videos
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 9a1382c3-4616-4011-9027-4ced28821d41 · outbound
DynFrame: Adaptive Reasoning-Driven Multimodal Framework with Dynamic Frame Augmentation for Complex Video Understanding Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation e6b53214-d894-4308-ad19-8923e8722b6b · outbound
DynFrame: Adaptive Reasoning-Driven Multimodal Framework with Dynamic Frame Augmentation for Complex Video Understanding Videozoomer: Reinforcement-learned temporal focusing for long video reasoning
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation b83c7ede-0965-4cb3-bd97-00d5d0cf9956 · outbound
DynFrame: Adaptive Reasoning-Driven Multimodal Framework with Dynamic Frame Augmentation for Complex Video Understanding Video-R1: Reinforcing Video Reasoning in MLLMs
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 7f4a8675-abfe-4692-a958-19d83c573a17 · outbound
DynFrame: Adaptive Reasoning-Driven Multimodal Framework with Dynamic Frame Augmentation for Complex Video Understanding Video-mme: The first-ever comprehensive evaluation benchmark of multi-modal llms in video analysis
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c78ab457-45da-4bcc-a101-5a727d13b26f · outbound
DynFrame: Adaptive Reasoning-Driven Multimodal Framework with Dynamic Frame Augmentation for Complex Video Understanding Love- r1: Advancing long video understanding with an adaptive zoom-in mechanism via multi-step reasoning
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation fb944604-a3fd-49d6-a8d4-57514e97112e · outbound
DynFrame: Adaptive Reasoning-Driven Multimodal Framework with Dynamic Frame Augmentation for Complex Video Understanding Tall: Temporal activity localization via language query
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 394b3f26-cdff-4683-9944-e0aa643365c6 · outbound
DynFrame: Adaptive Reasoning-Driven Multimodal Framework with Dynamic Frame Augmentation for Complex Video Understanding Gemini 3 Pro Model Card
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 72b1b12a-568c-471e-b8e7-cbc03326b264 · outbound
DynFrame: Adaptive Reasoning-Driven Multimodal Framework with Dynamic Frame Augmentation for Complex Video Understanding DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 0d0bc326-c0a6-4b3d-8765-bd9697045bbf · outbound
DynFrame: Adaptive Reasoning-Driven Multimodal Framework with Dynamic Frame Augmentation for Complex Video Understanding Framethinker: Learning to think with long videos via multi-turn frame spotlighting
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation c17be9e3-4f0d-4877-8f56-8a2df063a1b4 · outbound
DynFrame: Adaptive Reasoning-Driven Multimodal Framework with Dynamic Frame Augmentation for Complex Video Understanding Vision-R1: Incentivizing Reasoning Capability in Multimodal Large Language Models
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation ef8e1294-98b4-4ddf-bb3f-eeee7d438180 · outbound
DynFrame: Adaptive Reasoning-Driven Multimodal Framework with Dynamic Frame Augmentation for Complex Video Understanding GPT-4o System Card
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 8054942e-8613-4f4b-bc66-225174eb4613 · outbound
DynFrame: Adaptive Reasoning-Driven Multimodal Framework with Dynamic Frame Augmentation for Complex Video Understanding Chat-univi: Unified visual representation empowers large language models with image and video understanding
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e0eea99c-922c-43ce-a610-f15bafc3d2aa · outbound
DynFrame: Adaptive Reasoning-Driven Multimodal Framework with Dynamic Frame Augmentation for Complex Video Understanding Kimi-VL Technical Report
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 15a68bf9-b97b-4b4a-8ea3-abfaaf930ed2 · outbound
DynFrame: Adaptive Reasoning-Driven Multimodal Framework with Dynamic Frame Augmentation for Complex Video Understanding Dense-captioning events in videos
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1951d29b-8c83-4961-a868-06c5ff0cfc84 · outbound
DynFrame: Adaptive Reasoning-Driven Multimodal Framework with Dynamic Frame Augmentation for Complex Video Understanding LLaVA-OneVision: Easy Visual Task Transfer
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 5f329e02-4e84-42af-94a8-3b9914d312ba · outbound
DynFrame: Adaptive Reasoning-Driven Multimodal Framework with Dynamic Frame Augmentation for Complex Video Understanding Reinforcement Learning Tuning for VideoLLMs: Reward Design and Data Efficiency
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation ec21962f-3b64-4db1-883d-9015dceab48c · outbound
DynFrame: Adaptive Reasoning-Driven Multimodal Framework with Dynamic Frame Augmentation for Complex Video Understanding VideoChat-R1: Enhancing Spatio-Temporal Perception via Reinforcement Fine-Tuning
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation a3c115e8-2e73-4b43-85f0-1d82a03280be · outbound
DynFrame: Adaptive Reasoning-Driven Multimodal Framework with Dynamic Frame Augmentation for Complex Video Understanding VideoTemp-o3: Harmonizing Temporal Grounding and Video Understanding in Agentic Thinking-with-Videos
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 107391d0-6271-4530-93eb-512804475c8d · outbound
DynFrame: Adaptive Reasoning-Driven Multimodal Framework with Dynamic Frame Augmentation for Complex Video Understanding Video-rts: Rethinking reinforcement learning and test-time scaling for efficient and enhanced video reasoning.arXiv preprint arXiv:2507.06485, 2025
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 0c2ea2c2-2550-4613-8d19-f7645ee5f965 · outbound
DynFrame: Adaptive Reasoning-Driven Multimodal Framework with Dynamic Frame Augmentation for Complex Video Understanding Video-ChatGPT: Towards Detailed Video Understanding via Large Vision and Language Models
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation ff8b3412-52c2-45bb-a696-002569f432df · outbound
DynFrame: Adaptive Reasoning-Driven Multimodal Framework with Dynamic Frame Augmentation for Complex Video Understanding Open-o3 video: Grounded video reasoning with explicit spatio-temporal evidence
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 1644c19d-11a6-47f4-8706-7495d3d7a138 · outbound
DynFrame: Adaptive Reasoning-Driven Multimodal Framework with Dynamic Frame Augmentation for Complex Video Understanding Visual cot: Advancing multi-modal language models with a comprehensive dataset and benchmark for chain-of-thought reasoning
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a7303139-5c10-4923-a01b-e1d20d2ad53c · outbound
DynFrame: Adaptive Reasoning-Driven Multimodal Framework with Dynamic Frame Augmentation for Complex Video Understanding DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 0e945c59-2149-4293-8372-a9166d2c4bb0 · outbound
DynFrame: Adaptive Reasoning-Driven Multimodal Framework with Dynamic Frame Augmentation for Complex Video Understanding Temporal Grounding of Activities using Multimodal Large Language Models
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 4a0e107d-5a71-460e-9250-32ae08b5c9d0 · outbound
DynFrame: Adaptive Reasoning-Driven Multimodal Framework with Dynamic Frame Augmentation for Complex Video Understanding Adaptive keyframe sampling for long video understanding
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9b56a2e3-91fe-42c4-a10f-3c2b1b66a312 · outbound
DynFrame: Adaptive Reasoning-Driven Multimodal Framework with Dynamic Frame Augmentation for Complex Video Understanding GRPO-CARE: Consistency-Aware Reinforcement Learning for Multimodal Reasoning
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 5fbac1b6-d5cb-4ca0-b42d-7dab98014b6b · outbound
DynFrame: Adaptive Reasoning-Driven Multimodal Framework with Dynamic Frame Augmentation for Complex Video Understanding Videorft: Incentivizing video reasoning capability in mllms via reinforced fine-tuning
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 1d80d592-0544-49b3-8c17-2ea82b75c705 · outbound
DynFrame: Adaptive Reasoning-Driven Multimodal Framework with Dynamic Frame Augmentation for Complex Video Understanding LVBench: An Extreme Long Video Understanding Benchmark
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 92f3acad-7a0e-4a6e-85a3-d4edb8fd8e77 · outbound
DynFrame: Adaptive Reasoning-Driven Multimodal Framework with Dynamic Frame Augmentation for Complex Video Understanding Multimodal Chain-of-Thought Reasoning: A Comprehensive Survey
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 0391f5a4-5d2d-453f-829f-e644916d4f07 · outbound
DynFrame: Adaptive Reasoning-Driven Multimodal Framework with Dynamic Frame Augmentation for Complex Video Understanding Chain-of-thought prompting elicits reasoning in large language models.Advances in Neural Information Processing Systems, 35:24824–24837, 2022
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 36096b60-118f-48c3-8771-3ba85442b49e · outbound
DynFrame: Adaptive Reasoning-Driven Multimodal Framework with Dynamic Frame Augmentation for Complex Video Understanding Can I trust your answer? Visually grounded video question answering
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3d8fc918-8948-46c9-893b-e1e52b821ff8 · outbound
DynFrame: Adaptive Reasoning-Driven Multimodal Framework with Dynamic Frame Augmentation for Complex Video Understanding Videochat-r1.5: Visual test-time scaling to reinforce multimodal reasoning by iterative perception
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f51a4f7d-3e4e-4c37-8c49-fe4634c4e81d · outbound
DynFrame: Adaptive Reasoning-Driven Multimodal Framework with Dynamic Frame Augmentation for Complex Video Understanding LongVT: Incentivizing "Thinking with Long Videos" via Native Tool Calling
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 0029ac56-f577-430e-b1a3-1fbcf02a33d1 · outbound
DynFrame: Adaptive Reasoning-Driven Multimodal Framework with Dynamic Frame Augmentation for Complex Video Understanding Focus: Efficient keyframe selection for long video understanding
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 838367aa-8689-4197-9e11-14311c95be71 · outbound
DynFrame: Adaptive Reasoning-Driven Multimodal Framework with Dynamic Frame Augmentation for Complex Video Understanding Frame-voyager: Learning to query frames for video large language models
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f1224666-04e5-4bcb-be0a-0ba2d4c7743c · outbound
DynFrame: Adaptive Reasoning-Driven Multimodal Framework with Dynamic Frame Augmentation for Complex Video Understanding Video-o3: Native Interleaved Clue Seeking for Long Video Multi-Hop Reasoning
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 585e08e9-4107-4b7b-b8a9-e61d522f013e · outbound
DynFrame: Adaptive Reasoning-Driven Multimodal Framework with Dynamic Frame Augmentation for Complex Video Understanding Video-llama: An instruction-tuned audio-visual language model for video understanding
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c00334ef-161f-491f-be05-6bc80315fd53 · outbound
DynFrame: Adaptive Reasoning-Driven Multimodal Framework with Dynamic Frame Augmentation for Complex Video Understanding Thinking With Videos: Multimodal Tool-Augmented Reinforcement Learning for Long Video Reasoning
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation dbed8afa-7fa0-4bc1-af93-846a9da60741 · outbound
DynFrame: Adaptive Reasoning-Driven Multimodal Framework with Dynamic Frame Augmentation for Complex Video Understanding R1-VL: Learning to Reason with Multimodal Large Language Models via Step-wise Group Relative Policy Optimization
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation aff93130-9265-4d0a-b03c-6ca38dc78df7 · outbound
DynFrame: Adaptive Reasoning-Driven Multimodal Framework with Dynamic Frame Augmentation for Complex Video Understanding LLaVA-Video: Video Instruction Tuning With Synthetic Data
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 8f0e0d7e-22b1-4299-a4e3-4b78ef8ff049 · outbound
DynFrame: Adaptive Reasoning-Driven Multimodal Framework with Dynamic Frame Augmentation for Complex Video Understanding DeepEyes: Incentivizing "Thinking with Images" via Reinforcement Learning
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 4c05d73f-2aa9-4308-9574-8d453e49aa30 · outbound
DynFrame: Adaptive Reasoning-Driven Multimodal Framework with Dynamic Frame Augmentation for Complex Video Understanding MLVU: Benchmarking Multi-task Long Video Understanding
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 6afaf91e-c776-440d-8e98-62407c20a90d · outbound
DynFrame: Adaptive Reasoning-Driven Multimodal Framework with Dynamic Frame Augmentation for Complex Video Understanding InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 4931a756-9708-46db-88ca-18a2999ec714 · outbound
DynFrame: Adaptive Reasoning-Driven Multimodal Framework with Dynamic Frame Augmentation for Complex Video Understanding Unresolved cited work
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1ec86a77-a6e5-43c9-86e3-c15b10774c4c · outbound
DynFrame: Adaptive Reasoning-Driven Multimodal Framework with Dynamic Frame Augmentation for Complex Video Understanding Explain why this portion is critical
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4262fb4f-cbfe-4cb2-957e-1f27ade60ac6 · outbound
DynFrame: Adaptive Reasoning-Driven Multimodal Framework with Dynamic Frame Augmentation for Complex Video Understanding Unresolved cited work
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b5c64ec9-779e-47a3-9855-b09c6bfb9178 · outbound
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0e959838-c1dd-4c85-9232-371c320b6787 · outbound
DynFrame: Adaptive Reasoning-Driven Multimodal Framework with Dynamic Frame Augmentation for Complex Video Understanding Using the visual information in the segment, reason step-by-step to reach the answer
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 26696af3-b55b-4b4c-9e8a-e38be90fb1cf · outbound
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation abf719dc-f887-4c9f-abdb-568734f7f3b6 · outbound
DynFrame: Adaptive Reasoning-Driven Multimodal Framework with Dynamic Frame Augmentation for Complex Video Understanding Unresolved cited work
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e5c16279-b35f-4827-8bc8-6787148f2b7a · outbound
DynFrame: Adaptive Reasoning-Driven Multimodal Framework with Dynamic Frame Augmentation for Complex Video Understanding Unresolved cited work
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e0db0caf-164f-44a7-9846-382b920b2d3f · outbound
DynFrame: Adaptive Reasoning-Driven Multimodal Framework with Dynamic Frame Augmentation for Complex Video Understanding Keep the reasoning concise and non-repetitive
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0219894b-ceeb-4d60-8b6c-e4be36b5bfef · outbound
DynFrame: Adaptive Reasoning-Driven Multimodal Framework with Dynamic Frame Augmentation for Complex Video Understanding Unresolved cited work
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 84ab49aa-8ad4-4259-8c7b-1a2b2a60c609 · outbound
DynFrame: Adaptive Reasoning-Driven Multimodal Framework with Dynamic Frame Augmentation for Complex Video Understanding Max retrieval / injection
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.