Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T11:35:48.103145Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 41 of 41 outbound references and 14 inbound Pith citation observations for arXiv:2506.01908.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T11:35:48.103145Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-06T00:03:29.800596Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T16:49:57.219030Z
41 of 41 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 3e2ed415-a459-42f9-a09f-d03a53b8e48a · outbound
Reinforcement Learning Tuning for VideoLLMs: Reward Design and Data Efficiency Localizing moments in video with natural language
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5e6a877c-0211-4de8-ac75-94c96c2d6796 · outbound
Reinforcement Learning Tuning for VideoLLMs: Reward Design and Data Efficiency Qwen2.5-VL Technical Report
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eccda8be-c02a-419c-b524-25ba3ae25d04 · outbound
Reinforcement Learning Tuning for VideoLLMs: Reward Design and Data Efficiency VideoLLaMA 2: Advancing Spatial-Temporal Modeling and Audio Understanding in Video-LLMs
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8fc8c328-42a3-459d-8d07-088e5c15919c · outbound
Reinforcement Learning Tuning for VideoLLMs: Reward Design and Data Efficiency Video-R1: Reinforcing Video Reasoning in MLLMs
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8569e3b4-44da-4cba-a19a-2d97d92b075c · outbound
Reinforcement Learning Tuning for VideoLLMs: Reward Design and Data Efficiency Video-MME: The First-Ever Comprehensive Evaluation Benchmark of Multi-modal LLMs in Video Analysis
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 45f9f087-ac3c-4db6-88b4-1b9f4814e1cb · outbound
Reinforcement Learning Tuning for VideoLLMs: Reward Design and Data Efficiency Tall: Temporal activity localization via language query
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation df619bf6-970c-49cf-b621-4114a27adc8b · outbound
Reinforcement Learning Tuning for VideoLLMs: Reward Design and Data Efficiency DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d07b57b7-0ada-4364-a2cb-8ef7e1fd7278 · outbound
Reinforcement Learning Tuning for VideoLLMs: Reward Design and Data Efficiency Vtg-llm: Integrating timestamp knowledge into video llms for enhanced video temporal grounding
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2771756b-e8fa-4632-881e-95c545e86d39 · outbound
Reinforcement Learning Tuning for VideoLLMs: Reward Design and Data Efficiency VideoEspresso: A Large-Scale Chain-of-Thought Dataset for Fine-Grained Video Reasoning via Core Frame Selection
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9da80ad0-f63a-4bc9-a3db-7b3d1e21690c · outbound
Reinforcement Learning Tuning for VideoLLMs: Reward Design and Data Efficiency Vtimellm: Empower llm to grasp video moments
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 43b7121e-3cac-4838-88fd-b8361a7a9dc0 · outbound
Reinforcement Learning Tuning for VideoLLMs: Reward Design and Data Efficiency Lita: Language instructed temporal-localization assistant
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 45cfdbd5-8265-4231-a9f7-f3e8f5fd40f1 · outbound
Reinforcement Learning Tuning for VideoLLMs: Reward Design and Data Efficiency Unleashing the temporal-spatial reasoning capacity of gpt for training-free audio and language referenced video object segmentation
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3d982fbe-eb3f-4b05-8f52-62c741a52b0c · outbound
Reinforcement Learning Tuning for VideoLLMs: Reward Design and Data Efficiency Dense- captioning events in videos
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4a6e2bcb-30ea-4ec2-9188-066139becf08 · outbound
Reinforcement Learning Tuning for VideoLLMs: Reward Design and Data Efficiency LLaVA-OneVision: Easy Visual Task Transfer
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f79c5a1c-78e4-4915-8011-72b2ac8f5279 · outbound
Reinforcement Learning Tuning for VideoLLMs: Reward Design and Data Efficiency LLaVA-ST: A Multimodal Large Language Model for Fine-Grained Spatial-Temporal Understanding
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f4f70289-08d2-46c8-ae70-8bcb2c2b402d · outbound
Reinforcement Learning Tuning for VideoLLMs: Reward Design and Data Efficiency VideoChat: Chat-Centric Video Understanding
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ea610f96-09fd-40b8-9665-c8b5486b81d1 · outbound
Reinforcement Learning Tuning for VideoLLMs: Reward Design and Data Efficiency Mvbench: A comprehensive multi-modal video understanding benchmark
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4a53e36c-eb1c-4e6f-8fb3-3e6f8f8694de · outbound
Reinforcement Learning Tuning for VideoLLMs: Reward Design and Data Efficiency Llama-vid: An image is worth 2 tokens in large language models
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c7ba70d3-efd9-41c4-bb52-345f72de1a67 · outbound
Reinforcement Learning Tuning for VideoLLMs: Reward Design and Data Efficiency Kangaroo: A Powerful Video-Language Model Supporting Long-context Video Input
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 79e47c18-6a64-4be8-92c6-176f2ac99756 · outbound
Reinforcement Learning Tuning for VideoLLMs: Reward Design and Data Efficiency St-llm: Large language models are effective temporal learners
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 56dba7d3-10b7-4430-87a0-afb694e0234a · outbound
Reinforcement Learning Tuning for VideoLLMs: Reward Design and Data Efficiency TempCompass: Do Video LLMs Really Understand Videos?
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 123f7ac3-f57a-4fee-8903-251e01aa0edd · outbound
Reinforcement Learning Tuning for VideoLLMs: Reward Design and Data Efficiency Visual-RFT: Visual Reinforcement Fine-Tuning
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 573628b1-4839-4208-820e-ad9f759dbedb · outbound
Reinforcement Learning Tuning for VideoLLMs: Reward Design and Data Efficiency Perception test: A diagnostic benchmark for multimodal video models.Advances in Neural Information Processing Systems, 36:42748–42761, 2023
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a10982dd-14f0-48e5-a377-08aea53acc44 · outbound
Reinforcement Learning Tuning for VideoLLMs: Reward Design and Data Efficiency Momentor: Advancing Video Large Language Model with Fine-Grained Temporal Reasoning
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2e07006d-1dbc-4672-bfbc-14f921321bc9 · outbound
Reinforcement Learning Tuning for VideoLLMs: Reward Design and Data Efficiency Timechat: A time-sensitive multimodal large language model for long video understanding
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 886ba38e-c3c1-4fe8-ae62-1a6c3912c6d6 · outbound
Reinforcement Learning Tuning for VideoLLMs: Reward Design and Data Efficiency DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dc1fead3-176b-4331-8547-5c910e8d96d6 · outbound
Reinforcement Learning Tuning for VideoLLMs: Reward Design and Data Efficiency VLM-R1: A Stable and Generalizable R1-style Large Vision-Language Model
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ac4dc91a-c40b-4533-9d0e-c72615597896 · outbound
Reinforcement Learning Tuning for VideoLLMs: Reward Design and Data Efficiency Kimi k1.5: Scaling Reinforcement Learning with LLMs
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 921a662e-7183-44a5-b656-983f70e2ff53 · outbound
Reinforcement Learning Tuning for VideoLLMs: Reward Design and Data Efficiency Grounded-VideoLLM: Sharpening Fine-grained Temporal Grounding in Video Large Language Models
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c8498e34-385e-46d5-b571-f9cbbe09859c · outbound
Reinforcement Learning Tuning for VideoLLMs: Reward Design and Data Efficiency InternVideo2.5: Empowering Video MLLMs with Long and Rich Context Modeling
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 85ce6aae-bad9-42a4-9987-1de562e8821c · outbound
Reinforcement Learning Tuning for VideoLLMs: Reward Design and Data Efficiency HawkEye: Training Video-Text LLMs for Grounding Text in Videos
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e97de0ce-7f61-4652-b68c-917365ddae96 · outbound
Reinforcement Learning Tuning for VideoLLMs: Reward Design and Data Efficiency Next-qa: Next phase of question- answering to explaining temporal actions
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 295dcf46-87a4-4baf-b49c-c5725e07db3d · outbound
Reinforcement Learning Tuning for VideoLLMs: Reward Design and Data Efficiency Can i trust your answer? visually grounded video question answering
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0bb155bc-f509-48e6-9723-679ee5d23c62 · outbound
Reinforcement Learning Tuning for VideoLLMs: Reward Design and Data Efficiency mPLUG-Owl3: Towards Long Image-Sequence Understanding in Multi-Modal Large Language Models
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8493c7c0-257f-4882-ae9a-c2fa44dea18b · outbound
Reinforcement Learning Tuning for VideoLLMs: Reward Design and Data Efficiency Unhackable Temporal Rewarding for Scalable Video MLLMs
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 997cf40a-d098-4e74-9803-f8343e4597cb · outbound
Reinforcement Learning Tuning for VideoLLMs: Reward Design and Data Efficiency TimeSuite: Improving MLLMs for Long Video Understanding via Grounded Tuning
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b6525d68-4675-45d3-ad5f-a6f602a39e3a · outbound
Reinforcement Learning Tuning for VideoLLMs: Reward Design and Data Efficiency Vision-R1: Evolving Human-Free Alignment in Large Vision-Language Models via Vision-Guided Reinforcement Learning
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 03c4b6b9-57e9-42a5-ae6d-9513ed906ae0 · outbound
Reinforcement Learning Tuning for VideoLLMs: Reward Design and Data Efficiency R1-VL: Learning to Reason with Multimodal Large Language Models via Step-wise Group Relative Policy Optimization
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cc55fd6e-4219-447b-b325-f6fe6f62535c · outbound
Reinforcement Learning Tuning for VideoLLMs: Reward Design and Data Efficiency Long Context Transfer from Language to Vision
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 81a57794-523a-4f6d-aa33-8a70e34563d8 · outbound
Reinforcement Learning Tuning for VideoLLMs: Reward Design and Data Efficiency MMVU: Measuring Expert-Level Multi-Discipline Video Understanding
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 68083eb2-8d87-452c-9426-e39db36d209f · outbound
Reinforcement Learning Tuning for VideoLLMs: Reward Design and Data Efficiency R1-Zero's "Aha Moment" in Visual Reasoning on a 2B Non-SFT Model
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c2f4e65e-d940-42b8-8d43-1f68641fb498 · inbound
Empowering Nanoscale Connectivity through Molecular Communication: A Case Study of Virus Infection Reinforcement Learning Tuning for VideoLLMs: Reward Design and Data Efficiency
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9ab09eb4-570a-404a-9b79-903c34a99536 · inbound
TAR: Temporal Anchor-Constrained Reasoning for Video Temporal Grounding Reinforcement Learning Tuning for VideoLLMs: Reward Design and Data Efficiency
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 050345ff-e308-45d7-9f82-44827a2917c7 · inbound
OneThinker: All-in-one Reasoning Model for Image and Video Reinforcement Learning Tuning for VideoLLMs: Reward Design and Data Efficiency
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5d6fde9d-8de0-49ff-b154-7fd71dcbe877 · inbound
AdaTooler-V: Adaptive Tool-Use for Images and Videos Reinforcement Learning Tuning for VideoLLMs: Reward Design and Data Efficiency
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d62ebeaa-5f0d-4f4c-bb7e-d320650c0026 · inbound
GraphThinker: Reinforcing Temporally Grounded Video Reasoning with Event Graph Thinking Reinforcement Learning Tuning for VideoLLMs: Reward Design and Data Efficiency
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5f329e02-4e84-42af-94a8-3b9914d312ba · inbound
DynFrame: Adaptive Reasoning-Driven Multimodal Framework with Dynamic Frame Augmentation for Complex Video Understanding Reinforcement Learning Tuning for VideoLLMs: Reward Design and Data Efficiency
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2673ee9c-ee9e-4103-a9ff-3e0560d47143 · inbound
Temporal-Aware Reasoning Optimization for Video Temporal Grounding Reinforcement Learning Tuning for VideoLLMs: Reward Design and Data Efficiency
Reference 122
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 78b1981f-1631-4e84-8e43-db819365475a · inbound
Reasoning as Intersection: Consensus-Frame Alignment for Visual Focus in Video-MLLMs Reinforcement Learning Tuning for VideoLLMs: Reward Design and Data Efficiency
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5fe7afff-ff8f-4d04-bc81-d73212b62804 · inbound
CARE: Competence-Aware Reward Shaping for Adaptive Reasoning Length in Video-MLLMs Reinforcement Learning Tuning for VideoLLMs: Reward Design and Data Efficiency
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 17baca14-be29-4760-8abd-79b8b54c4837 · inbound
DramaDirector: Geometry-Guided Short Drama Generation Reinforcement Learning Tuning for VideoLLMs: Reward Design and Data Efficiency
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c728d01b-29e6-4990-bdbd-fdb58bd8a711 · inbound
SER: Learning to Ground Video Reasoning with Semantic Evidence Rewards Reinforcement Learning Tuning for VideoLLMs: Reward Design and Data Efficiency
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2ba040e6-b12e-4f6b-a703-ede10224c1a3 · inbound
Video-MME-Logical: A Controlled Diagnostic Benchmark for Video Temporal-Logical Reasoning Reinforcement Learning Tuning for VideoLLMs: Reward Design and Data Efficiency
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2c46cb5e-b4d1-46c7-a68f-ac8bc4892e0c · inbound
TimeLens2: Generalist Video Temporal Grounding with Multimodal LLMs Reinforcement Learning Tuning for VideoLLMs: Reward Design and Data Efficiency
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1ed0b2c7-1ed2-42c3-be13-e58a000da904 · inbound
AdaThinkV: Adaptive Thinking for Token-Efficient Video Reasoning Reinforcement Learning Tuning for VideoLLMs: Reward Design and Data Efficiency
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.