Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-17T02:18:21.718091Z
Paper Citation Record · LEDGER
As of 4 August 2026, this Paper Citation Record lists 72 of 72 outbound references and 3 inbound Pith citation observations for arXiv:2512.03963.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-17T02:18:21.718091Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-03T06:30:56.289259+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-03T04:17:27.916261Z
A source-named dated measurement, never combined with another source.
Source: cited_works
72 of 72 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 9e924d67-296f-4bc0-97c4-95c6e387773f · outbound
TempR1: Improving Temporal Understanding of MLLMs via Temporal-Aware Multi-Task Reinforcement Learning GPT-4 Technical Report
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation ca102fe4-0294-4bc1-9fbe-3246bf0dcebe · outbound
TempR1: Improving Temporal Understanding of MLLMs via Temporal-Aware Multi-Task Reinforcement Learning Localizing mo- ments in video with natural language
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation c2da755f-03bf-48c4-8594-4bb41a80e636 · outbound
TempR1: Improving Temporal Understanding of MLLMs via Temporal-Aware Multi-Task Reinforcement Learning Qwen2.5-VL Technical Report
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 2d14289e-c51c-4387-99ec-71096cad233b · outbound
TempR1: Improving Temporal Understanding of MLLMs via Temporal-Aware Multi-Task Reinforcement Learning UniVG-R1: Reasoning Guided Universal Visual Grounding with Reinforcement Learning
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 23aaa48a-bfdf-4790-98c1-42b278245487 · outbound
TempR1: Improving Temporal Understanding of MLLMs via Temporal-Aware Multi-Task Reinforcement Learning Dense events grounding in video
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation c6fec625-18a3-4d76-9307-8443d9984044 · outbound
TempR1: Improving Temporal Understanding of MLLMs via Temporal-Aware Multi-Task Reinforcement Learning Activitynet: A large-scale video benchmark for human activity understanding
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 6a65651b-f8ad-4a84-80c0-e377032b98ea · outbound
TempR1: Improving Temporal Understanding of MLLMs via Temporal-Aware Multi-Task Reinforcement Learning Flashvtg: Feature layering and adaptive score handling network for video temporal grounding
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation cd5c16b1-0bb6-4f72-b4ca-8e53382e30c7 · outbound
TempR1: Improving Temporal Understanding of MLLMs via Temporal-Aware Multi-Task Reinforcement Learning Scaling rl to long videos
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 42e182a2-3342-41fe-9686-0d1cf0474f6d · outbound
TempR1: Improving Temporal Understanding of MLLMs via Temporal-Aware Multi-Task Reinforcement Learning VisRL: Intention-Driven Visual Perception via Reinforced Reasoning
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 1b69e8a0-90a9-4dce-8cd0-419b334bec29 · outbound
TempR1: Improving Temporal Understanding of MLLMs via Temporal-Aware Multi-Task Reinforcement Learning Video-R1: Reinforcing Video Reasoning in MLLMs
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation ebccfd11-141e-4040-af8c-0d4fabdf6973 · outbound
TempR1: Improving Temporal Understanding of MLLMs via Temporal-Aware Multi-Task Reinforcement Learning Video-mme: The first-ever comprehensive evaluation benchmark of multi-modal llms in video analysis
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 68b7ba79-714e-4aeb-92c4-ee1a6901d77f · outbound
TempR1: Improving Temporal Understanding of MLLMs via Temporal-Aware Multi-Task Reinforcement Learning Tall: Temporal activity localization via language query
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation ecea151b-2de8-4a09-91a9-e993788542b7 · outbound
TempR1: Improving Temporal Understanding of MLLMs via Temporal-Aware Multi-Task Reinforcement Learning TAR: Temporal Anchor-Constrained Reasoning for Video Temporal Grounding
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 520d7b46-2d7f-4254-a6a2-34761e1c5191 · outbound
TempR1: Improving Temporal Understanding of MLLMs via Temporal-Aware Multi-Task Reinforcement Learning DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 169e689b-ea81-47b2-ae35-2c795d170f3c · outbound
TempR1: Improving Temporal Understanding of MLLMs via Temporal-Aware Multi-Task Reinforcement Learning TRACE: Temporal Grounding Video LLM via Causal Event Modeling
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation c545c391-4de8-4bf9-9af6-5814e52614dd · outbound
TempR1: Improving Temporal Understanding of MLLMs via Temporal-Aware Multi-Task Reinforcement Learning Vtimellm: Empower llm to grasp video moments
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation a01b7b66-7d29-4ad0-94ec-e21dfb402f7b · outbound
TempR1: Improving Temporal Understanding of MLLMs via Temporal-Aware Multi-Task Reinforcement Learning Vision-R1: Incentivizing Reasoning Capability in Multimodal Large Language Models
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 9807d542-83f1-4aec-9b6c-a81ac8618563 · outbound
TempR1: Improving Temporal Understanding of MLLMs via Temporal-Aware Multi-Task Reinforcement Learning Online Video Understanding: OVBench and VideoChat-Online
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 4623874d-a6f0-4582-b3d5-c36b164e0407 · outbound
TempR1: Improving Temporal Understanding of MLLMs via Temporal-Aware Multi-Task Reinforcement Learning in the wild
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation f2974a46-ed9a-4530-a305-09ee1de41461 · outbound
TempR1: Improving Temporal Understanding of MLLMs via Temporal-Aware Multi-Task Reinforcement Learning OpenAI o1 System Card
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation a1bfd47c-1aea-4da4-a8d9-9ffb80f8ef83 · outbound
TempR1: Improving Temporal Understanding of MLLMs via Temporal-Aware Multi-Task Reinforcement Learning Knowing where to focus: Event-aware transformer for video grounding
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation de5e88f1-43a6-4c72-9611-8fbe04d39583 · outbound
TempR1: Improving Temporal Understanding of MLLMs via Temporal-Aware Multi-Task Reinforcement Learning Dense-captioning events in videos
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 6c2f1939-f828-48b5-941a-67ca5bde11a3 · outbound
TempR1: Improving Temporal Understanding of MLLMs via Temporal-Aware Multi-Task Reinforcement Learning Detecting mo- ments and highlights in videos via natural language queries
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation c12dd039-c2f9-4ce8-818e-28bf6ffea39e · outbound
TempR1: Improving Temporal Understanding of MLLMs via Temporal-Aware Multi-Task Reinforcement Learning LLaVA-OneVision: Easy Visual Task Transfer
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation b52f970c-bf0a-4400-9c52-e9db42d95dfe · outbound
TempR1: Improving Temporal Understanding of MLLMs via Temporal-Aware Multi-Task Reinforcement Learning Unmasked teacher: Towards training-efficient video foundation models
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 01f84cc9-2158-465c-85cd-b87dd00ccf88 · outbound
TempR1: Improving Temporal Understanding of MLLMs via Temporal-Aware Multi-Task Reinforcement Learning Mo- mentdiff: Generative video moment retrieval from random to real.Advances in neural information processing systems, 36:65948–65966
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 272ecf85-534d-471f-9d58-5b4347c00ece · outbound
TempR1: Improving Temporal Understanding of MLLMs via Temporal-Aware Multi-Task Reinforcement Learning VideoChat-R1: Enhancing Spatio-Temporal Perception via Reinforcement Fine-Tuning
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 93ac2896-e900-4888-a00a-4a25f5a5458d · outbound
TempR1: Improving Temporal Understanding of MLLMs via Temporal-Aware Multi-Task Reinforcement Learning Ground- inggpt: Language enhanced multi-modal grounding model
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation d7c17fae-3323-4fa8-bc27-438b4c46233a · outbound
TempR1: Improving Temporal Understanding of MLLMs via Temporal-Aware Multi-Task Reinforcement Learning Video-llava: Learning united visual repre- sentation by alignment before projection
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation e56712fc-fcc8-4a34-96fa-f3204fd3b467 · outbound
TempR1: Improving Temporal Understanding of MLLMs via Temporal-Aware Multi-Task Reinforcement Learning Univtg: Towards unified video- language temporal grounding
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 681f3c37-9bc1-4d10-bcc4-050b33aa5877 · outbound
TempR1: Improving Temporal Understanding of MLLMs via Temporal-Aware Multi-Task Reinforcement Learning Visual instruction tuning.Advances in neural information processing systems, 36:34892–34916
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation b8c9c9b0-1b9b-45fe-920c-371565cd023a · outbound
TempR1: Improving Temporal Understanding of MLLMs via Temporal-Aware Multi-Task Reinforcement Learning Umt: Unified multi-modal transformers for joint video moment retrieval and highlight detection
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 672cd8b6-56a2-4a32-a61b-78fd726cba21 · outbound
TempR1: Improving Temporal Understanding of MLLMs via Temporal-Aware Multi-Task Reinforcement Learning r 2-tuning: Ef- ficient image-to-video transfer learning for video temporal grounding
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 28b801d3-bcaa-4514-b818-0c7cdf0e3eab · outbound
TempR1: Improving Temporal Understanding of MLLMs via Temporal-Aware Multi-Task Reinforcement Learning TempCompass: Do Video LLMs Really Understand Videos?
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 442ec364-5fa3-4d06-918a-45a016735ab8 · outbound
TempR1: Improving Temporal Understanding of MLLMs via Temporal-Aware Multi-Task Reinforcement Learning Unresolved cited work
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 4f370744-bb59-4f16-9eb1-72d9333d5167 · outbound
TempR1: Improving Temporal Understanding of MLLMs via Temporal-Aware Multi-Task Reinforcement Learning Visual-RFT: Visual Reinforcement Fine-Tuning
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation a810bf84-4b8e-4662-a534-5e19cae1e3c7 · outbound
TempR1: Improving Temporal Understanding of MLLMs via Temporal-Aware Multi-Task Reinforcement Learning MUSEG: Reinforcing Video Temporal Understanding via Timestamp-Aware Multi-Segment Grounding
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation ce32cea4-58d5-429f-acab-add86dd88f39 · outbound
TempR1: Improving Temporal Understanding of MLLMs via Temporal-Aware Multi-Task Reinforcement Learning Correlation-guided query-dependency calibration in video representation learning for temporal grounding.CoRR
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation a2f9eeaa-84bd-46d6-ac96-24c143f27db4 · outbound
TempR1: Improving Temporal Understanding of MLLMs via Temporal-Aware Multi-Task Reinforcement Learning Query-dependent video representa- tion for moment retrieval and highlight detection
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 9dc16043-078a-4a3c-8d20-3d1239b663b5 · outbound
TempR1: Improving Temporal Understanding of MLLMs via Temporal-Aware Multi-Task Reinforcement Learning SpaceR: Reinforcing MLLMs in Video Spatial Reasoning
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 1a5daa42-d2f4-4fb3-af9d-0cf2f6326335 · outbound
TempR1: Improving Temporal Understanding of MLLMs via Temporal-Aware Multi-Task Reinforcement Learning Per- ception test: A diagnostic benchmark for multimodal video models.Advances in Neural Information Processing Sys- tems, 36:42748–42761
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 11d2dfa6-32f8-47de-a608-34136f45a500 · outbound
TempR1: Improving Temporal Understanding of MLLMs via Temporal-Aware Multi-Task Reinforcement Learning Chatvtg: Video temporal grounding via chat with video dialogue large language models
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 59f2e683-7596-40d4-8c8c-769b76484e69 · outbound
TempR1: Improving Temporal Understanding of MLLMs via Temporal-Aware Multi-Task Reinforcement Learning Learning transferable visual models from natural language supervi- sion
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 6dbe3818-759c-4abe-ae7f-eb186a401164 · outbound
TempR1: Improving Temporal Understanding of MLLMs via Temporal-Aware Multi-Task Reinforcement Learning Timechat: A time-sensitive multimodal large lan- guage model for long video understanding
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 4f603076-e551-4285-867f-d759b3d35f27 · outbound
TempR1: Improving Temporal Understanding of MLLMs via Temporal-Aware Multi-Task Reinforcement Learning DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 6d828343-faca-4796-998e-a3172586f1a0 · outbound
TempR1: Improving Temporal Understanding of MLLMs via Temporal-Aware Multi-Task Reinforcement Learning VLM-R1: A Stable and Generalizable R1-style Large Vision-Language Model
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation e4818a9e-30d2-4592-9db8-3dd509e5649c · outbound
TempR1: Improving Temporal Understanding of MLLMs via Temporal-Aware Multi-Task Reinforcement Learning End-to-end dense video grounding via parallel regression.Computer Vi- sion and Image Understanding, 242:103980
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 1459d225-e80c-4a46-bab7-4e39c11f7548 · outbound
TempR1: Improving Temporal Understanding of MLLMs via Temporal-Aware Multi-Task Reinforcement Learning Moviechat: From dense token to sparse memory for long video understanding
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 7e7cee7c-7670-4ba0-9e21-e339cd5aec3f · outbound
TempR1: Improving Temporal Understanding of MLLMs via Temporal-Aware Multi-Task Reinforcement Learning Tr- detr: Task-reciprocal transformer for joint moment retrieval and highlight detection
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation f233cbca-53fe-4135-bf8b-a2ad4156c04e · outbound
TempR1: Improving Temporal Understanding of MLLMs via Temporal-Aware Multi-Task Reinforcement Learning Hierarchical semantic correspondence net- works for video paragraph grounding
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 15dd6e88-c57e-4c4c-837a-008464009a0f · outbound
TempR1: Improving Temporal Understanding of MLLMs via Temporal-Aware Multi-Task Reinforcement Learning Hierarchical semantic correspondence net- works for video paragraph grounding
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation dc84d14d-1295-446c-b300-38fb62d65e0f · outbound
TempR1: Improving Temporal Understanding of MLLMs via Temporal-Aware Multi-Task Reinforcement Learning Tspo: Temporal sampling policy optimization for long- form video language understanding
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 862cdbb2-b580-43cb-b6df-242a8ec5a795 · outbound
TempR1: Improving Temporal Understanding of MLLMs via Temporal-Aware Multi-Task Reinforcement Learning Gemini: A Family of Highly Capable Multimodal Models
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation c6333379-37db-4fd4-85b4-7502732d183e · outbound
TempR1: Improving Temporal Understanding of MLLMs via Temporal-Aware Multi-Task Reinforcement Learning Videorft: Incentivizing video reasoning capability in mllms via reinforced fine-tuning
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation a8c5d2f2-9c0e-49b3-b0e0-b42e0379d4fd · outbound
TempR1: Improving Temporal Understanding of MLLMs via Temporal-Aware Multi-Task Reinforcement Learning InternVideo: General Video Foundation Models via Generative and Discriminative Learning
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation f8404e5c-5d49-4305-b484-126752c65f39 · outbound
TempR1: Improving Temporal Understanding of MLLMs via Temporal-Aware Multi-Task Reinforcement Learning Internvideo2: Scaling foundation models for mul- timodal video understanding
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation c2633b38-8492-4fae-b0b6-62c18d264932 · outbound
TempR1: Improving Temporal Understanding of MLLMs via Temporal-Aware Multi-Task Reinforcement Learning HawkEye: Training Video-Text LLMs for Grounding Text in Videos
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 5de38293-abb8-4a48-8d1d-41bbe9413f47 · outbound
TempR1: Improving Temporal Understanding of MLLMs via Temporal-Aware Multi-Task Reinforcement Learning Efficient Temporal Extrapolation of Multimodal Large Language Models with Temporal Grounding Bridge
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation fb327704-a237-408e-84ee-7a14cf52e717 · outbound
TempR1: Improving Temporal Understanding of MLLMs via Temporal-Aware Multi-Task Reinforcement Learning Time-R1: Post-Training Large Vision Language Model for Temporal Video Grounding
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 1a187403-cbb5-4c2f-8eba-367d2645778b · outbound
TempR1: Improving Temporal Understanding of MLLMs via Temporal-Aware Multi-Task Reinforcement Learning Visionary-r1: Mitigating shortcuts in visual reasoning with reinforcement learning
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 6a123d00-0d84-4dba-92cd-0ef2c475cd52 · outbound
TempR1: Improving Temporal Understanding of MLLMs via Temporal-Aware Multi-Task Reinforcement Learning Can i trust your answer? visually grounded video question answering
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 1b539996-056c-4907-9e4c-50a5e54ba3c3 · outbound
TempR1: Improving Temporal Understanding of MLLMs via Temporal-Aware Multi-Task Reinforcement Learning Bridging the gap: A unified video comprehension framework for mo- ment retrieval and highlight detection
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation a30d6193-0b77-4a33-8a28-fdcf02fa13f9 · outbound
TempR1: Improving Temporal Understanding of MLLMs via Temporal-Aware Multi-Task Reinforcement Learning VideoCLIP: Contrastive Pre-training for Zero-shot Video-Text Understanding
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation dbb608f1-cb63-4ed3-ad90-f84c4bacb0c0 · outbound
TempR1: Improving Temporal Understanding of MLLMs via Temporal-Aware Multi-Task Reinforcement Learning Videochat-r1
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 067c528d-c020-458e-8a3a-3237b3e1d735 · outbound
TempR1: Improving Temporal Understanding of MLLMs via Temporal-Aware Multi-Task Reinforcement Learning DAPO: An Open-Source LLM Reinforcement Learning System at Scale
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation e0fb155f-e300-4ee5-a867-e80b8a875f35 · outbound
TempR1: Improving Temporal Understanding of MLLMs via Temporal-Aware Multi-Task Reinforcement Learning TimeSuite: Improving MLLMs for Long Video Understanding via Grounded Tuning
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 3b27a866-be10-47e8-90f6-c954341edf38 · outbound
TempR1: Improving Temporal Understanding of MLLMs via Temporal-Aware Multi-Task Reinforcement Learning Video-LLaMA: An Instruction-tuned Audio-Visual Language Model for Video Understanding
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation be23abb0-5110-4e32-841d-bccc4487b546 · outbound
TempR1: Improving Temporal Understanding of MLLMs via Temporal-Aware Multi-Task Reinforcement Learning Sc-captioner: Improving image captioning with self- correction by reinforcement learning
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 7e940b54-6378-4f98-b9df-28dc33c57448 · outbound
TempR1: Improving Temporal Understanding of MLLMs via Temporal-Aware Multi-Task Reinforcement Learning TinyLLaVA-Video-R1: Towards Smaller LMMs for Video Reasoning
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 6ff04040-ff6f-49d2-9d95-6d3bd7fc1f47 · outbound
TempR1: Improving Temporal Understanding of MLLMs via Temporal-Aware Multi-Task Reinforcement Learning Hacs: Human action clips and segments dataset for recognition and temporal localization
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation b5e099b8-c0de-49f0-bddb-704e8764274b · outbound
TempR1: Improving Temporal Understanding of MLLMs via Temporal-Aware Multi-Task Reinforcement Learning Group Sequence Policy Optimization
Reference 72
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 72669cb1-ea7a-4f24-8e64-5a80eb90dc0f · outbound
TempR1: Improving Temporal Understanding of MLLMs via Temporal-Aware Multi-Task Reinforcement Learning Rethinking the video sampling and reasoning strategies for temporal sentence grounding
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 38fb59fc-0b9b-4067-9d2c-c447183f7ddc · inbound
Multi-Task GRPO: Reliable LLM Reasoning Across Tasks TempR1: Improving Temporal Understanding of MLLMs via Temporal-Aware Multi-Task Reinforcement Learning
Reference 73
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ce54db9b-5438-4540-9a1b-cbffd719bea6 · inbound
DELTAVID: Enhancing Fine-Grained Spatiotemporal Perception with Cross-Video Differences TempR1: Improving Temporal Understanding of MLLMs via Temporal-Aware Multi-Task Reinforcement Learning
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 31e83b57-04d5-45c2-93b1-b4bffa41c296 · inbound
TimeLens2: Generalist Video Temporal Grounding with Multimodal LLMs TempR1: Improving Temporal Understanding of MLLMs via Temporal-Aware Multi-Task Reinforcement Learning
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.