Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T05:30:37.955324Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 65 of 65 outbound references and 2 inbound Pith citation observations for arXiv:2508.01699.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T05:30:37.955324Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-07-14T19:27:29.843866Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-11T19:06:08.881592Z
65 of 65 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 623f6c64-f604-4f32-8938-a8446102c4c8 · outbound
TimeExpert: An Expert-Guided Video LLM for Video Temporal Grounding Meteor: An automatic metric for mt evaluation with improved correlation with hu- man judgments
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 05c00db8-871c-4775-a3a2-3c3ce4dec1b0 · outbound
TimeExpert: An Expert-Guided Video LLM for Video Temporal Grounding Activitynet: A large-scale video benchmark for human activity understanding
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c78fdb2e-00a1-4b18-bd51-4781f388e678 · outbound
TimeExpert: An Expert-Guided Video LLM for Video Temporal Grounding Sharegpt4video: Improving video understand- ing and generation with better captions
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d594ea80-d3fa-48c4-b294-be2a229abc9a · outbound
TimeExpert: An Expert-Guided Video LLM for Video Temporal Grounding Vast: A vision-audio-subtitle-text omni-modality foundation model and dataset
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3ef79b53-06a9-4b21-9f6c-85d5dd158b02 · outbound
TimeExpert: An Expert-Guided Video LLM for Video Temporal Grounding VideoLLaMA 2: Advancing Spatial-Temporal Modeling and Audio Understanding in Video-LLMs
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b9a04bde-8af6-4c7d-b103-ceec727b8e87 · outbound
TimeExpert: An Expert-Guided Video LLM for Video Temporal Grounding Uni- fied scaling laws for routed language models
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1a072bc1-35a1-4940-a54e-7d53b065ba84 · outbound
TimeExpert: An Expert-Guided Video LLM for Video Temporal Grounding An image is worth 16x16 words: Trans- formers for image recognition at scale
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f6ea2427-d8b2-44f8-9eec-6ee7fca202fc · outbound
TimeExpert: An Expert-Guided Video LLM for Video Temporal Grounding Learning Factored Representations in a Deep Mixture of Experts
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c0c05667-94c4-484c-95d8-f322f64d16cc · outbound
TimeExpert: An Expert-Guided Video LLM for Video Temporal Grounding To- wards an empirical understanding of moe design choices
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 157887cb-5000-4b81-8b97-0f9409a752e7 · outbound
TimeExpert: An Expert-Guided Video LLM for Video Temporal Grounding Switch transformers: Scaling to trillion parameter models with sim- ple and efficient sparsity
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1113bebe-bf15-4fd8-894d-e77ba17abd5c · outbound
TimeExpert: An Expert-Guided Video LLM for Video Temporal Grounding Video-mme: The first-ever comprehensive evaluation benchmark of multi-modal llms in video analysis
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3df24c47-2fca-46c6-abe5-2bc47edf56da · outbound
TimeExpert: An Expert-Guided Video LLM for Video Temporal Grounding Soda: Story oriented dense video captioning evaluation framework
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 605203f0-9526-4c6f-9f37-717c149fb719 · outbound
TimeExpert: An Expert-Guided Video LLM for Video Temporal Grounding Tall: Temporal activity localization via language query
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 093ce128-28cb-4f61-8c89-8d03aad676a0 · outbound
TimeExpert: An Expert-Guided Video LLM for Video Temporal Grounding Dynamic mixture of experts: An auto- tuning approach for efficient transformer models
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6889abd9-af84-4ed9-bb06-8b189a4ebf88 · outbound
TimeExpert: An Expert-Guided Video LLM for Video Temporal Grounding Vtg-llm: Integrating timestamp knowledge into video llms for enhanced video temporal grounding
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a1f15d89-1ff6-4c8b-871f-0de6772c6963 · outbound
TimeExpert: An Expert-Guided Video LLM for Video Temporal Grounding Trace: Temporal grounding video llm via causal event modeling
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 131fdcac-9151-4a86-b262-05f56902c143 · outbound
TimeExpert: An Expert-Guided Video LLM for Video Temporal Grounding Creating summaries from user videos
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 17d0bea3-7751-402c-b7f0-41c96bcaaf06 · outbound
TimeExpert: An Expert-Guided Video LLM for Video Temporal Grounding Unleash the potential of clip for video highlight detection
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation eb98ddcc-bbff-493f-8710-4580b8875508 · outbound
TimeExpert: An Expert-Guided Video LLM for Video Temporal Grounding Vtimellm: Empower llm to grasp video moments
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2f14ab98-16ba-440c-ac4c-2eae1cc543bd · outbound
TimeExpert: An Expert-Guided Video LLM for Video Temporal Grounding Lita: Language instructed temporal-localization assistant
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ea59856b-3465-421e-893d-13195462547c · outbound
TimeExpert: An Expert-Guided Video LLM for Video Temporal Grounding Harder tasks need more experts: Dynamic routing in moe models
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2b6eb25a-64d9-42c0-b027-0f29de08ed76 · outbound
TimeExpert: An Expert-Guided Video LLM for Video Temporal Grounding Do you remember? dense video captioning with cross-modal memory retrieval
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 087627ca-d238-4b66-b779-31be644de0bd · outbound
TimeExpert: An Expert-Guided Video LLM for Video Temporal Grounding Dense-captioning events in videos
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0d490c18-b50d-4187-83f2-910cd91a7aa5 · outbound
TimeExpert: An Expert-Guided Video LLM for Video Temporal Grounding Detecting mo- ments and highlights in videos via natural language queries
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 25639214-899c-434e-956e-1357c905573c · outbound
TimeExpert: An Expert-Guided Video LLM for Video Temporal Grounding Aria: An Open Multimodal Native Mixture-of-Experts Model
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f1aa857f-33c9-40a7-b022-50c49599aa1f · outbound
TimeExpert: An Expert-Guided Video LLM for Video Temporal Grounding Unmasked teacher: Towards training-efficient video foundation models
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 17249dd6-7a2b-4b77-bdc5-01035a740a7d · outbound
TimeExpert: An Expert-Guided Video LLM for Video Temporal Grounding Mvbench: A comprehensive multi-modal video understand- ing benchmark
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 38d5f266-c7c9-4f71-a086-1d79f9b2507a · outbound
TimeExpert: An Expert-Guided Video LLM for Video Temporal Grounding Uni- moe: Scaling unified multimodal llms with mixture of ex- perts
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0bf732be-fc47-43d5-94dc-a48d507c7e89 · outbound
TimeExpert: An Expert-Guided Video LLM for Video Temporal Grounding Video-llava: Learning united visual repre- sentation by alignment before projection
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 74721bab-b9d3-4406-83fb-14975db5b8d5 · outbound
TimeExpert: An Expert-Guided Video LLM for Video Temporal Grounding Univtg: Towards unified video- language temporal grounding
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3919a88f-713a-407b-b5e3-5031f14b8e5e · outbound
TimeExpert: An Expert-Guided Video LLM for Video Temporal Grounding Visual instruction tuning
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6fa5e354-bf29-4290-84af-8d37a5a3f3be · outbound
TimeExpert: An Expert-Guided Video LLM for Video Temporal Grounding Umt: Unified multi-modal transformers for joint video moment retrieval and highlight detection
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f21d3b11-8dd8-4f2a-8183-1df5bf4165df · outbound
TimeExpert: An Expert-Guided Video LLM for Video Temporal Grounding Valley: Video Assistant with Large Language model Enhanced abilitY
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4f7a096b-b4d3-44bb-8fae-6767e6fa3049 · outbound
TimeExpert: An Expert-Guided Video LLM for Video Temporal Grounding Video-chatgpt: Towards detailed video understanding via large vision and language models
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation aad3652a-5708-4e5d-8c17-6efe370f64ca · outbound
TimeExpert: An Expert-Guided Video LLM for Video Temporal Grounding Correlation-Guided Query-Dependency Calibration for Video Temporal Grounding
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 946c2b1c-897d-422a-a44e-bf2b23a11a99 · outbound
TimeExpert: An Expert-Guided Video LLM for Video Temporal Grounding Query-dependent video representa- tion for moment retrieval and highlight detection
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 951b66dd-ded4-46e1-b914-0e7ae09582af · outbound
TimeExpert: An Expert-Guided Video LLM for Video Temporal Grounding En- coding and controlling global semantics for long-form video question answering
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5d933afe-c99b-4224-ba99-330958abac2c · outbound
TimeExpert: An Expert-Guided Video LLM for Video Temporal Grounding Queryd: A video dataset with high-quality text and audio narrations
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d99fb04b-5f3b-4521-9298-0da608427cec · outbound
TimeExpert: An Expert-Guided Video LLM for Video Temporal Grounding Momen- tor: Advancing video large language model with fine-grained temporal reasoning
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5147b790-31ef-4d86-aa40-411c7b036680 · outbound
TimeExpert: An Expert-Guided Video LLM for Video Temporal Grounding Timechat: A time-sensitive multimodal large language model for long video understanding
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d5300c57-1ae8-49f0-8ef3-53426d8d5c33 · outbound
TimeExpert: An Expert-Guided Video LLM for Video Temporal Grounding Outra- geously large neural networks: The sparsely-gated mixture- of-experts layer
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e0ed41b6-1d91-4256-9b4f-85882b0e8c6e · outbound
TimeExpert: An Expert-Guided Video LLM for Video Temporal Grounding Tvsum: Summarizing web videos using titles
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation fe940160-b914-400f-ab09-e401a411b2b2 · outbound
TimeExpert: An Expert-Guided Video LLM for Video Temporal Grounding Coin: A large-scale dataset for comprehensive instructional video analysis
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b25f3dc0-f091-4f80-a22d-fb0b32701f85 · outbound
TimeExpert: An Expert-Guided Video LLM for Video Temporal Grounding Cider: Consensus-based image description evalua- tion
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation cb8d8afd-6d9a-4832-a6ef-ffdc5fd7d998 · outbound
TimeExpert: An Expert-Guided Video LLM for Video Temporal Grounding Grounded-VideoLLM: Sharpening Fine-grained Temporal Grounding in Video Large Language Models
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 93fa13a4-7f4b-47c4-a171-aad8de3a708e · outbound
TimeExpert: An Expert-Guided Video LLM for Video Temporal Grounding Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d2919905-838c-4982-8446-f1457ca57e9c · outbound
TimeExpert: An Expert-Guided Video LLM for Video Temporal Grounding End-to-end dense video captioning with parallel decoding
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d3969c71-fd3e-4fd9-a942-9c16d6e17c84 · outbound
TimeExpert: An Expert-Guided Video LLM for Video Temporal Grounding InternVideo: General Video Foundation Models via Generative and Discriminative Learning
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 87fcc106-ab07-4f0a-8792-1a6c1b708276 · outbound
TimeExpert: An Expert-Guided Video LLM for Video Temporal Grounding HawkEye: Training Video-Text LLMs for Grounding Text in Videos
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b558d7db-a77f-49b3-bc16-66330f50fb08 · outbound
TimeExpert: An Expert-Guided Video LLM for Video Temporal Grounding Star: A benchmark for situated reasoning in real-world videos
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3726fe4c-6468-4c6f-bd24-70914bd28998 · outbound
TimeExpert: An Expert-Guided Video LLM for Video Temporal Grounding A large cross- modal video retrieval dataset with reading comprehension
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 63d58ae2-fa10-454f-98de-3299a0f51c49 · outbound
TimeExpert: An Expert-Guided Video LLM for Video Temporal Grounding Multi-head mixture-of-experts
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4c366bce-39e4-4cd0-8593-0e30cda2ded0 · outbound
TimeExpert: An Expert-Guided Video LLM for Video Temporal Grounding Next-qa: Next phase of question-answering to explaining temporal actions
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ba09c676-ddfd-4cd9-a31b-72a8c20d5874 · outbound
TimeExpert: An Expert-Guided Video LLM for Video Temporal Grounding Videoclip: Contrastive pre-training for zero-shot video-text understanding
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7af50d6b-ec72-4721-b2f1-72a2dbe5712b · outbound
TimeExpert: An Expert-Guided Video LLM for Video Temporal Grounding PLLaVA : Parameter-free LLaVA Extension from Images to Videos for Video Dense Captioning
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ea7a8c66-e388-4263-897b-46e884175c77 · outbound
TimeExpert: An Expert-Guided Video LLM for Video Temporal Grounding M6-T: Exploring Sparse Expert Models and Beyond
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 851e325b-d961-44ab-8b75-d640d8e482e4 · outbound
TimeExpert: An Expert-Guided Video LLM for Video Temporal Grounding Vid2seq: Large-scale pretraining of a vi- sual language model for dense video captioning
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5a8dcb2a-b125-4ab5-81b9-631efe2ed276 · outbound
TimeExpert: An Expert-Guided Video LLM for Video Temporal Grounding Xmoe: Sparse models with fine-grained and adaptive expert selection
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e3580b97-3d5a-4a1a-8e90-e0f4f2709d2b · outbound
TimeExpert: An Expert-Guided Video LLM for Video Temporal Grounding Hierarchical video-moment retrieval and step-captioning
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4aaf3c91-7ccc-49bc-a41f-3d144cf8a627 · outbound
TimeExpert: An Expert-Guided Video LLM for Video Temporal Grounding Unimd: Towards unifying moment retrieval and temporal ac- tion detection
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 446f505a-9a3c-48ad-b300-2ea55a96d67f · outbound
TimeExpert: An Expert-Guided Video LLM for Video Temporal Grounding Adamoe: Token-adaptive routing with null ex- perts for mixture-of-experts language models
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 40fedbaa-ed58-438a-ac61-d11118b75316 · outbound
TimeExpert: An Expert-Guided Video LLM for Video Temporal Grounding Sigmoid loss for language image pre-training
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation fc61fe64-0088-4c2f-b95f-349400158007 · outbound
TimeExpert: An Expert-Guided Video LLM for Video Temporal Grounding LLaVA-Video: Video Instruction Tuning With Synthetic Data
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 01d9109f-96a0-482e-a24e-52e57a29f998 · outbound
TimeExpert: An Expert-Guided Video LLM for Video Temporal Grounding Towards automatic learning of procedures from web instructional videos
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e2d15f56-e33b-49f5-9a5f-b9ea67da2ed7 · outbound
TimeExpert: An Expert-Guided Video LLM for Video Temporal Grounding ST-MoE: Designing Stable and Transferable Sparse Expert Models
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8ca6f62c-11c8-46a6-8623-e65dd0c21a46 · inbound
Towards Temporal Compositional Reasoning in Long-Form Sports Videos TimeExpert: An Expert-Guided Video LLM for Video Temporal Grounding
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a4663677-3956-4d70-991c-cadd2285c595 · inbound
Towards Temporal Compositional Reasoning in Long-Form Sports Videos TimeExpert: An Expert-Guided Video LLM for Video Temporal Grounding
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.