Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-05T23:32:17.911414Z
Paper Citation Record · LEDGER
As of 20 August 2026, this Paper Citation Record lists 100 of 164 outbound references and 3 inbound Pith citation observations for arXiv:2508.10922.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-05T23:32:17.911414Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-04T13:27:58.123306Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T03:29:31.264199Z
100 of 164 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation d47efee2-bed4-4c74-8608-84941efa5a9a · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model Unresolved cited work
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b6867581-5b02-4758-9364-ed0c7a825638 · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model InternVideo: General Video Foundation Models via Generative and Discriminative Learning
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f1c7ca98-9471-4fc4-aba7-1f20aa7d5cbd · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model Unresolved cited work
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 710efd66-7ff8-480c-b180-e74d17545b68 · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model Unresolved cited work
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d3dd2b83-3a52-41f0-b953-5241e969b2e0 · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model Unresolved cited work
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 95f37fb3-323d-4daf-9765-a459f1fadd24 · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model Regneri, M
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b3dd524e-b921-401e-99ae-cc67967b3084 · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model Anne Hendricks, O
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 161fa926-d0a6-4d94-81f0-88a26f8fd460 · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model Krishna, K
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f0bdd02f-48d0-4c82-86ba-4d65edb917a7 · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model Unresolved cited work
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 49472454-c85a-4843-8d08-80e58e3c1e08 · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model Unresolved cited work
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b64cca6f-fa04-4422-bbd6-4e97073754eb · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model Xiong, Y
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 07245b3a-70d3-41fa-9714-4e67115b95c7 · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model Unresolved cited work
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 066bf953-70a3-4a21-b9b4-ef8ba83ced6e · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model Chen, Y.-C
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3e5bc3c0-2e9a-4d2e-b5e2-a760029593be · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model Zhang, X
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0d5c2cb1-6af6-4817-8dd1-d749ecdf7b34 · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model Unresolved cited work
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d7347c27-3b3c-47fa-8d06-1442d2c1b9b8 · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model Unresolved cited work
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1964705f-3d3e-41aa-b950-c773e4027ee4 · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model Zhang, Y
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 358ddc0a-c1fb-498f-8ba1-e81a049dc7bd · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model Unresolved cited work
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1cfede5a-2869-407c-a0a5-f8716337bc1f · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model Unresolved cited work
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 65bd9630-41a3-46b9-916e-bed277a27ef1 · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model OPT: Open Pre-trained Transformer Language Models
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2bcbd8b2-c144-4457-a91c-3df27b2dc296 · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model Unresolved cited work
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 907589c4-ac8a-4711-86b1-82450819d6cd · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 33dc5993-1d22-4d4d-ad40-f50ad8e7cc77 · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model Unresolved cited work
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f3e11a11-382d-4dc5-a13a-9dc3717a4ab0 · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model Unresolved cited work
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a2e85720-5584-4f24-bd07-71b9ebab44f8 · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model VideoLLaMA 2: Advancing Spatial-Temporal Modeling and Audio Understanding in Video-LLMs
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation de74602a-762c-4ac9-9b46-cca957b72f45 · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model PLLaVA : Parameter-free LLaVA Extension from Images to Videos for Video Dense Captioning
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 44d77a73-e0ac-48f4-822b-474fab802c87 · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model Huang, X
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c8e0f96d-0b6c-408f-aede-f6307fd30b97 · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model Unresolved cited work
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d330d962-caea-4a7e-bd1c-fbf39944b073 · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model Unresolved cited work
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 21ab5dd0-d1ef-4df7-b288-9b8eb5d86e05 · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model Unresolved cited work
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8094023f-5c7d-4f0c-95cf-a041a5174776 · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model Question-Answering Dense Video Events
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 434ed86b-354c-41ac-bbfa-ed746c68d012 · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model Unresolved cited work
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 989f777a-e959-4819-8ad1-662748ee862c · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model HawkEye: Training Video-Text LLMs for Grounding Text in Videos
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a08790b8-dd72-424b-85b7-b1c9ae0dffc5 · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model Unresolved cited work
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5ee27c7d-fb51-4241-8eaf-ef831dfba3f3 · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model Grounded-VideoLLM: Sharpening Fine-grained Temporal Grounding in Video Large Language Models
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 95e527c9-5938-419e-8dbf-f4abc8847158 · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model Unresolved cited work
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ef7c5fcc-0cc6-4ebd-b0cc-dc5daecb2176 · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model Multimodal Foundation Models: From Specialists to General-Purpose Assistants
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a1115ff8-dc7e-4c5d-b44e-ff31a2438b41 · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model A Comprehensive Study of Deep Video Action Recognition
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2975d617-022c-4f22-86bf-9aa15455e0a7 · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model Abdar, M
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 846e7eb6-fa64-48a2-b450-500334a692f1 · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model Unresolved cited work
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d4881420-e7fa-4c70-b3e5-fa63564f43e9 · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model Zhang, J
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 014b09cb-8115-4e53-a693-aabaadb268e0 · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model A Survey on Natural Language Video Localization
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 21fbba2c-03c7-447c-bdd3-e3ca6c267ddf · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model Unresolved cited work
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 55320261-df96-4821-9363-464649fb13b0 · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model Unresolved cited work
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8ffb8ef7-7d1a-4ad8-8f33-38b013ebddb5 · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model Zhang, A
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d2d29c30-0c5b-47d8-9ec2-111db965fa1a · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model Zhang, H
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 559c54d0-d5d1-4c3a-8da5-f6fbbc1c6f72 · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model Unresolved cited work
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a9dbd8d1-b3d8-4adb-be10-5b0c26411e1a · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model Unresolved cited work
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c93a3c89-3f11-48d1-9b9f-3b732bd20044 · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model Unresolved cited work
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 471c449e-de1f-4052-98be-129190e86b9a · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model Unresolved cited work
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 785bfcf4-c5d7-45b1-b340-75bc7c769e3f · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model Zhang, A
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f610a675-51b8-4e4d-ba0b-882b260d5c68 · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model Unresolved cited work
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 10c6805a-8eef-4283-ace3-ca14d249584a · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model Unresolved cited work
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 544ec9ee-5746-496c-a53c-15a78c722964 · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model Iashin and E
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4c3be1a9-f883-4e45-9cc1-4acb42f792f8 · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model Unresolved cited work
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 19e47eb9-e53e-4ef6-b09f-014ff4bff5ee · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model CG-Bench: Clue-grounded Question Answering Benchmark for Long Video Understanding
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eb7c405f-827e-40fe-9f4e-d9a7da9559bd · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model Unresolved cited work
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3d3bdab4-1c93-4a54-b42b-ecc0b7b0bbff · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model Zhang, X
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c7011462-3c0d-4498-bc56-d0c1e7897347 · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model Video-LLaVA: Learning United Visual Representation by Alignment Before Projection
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 03c0c3d3-0db8-4f80-8ec7-9fd97c3db266 · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model VideoChat: Chat-Centric Video Understanding
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6b544616-83ca-4226-8f75-51a7f3224619 · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model Dosovitskiy, L
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3168c6c6-3f0d-4066-88d2-dc460577a500 · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model EVA-CLIP: Improved Training Techniques for CLIP at Scale
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 47c5f693-624f-4adc-99b1-43066b9d9b89 · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model Unresolved cited work
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d9769dd9-7495-450c-bd61-cbe0690111c7 · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model Radford, J
Reference 64
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1d6a3f28-cf46-4f13-be85-7c7b3ff04869 · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model Feichtenhofer, Y
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1eb1a059-d6de-4815-a9a2-f124e60d9c9a · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model Unresolved cited work
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8a9ae72e-bdc2-40b4-b00d-fb97646cdeec · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model Unresolved cited work
Reference 67
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e0f9fe1b-f90e-4d06-9af1-480b6a5ef781 · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model Unresolved cited work
Reference 68
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6d3917f6-fd6a-4437-b08e-e1c464182e5f · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model Unresolved cited work
Reference 69
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aecfe681-e828-44a6-94f9-67d088684a43 · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model HierarQ: Task-Aware Hierarchical Q-Former for Enhanced Video Understanding
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 90b69520-e057-43d9-8cc0-42c77af96758 · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model Alayrac, J
Reference 71
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e5f12779-6a3e-4707-ad9e-4d98c238e7a1 · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model Unresolved cited work
Reference 72
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation af7cf594-9292-4f89-8d3e-4663faab3b93 · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model mPLUG-Owl: Modularization Empowers Large Language Models with Multimodality
Reference 73
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 795bc96e-a601-40d7-982a-70f0bc170e33 · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model Huang, L
Reference 74
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9c32c71b-70ce-46b2-b471-906b0464179a · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model Unresolved cited work
Reference 75
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8d9f69a4-1136-4732-8936-b794417cb6bd · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model LLaVA-NeXT-Interleave: Tackling Multi-image, Video, and 3D in Large Multimodal Models
Reference 76
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1949ae06-50bd-4c89-a926-48883de00917 · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model mPLUG-Owl3: Towards Long Image-Sequence Understanding in Multi-Modal Large Language Models
Reference 77
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 865f3073-7922-4280-8fb8-51adb5bb41d0 · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model EVLM: An Efficient Vision-Language Model for Visual Understanding
Reference 78
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 89728ef8-fc4a-4e38-ae24-691b01dab7f8 · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model Slow-Fast Architecture for Video Multi-Modal Large Language Models
Reference 79
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 819149ec-2a90-4398-bce9-4c9b07ecff66 · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model Miech, D
Reference 80
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d9c08757-20ed-4c5c-bfb7-01ad525618cd · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model Unresolved cited work
Reference 81
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f1de355d-a83b-4b4c-98ec-849371a1a58c · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model Schuhmann, R
Reference 82
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4d0bd7e3-10c6-4259-acb1-704fa9e30622 · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model M$^3$IT: A Large-Scale Dataset towards Multi-Modal Multilingual Instruction Tuning
Reference 83
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f1561613-df99-4fd2-96ae-a26e9ea1434e · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model Unresolved cited work
Reference 84
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bbd18ff9-bd60-41e7-8f7d-00d4155bd2e8 · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model Unresolved cited work
Reference 85
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cde0115a-451a-46d7-9c81-bc72ae14062d · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model Unresolved cited work
Reference 86
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 45f422a6-bf94-4aba-a52a-2f60719e77aa · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model Visual Prompting in Multimodal Large Language Models: A Survey
Reference 87
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0db8b3c9-30bd-480e-a287-ede2fcbd48a3 · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model Unresolved cited work
Reference 88
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8aff295d-f45c-49bf-8711-5fbe9bc37e66 · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model Unresolved cited work
Reference 89
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 90112b56-42bf-4d27-b410-da5d77c56cbe · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model Liu, C.-Y
Reference 90
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7ffdd6e3-785b-4995-a9af-7dd2950cb6fb · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model MLLM as Video Narrator: Mitigating Modality Imbalance in Video Moment Retrieval
Reference 91
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 40ef3941-45ca-4495-967b-b3e6ee5cf6a0 · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model Di and W
Reference 92
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 156dfd67-c440-4efa-851d-09709eb8606e · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model Zheng, X
Reference 93
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5ddbcca4-14eb-4c7a-b53e-c010ce2d2124 · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model Grounding-Prompter: Prompting LLM with Multimodal Information for Temporal Sentence Grounding in Long Videos
Reference 94
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 52300739-7cee-474e-977b-5d4e9eb60b6b · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model Unresolved cited work
Reference 95
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 96585e33-88e0-4671-a87b-258ad57956e1 · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model VERIFIED: A Video Corpus Moment Retrieval Benchmark for Fine-Grained Video Understanding
Reference 96
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 20f51e9d-5358-4905-ac83-e8e182b11593 · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model Infusing Environmental Captions for Long-Form Video Language Grounding
Reference 97
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation b960ac46-3284-4199-99bf-247be9d6cc91 · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model Unresolved cited work
Reference 98
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 73d51dbf-dc62-48a4-9c7e-12dc4c4c26af · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model Unresolved cited work
Reference 99
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a55f8afc-6794-49d6-b2e8-f7c444f1765a · outbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model Context-Enhanced Video Moment Retrieval with Large Language Models
Reference 100
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 88ca922f-510e-4504-9f08-5e73b4c37d34 · inbound
Training-free Uncertainty Guidance for Complex Visual Tasks with MLLMs A Survey on Video Temporal Grounding with Multimodal Large Language Model
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c62788a6-7f32-4a43-bd09-750b057d5d86 · inbound
Natural-Language Temporal Grounding in Hour-Long Videos is a Search Problem: A Benchmark and Empirical Decomposition A Survey on Video Temporal Grounding with Multimodal Large Language Model
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation b4dd3e4b-6a31-4ccd-b5c9-eab64f69718f · inbound
NEST: Narrative Event Structures in Time for Long Video Understanding A Survey on Video Temporal Grounding with Multimodal Large Language Model
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.