Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 27 inbound Pith citation observations for arXiv:2312.02051.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-08T17:37:38.120634Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T06:39:37.643725Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 2894a04d-1d74-4417-a654-57b721d1551d · inbound
TempCompass: Do Video LLMs Really Understand Videos? TimeChat: A Time-sensitive Multimodal Large Language Model for Long Video Understanding
Reference 116
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation cd0c9261-42f4-4beb-8ad3-e83f25ecae50 · inbound
MLVU: Benchmarking Multi-task Long Video Understanding TimeChat: A Time-sensitive Multimodal Large Language Model for Long Video Understanding
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation e1fd15f6-ca78-464e-9243-f4693104d822 · inbound
VideoLLaMA 2: Advancing Spatial-Temporal Modeling and Audio Understanding in Video-LLMs TimeChat: A Time-sensitive Multimodal Large Language Model for Long Video Understanding
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation ed7db905-7503-481c-b1d0-c2d6a37bc33b · inbound
LVBench: An Extreme Long Video Understanding Benchmark TimeChat: A Time-sensitive Multimodal Large Language Model for Long Video Understanding
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation ca324a57-d643-4590-a8bb-7f477c6345fe · inbound
CogVLM2: Visual Language Models for Image and Video Understanding TimeChat: A Time-sensitive Multimodal Large Language Model for Long Video Understanding
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 74733113-350f-4079-852c-d6a93c738232 · inbound
InternVideo2.5: Empowering Video MLLMs with Long and Rich Context Modeling TimeChat: A Time-sensitive Multimodal Large Language Model for Long Video Understanding
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation b7bb7727-ba7d-4ecc-ba1e-6079cc4e9f21 · inbound
MTPChat: A Multimodal Time-Aware Persona Dataset for Conversational Agents TimeChat: A Time-sensitive Multimodal Large Language Model for Long Video Understanding
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0bbf7527-2e0a-43dd-8862-846ec2812208 · inbound
RICO: Improving Accuracy and Completeness in Image Recaptioning via Visual Reconstruction TimeChat: A Time-sensitive Multimodal Large Language Model for Long Video Understanding
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation db37397a-99a4-4b89-a1ef-8f05e1286840 · inbound
EASG-Bench: Video Q&A Benchmark with Egocentric Action Scene Graphs TimeChat: A Time-sensitive Multimodal Large Language Model for Long Video Understanding
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 039f19b2-59e5-41e2-b0c1-42a3e1a0c856 · inbound
LLaVA-Scissor: Token Compression with Semantic Connected Components for Video LLMs TimeChat: A Time-sensitive Multimodal Large Language Model for Long Video Understanding
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c76d8af5-6ecf-405e-8499-77ad6f20ccc0 · inbound
DIVE: Deep-search Iterative Video Exploration A Technical Report for the CVRR Challenge at CVPR 2025 TimeChat: A Time-sensitive Multimodal Large Language Model for Long Video Understanding
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ce5d04b7-30d0-4adc-8fce-0d5913fe9b60 · inbound
MANTA: Cross-Modal Semantic Alignment and Information-Theoretic Optimization for Long-form Multimodal Understanding TimeChat: A Time-sensitive Multimodal Large Language Model for Long Video Understanding
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b85caa48-9723-4b62-a11a-6e2c148e0057 · inbound
A Paradigm Shift: Fully End-to-End Training for Temporal Sentence Grounding in Videos TimeChat: A Time-sensitive Multimodal Large Language Model for Long Video Understanding
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation bbc5f630-013a-40d7-a1a7-2ec06f18f519 · inbound
MarkIt: Training-Free Visual Markers for Precise Video Temporal Grounding TimeChat: A Time-sensitive Multimodal Large Language Model for Long Video Understanding
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 7ba36f90-8d5b-4224-860b-09a8bac66510 · inbound
MarkIt: Training-Free Visual Markers for Precise Video Temporal Grounding TimeChat: A Time-sensitive Multimodal Large Language Model for Long Video Understanding
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 62853dd5-b767-49ba-a9a9-4a13f37bcfab · inbound
LLaVA-OneVision-2: Towards Next-Generation Perceptual Intelligence TimeChat: A Time-sensitive Multimodal Large Language Model for Long Video Understanding
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation c5036575-fc9c-419e-84c4-37101201c4e5 · inbound
VTI-CoT: Visual-Textual Interleaved Chain of Thought for Video Reasoning TimeChat: A Time-sensitive Multimodal Large Language Model for Long Video Understanding
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 99f19cce-399a-45ea-af49-7a991abf30fc · inbound
Watch, Remember, Reason: Human-View Video Understanding with MLLMs TimeChat: A Time-sensitive Multimodal Large Language Model for Long Video Understanding
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 720ab0c3-e66d-454c-a5d5-e85fc57d5822 · inbound
HPP: Hierarchical Programmatic Probing for Long Video Understanding by Decoupling Perception and Reasoning TimeChat: A Time-sensitive Multimodal Large Language Model for Long Video Understanding
Reference 193
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 3e1094fb-58b5-43cf-8056-6f2b0e2139b4 · inbound
EFlow: Learning Evidence Flow for Long-Video Reasoning with Adaptive Reflection TimeChat: A Time-sensitive Multimodal Large Language Model for Long Video Understanding
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation af988246-9c7f-4120-a332-cdde69102010 · inbound
EFlow: Learning Evidence Flow for Long-Video Reasoning with Adaptive Reflection TimeChat: A Time-sensitive Multimodal Large Language Model for Long Video Understanding
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 66d9599d-a935-4146-8499-6ec150eea4e5 · inbound
EFlow: Learning Evidence Flow for Long-Video Reasoning with Adaptive Reflection TimeChat: A Time-sensitive Multimodal Large Language Model for Long Video Understanding
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0e22faf1-5308-4869-a1b5-4ca00032fe8c · inbound
DELTAVID: Enhancing Fine-Grained Spatiotemporal Perception with Cross-Video Differences TimeChat: A Time-sensitive Multimodal Large Language Model for Long Video Understanding
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ddf24105-f1b7-4d79-bbad-836acba77bba · inbound
Mixture of Probes: Learning from Privileged Modalities in Multimodal LLMs Through Probing TimeChat: A Time-sensitive Multimodal Large Language Model for Long Video Understanding
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b0c52bc0-b812-4f45-9d1b-c2b9ab76835b · inbound
TimeLens2: Generalist Video Temporal Grounding with Multimodal LLMs TimeChat: A Time-sensitive Multimodal Large Language Model for Long Video Understanding
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 75c7b714-5509-4238-bc47-182e8a457c22 · inbound
TimePLE: Rethinking Temporal Representation for Video Temporal Grounding TimeChat: A Time-sensitive Multimodal Large Language Model for Long Video Understanding
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7373f027-c75e-455f-a08b-410c4054f616 · inbound
AdaThinkV: Adaptive Thinking for Token-Efficient Video Reasoning TimeChat: A Time-sensitive Multimodal Large Language Model for Long Video Understanding
Reference 67
Source-reported events for the cited work
Unavailable: canonical work link unavailable.