Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 24 inbound Pith citation observations for arXiv:2412.10360.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T15:41:08.568366Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-03T10:48:02.918791Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 51649a7d-f4f2-4ec0-8e87-21456cbedbd9 · inbound
InternVideo2.5: Empowering Video MLLMs with Long and Rich Context Modeling Apollo: An Exploration of Video Understanding in Large Multimodal Models
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3ca642b1-742d-46e3-9920-106d019fb4c1 · inbound
VideoLLaMA 3: Frontier Multimodal Foundation Models for Image and Video Understanding Apollo: An Exploration of Video Understanding in Large Multimodal Models
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 73db5c87-2b89-4b6e-abd9-9da8866b837f · inbound
Breaking Down Video LLM Benchmarks: Knowledge, Spatial Perception, or True Temporal Understanding? Apollo: An Exploration of Video Understanding in Large Multimodal Models
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 208b340b-771f-4eda-aa3a-b4fde362b2f0 · inbound
UniGen: Enhanced Training & Test-Time Strategies for Unified Multimodal Understanding and Generation Apollo: An Exploration of Video Understanding in Large Multimodal Models
Reference 97
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3e01e98a-c288-4951-bc8a-ad6f1c0fdcf2 · inbound
Modality Curation: Building Universal Embeddings for Advanced Multimodal Information Retrieval Apollo: An Exploration of Video Understanding in Large Multimodal Models
Reference 120
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 447bbede-b0be-4fbc-9feb-22827ba5a7bd · inbound
Threading Keyframe with Narratives: MLLMs as Strong Long Video Comprehenders Apollo: An Exploration of Video Understanding in Large Multimodal Models
Reference 84
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0ff19277-4d81-42ea-a6ca-e0b1fdb91644 · inbound
FlexSelect: Flexible Token Selection for Efficient Long Video Understanding Apollo: An Exploration of Video Understanding in Large Multimodal Models
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1ab9b6da-ae28-476b-9d98-ef9bd5fe884c · inbound
Beyond Text Compression: Evaluating Tokenizers Across Scales Apollo: An Exploration of Video Understanding in Large Multimodal Models
Reference 116
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1645717c-29a1-4eb3-8e9f-e7e2978c2325 · inbound
ARGUS: Hallucination and Omission Evaluation in Video-LLMs Apollo: An Exploration of Video Understanding in Large Multimodal Models
Reference 68
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aff8f26f-bed7-41c0-a3c2-a8a286242dcd · inbound
LLaVA-Scissor: Token Compression with Semantic Connected Components for Video LLMs Apollo: An Exploration of Video Understanding in Large Multimodal Models
Reference 115
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0a20b61a-a85a-4611-93dc-11fbe6e68ff3 · inbound
AuroraLong: Bringing RNNs Back to Efficient Open-Ended Video Understanding Apollo: An Exploration of Video Understanding in Large Multimodal Models
Reference 123
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5f9be675-8665-4051-ba81-0ff362f1473c · inbound
ExpStar: Towards Automatic Commentary Generation for Multi-discipline Scientific Experiments Apollo: An Exploration of Video Understanding in Large Multimodal Models
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 44079027-0ad8-4f25-8acd-0306ad0d4fb3 · inbound
"Harmless to You, Hurtful to Me!": Investigating the Detection of Toxic Languages Grounded in the Perspective of Youth Apollo: An Exploration of Video Understanding in Large Multimodal Models
Reference 104
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6cf8de07-4242-40f0-a445-23d5848dd7d5 · inbound
VLM4D: Towards Spatiotemporal Awareness in Vision Language Models Apollo: An Exploration of Video Understanding in Large Multimodal Models
Reference 105
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4ebc1d82-7cb1-49f5-9264-9aaf78b8426d · inbound
PixFoundation 2.0: Do Video Multi-Modal LLMs Use Motion in Visual Grounding? Apollo: An Exploration of Video Understanding in Large Multimodal Models
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bf372c0d-9bca-492b-a2ab-78d4e97b3cb7 · inbound
Adapting MLLMs for Nuanced Video Retrieval Apollo: An Exploration of Video Understanding in Large Multimodal Models
Reference 92
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 318777ff-820e-4174-8411-3e46e8365194 · inbound
When Text Hijacks Vision: Benchmarking and Mitigating Text Overlay-Induced Hallucination in Vision Language Models Apollo: An Exploration of Video Understanding in Large Multimodal Models
Reference 79
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e639179c-8052-4189-8f40-7fd587b4e7c3 · inbound
Flat-Pack Bench: Evaluating Spatio-Temporal Understanding in Large Vision-Language Models through Furniture Assembly Apollo: An Exploration of Video Understanding in Large Multimodal Models
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0a37124e-a73b-4485-b966-0c8d1371122d · inbound
LLaVA-OneVision-2: Towards Next-Generation Perceptual Intelligence Apollo: An Exploration of Video Understanding in Large Multimodal Models
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3113e5cd-019c-4b5e-9caf-e93928e905c7 · inbound
Multi-Scale Separable Fourier Neural Networks for Solving High-Frequency PDEs Apollo: An Exploration of Video Understanding in Large Multimodal Models
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 842d6468-15ae-4852-8695-f90881a5b040 · inbound
Multi-Scale Separable Fourier Neural Networks for Solving High-Frequency PDEs Apollo: An Exploration of Video Understanding in Large Multimodal Models
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e62a9e4e-ff01-4793-b538-68d25de14127 · inbound
PEEK: Picking Essential frames via Efficient Knowledge distillation Apollo: An Exploration of Video Understanding in Large Multimodal Models
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1b33d8c6-4bb1-49b1-bbfa-ae9e82327cbc · inbound
InternVideo3: Agentify Foundation Models with Multimodal Contextual Reasoning Apollo: An Exploration of Video Understanding in Large Multimodal Models
Reference 263
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c9dbe76f-b26c-4668-9b61-2ec3dea2696c · inbound
Agent-Computer Observation Interfaces Enable Dynamic Computer Use Apollo: An Exploration of Video Understanding in Large Multimodal Models
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.