Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 10 inbound Pith citation observations for arXiv:2406.06040.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T04:58:10.639605Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-18T04:02:43.573437Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 9138f91f-6504-4749-a4bc-281226eec477 · inbound
Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling Vript: A Video Is Worth Thousands of Words
Reference 270
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 717bf914-3f9a-417d-8baa-b96a6588e2d0 · inbound
Open-Sora: Democratizing Efficient Video Production for All Vript: A Video Is Worth Thousands of Words
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5cfa7a2c-3401-4e26-b69e-c5672311255c · inbound
VideoChat-Flash: Hierarchical Compression for Long-Context Video Modeling Vript: A Video Is Worth Thousands of Words
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 50d8304b-f818-4452-a6c0-04aa645e7471 · inbound
InternVideo2.5: Empowering Video MLLMs with Long and Rich Context Modeling Vript: A Video Is Worth Thousands of Words
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 494bda89-ca82-47cf-91b1-46a129d9f1df · inbound
SmolVLM: Redefining small and efficient multimodal models Vript: A Video Is Worth Thousands of Words
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 049d0006-7368-460b-a253-3cfe11c2066f · inbound
Ming-Omni: A Unified Multimodal Model for Perception and Generation Vript: A Video Is Worth Thousands of Words
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ae47c4a9-105c-400c-bcc5-1530b30d7e8f · inbound
CI-VID: A Coherent Interleaved Text-Video Dataset Vript: A Video Is Worth Thousands of Words
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aec39a04-ce61-4d56-92ae-d28ea6287923 · inbound
Adversarial Distribution Matching for Diffusion Distillation Towards Efficient Image and Video Synthesis Vript: A Video Is Worth Thousands of Words
Reference 82
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a1cf3dde-b5ac-47fb-af1c-7cd6dcf65ebe · inbound
MESH -- Understanding Videos Like Human: Measuring Hallucinations in Large Video Models Vript: A Video Is Worth Thousands of Words
Reference 81
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 696b1e40-7be9-4046-870a-3400740071da · inbound
VideoChat3: Fully Open Video MLLM for Efficient and Generalist Video Understanding Vript: A Video Is Worth Thousands of Words
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.