Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 15 inbound Pith citation observations for arXiv:2412.09596.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-08T22:47:39.376946Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T03:19:29.961741Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation c4b67586-0fda-4dfd-a13e-849b0c11f5c8 · inbound
Ola: Pushing the Frontiers of Omni-Modal Language Model InternLM-XComposer2.5-OmniLive: A Comprehensive Multimodal System for Long-term Streaming Video and Audio Interactions
Reference 77
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 75b57934-9208-4775-af26-56708e043fd2 · inbound
Phi-4-Mini Technical Report: Compact yet Powerful Multimodal Language Models via Mixture-of-LoRAs InternLM-XComposer2.5-OmniLive: A Comprehensive Multimodal System for Long-term Streaming Video and Audio Interactions
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation e7381476-3277-4121-a530-b252841a80b1 · inbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence InternLM-XComposer2.5-OmniLive: A Comprehensive Multimodal System for Long-term Streaming Video and Audio Interactions
Reference 74
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0adb5e62-b5c0-4719-9158-042dcd1b604f · inbound
Know-MRI: A Knowledge Mechanisms Revealer&Interpreter for Large Language Models InternLM-XComposer2.5-OmniLive: A Comprehensive Multimodal System for Long-term Streaming Video and Audio Interactions
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4cef6c13-6340-4263-8359-99ce2afbf1c6 · inbound
Native Visual Understanding: Resolving Resolution Dilemmas in Vision-Language Models InternLM-XComposer2.5-OmniLive: A Comprehensive Multimodal System for Long-term Streaming Video and Audio Interactions
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3ec31ea1-a90e-4493-a56b-c09be0b7b9bb · inbound
HumanOmniV2: From Understanding to Omni-Modal Reasoning with Context InternLM-XComposer2.5-OmniLive: A Comprehensive Multimodal System for Long-term Streaming Video and Audio Interactions
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6b46ae8c-9509-456b-aa26-197f2265ce28 · inbound
HumanPCR: Probing MLLM Capabilities in Diverse Human-Centric Scenes InternLM-XComposer2.5-OmniLive: A Comprehensive Multimodal System for Long-term Streaming Video and Audio Interactions
Reference 103
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c6146844-f8ee-49d0-9297-b77bace798b1 · inbound
Skyra: AI-Generated Video Detection via Grounded Artifact Reasoning InternLM-XComposer2.5-OmniLive: A Comprehensive Multimodal System for Long-term Streaming Video and Audio Interactions
Reference 84
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 12da96cb-a329-43fa-9797-2526cd9f7f49 · inbound
Character Beyond Speech: Leveraging Role-Playing Evaluation in Audio Large Language Models via Reinforcement Learning InternLM-XComposer2.5-OmniLive: A Comprehensive Multimodal System for Long-term Streaming Video and Audio Interactions
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation e404bb10-364e-4aa7-977c-1807aaf36f69 · inbound
Beyond the Current Observation: Evaluating Multimodal Large Language Models in Controllable Non-Markov Games InternLM-XComposer2.5-OmniLive: A Comprehensive Multimodal System for Long-term Streaming Video and Audio Interactions
Reference 92
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation d8002c6b-4f07-46a7-b86c-10a96f7a00a3 · inbound
ViCoStream: Streaming VideoLLMs Can Run Beyond 100 FPS with Stage-Wise Coordinated Inference InternLM-XComposer2.5-OmniLive: A Comprehensive Multimodal System for Long-term Streaming Video and Audio Interactions
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation a25e7ef2-4c1f-492c-8e1e-295760706f59 · inbound
Light-Omni: Reflex over Reasoning in Agentic Video Understanding with Long-Term Memory InternLM-XComposer2.5-OmniLive: A Comprehensive Multimodal System for Long-term Streaming Video and Audio Interactions
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c7153a65-386a-4160-b939-63be0ae1703d · inbound
FOLIO: Focused Semantic Memory for Streaming Video Understanding InternLM-XComposer2.5-OmniLive: A Comprehensive Multimodal System for Long-term Streaming Video and Audio Interactions
Reference 79
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8a354390-4034-47f5-b432-5c9212b803d4 · inbound
X$^3$-OPD: Distilling Reasoning into Large Audio-Language Models via On-Policy Alignment InternLM-XComposer2.5-OmniLive: A Comprehensive Multimodal System for Long-term Streaming Video and Audio Interactions
Reference 134
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 23380163-05c6-4472-aa1b-ba3150efe153 · inbound
Mage-VL: An Efficient Codec-Native Streaming Multimodal Foundation Model InternLM-XComposer2.5-OmniLive: A Comprehensive Multimodal System for Long-term Streaming Video and Audio Interactions
Reference 76
Source-reported events for the cited work
Unavailable: canonical work link unavailable.