Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 11 inbound Pith citation observations for arXiv:2404.05726.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:46:44.831913Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T06:39:37.683925Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 10ff7284-3e28-4fa0-9804-a77c502c2499 · inbound
DriveVLM: The Convergence of Autonomous Driving and Large Vision-Language Models MA-LMM: Memory-Augmented Large Multimodal Model for Long-Term Video Understanding
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1fda6659-f351-4b47-9145-27bc985f5b95 · inbound
MLVU: Benchmarking Multi-task Long Video Understanding MA-LMM: Memory-Augmented Large Multimodal Model for Long-Term Video Understanding
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e3558315-0c00-4d59-819c-9fbc06c7dbbb · inbound
VideoLLaMA 2: Advancing Spatial-Temporal Modeling and Audio Understanding in Video-LLMs MA-LMM: Memory-Augmented Large Multimodal Model for Long-Term Video Understanding
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e91f77fe-e0a8-43dd-a17d-f75572f8d718 · inbound
InternLM-XComposer-2.5: A Versatile Large Vision Language Model Supporting Long-Contextual Input and Output MA-LMM: Memory-Augmented Large Multimodal Model for Long-Term Video Understanding
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9eb46eb0-8e47-4268-81dc-38ff53e2fe95 · inbound
Towards General Continuous Memory for Vision-Language Models MA-LMM: Memory-Augmented Large Multimodal Model for Long-Term Video Understanding
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fdc0e26e-7bbc-413b-abce-bc2b993963ca · inbound
ViFusion: In-Network Tensor Fusion for Scalable Video Feature Indexing MA-LMM: Memory-Augmented Large Multimodal Model for Long-Term Video Understanding
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8ecbc304-40b5-428f-8234-871d5a1d251c · inbound
Video-XL-2: Towards Very Long-Video Understanding Through Task-Aware KV Sparsification MA-LMM: Memory-Augmented Large Multimodal Model for Long-Term Video Understanding
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2800f197-06c6-49eb-bd7d-99a1b40740be · inbound
Task-Aware KV Compression For Cost-Effective Long Video Understanding MA-LMM: Memory-Augmented Large Multimodal Model for Long-Term Video Understanding
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 604b0fd2-5bef-4b08-adb0-c9b68053db71 · inbound
Seeing the Scene Matters: Revealing Forgetting in Video Understanding Models with a Scene-Aware Long-Video Benchmark MA-LMM: Memory-Augmented Large Multimodal Model for Long-Term Video Understanding
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation bb49de16-191a-4ae0-9658-6af5cbc4c952 · inbound
PyraVid: Hierarchical Multimodal Memory for Long-Horizon Video Reasoning MA-LMM: Memory-Augmented Large Multimodal Model for Long-Term Video Understanding
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6a8ce2a8-7760-465a-8253-1ddea61a5315 · inbound
HPP: Hierarchical Programmatic Probing for Long Video Understanding by Decoupling Perception and Reasoning MA-LMM: Memory-Augmented Large Multimodal Model for Long-Term Video Understanding
Reference 156
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.