Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 23 inbound Pith citation observations for arXiv:2410.03051.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:14:37.686789Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T15:09:54.928668Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 831294d4-84b3-4832-bc3a-b50709e32390 · inbound
HumanVBench: Probing Human-Centric Video Understanding in MLLMs with Automatically Synthesized Benchmarks AuroraCap: Efficient, Performant Video Detailed Captioning and a New Benchmark
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 68652cac-f1a8-4700-8de1-5f779442999b · inbound
Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos AuroraCap: Efficient, Performant Video Detailed Captioning and a New Benchmark
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2537b212-0d3b-463a-bec1-b35f1d15dcc5 · inbound
Growing a Multi-head Twig via Distillation and Reinforcement Learning to Accelerate Large Vision-Language Models AuroraCap: Efficient, Performant Video Detailed Captioning and a New Benchmark
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation caca77dd-0c7f-436f-a4a7-f8dbfe66e7db · inbound
Modality Curation: Building Universal Embeddings for Advanced Multimodal Information Retrieval AuroraCap: Efficient, Performant Video Detailed Captioning and a New Benchmark
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 286408a4-2bf5-439d-b8c2-c1f170a644df · inbound
TUNA: Comprehensive Fine-grained Temporal Understanding Evaluation on Dense Dynamic Videos AuroraCap: Efficient, Performant Video Detailed Captioning and a New Benchmark
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 14d775da-00a6-4a49-9a83-208061a1d3d8 · inbound
Vid-SME: Membership Inference Attacks against Large Video Understanding Models AuroraCap: Efficient, Performant Video Detailed Captioning and a New Benchmark
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ea527a28-a167-4066-bada-0ccd3e18f566 · inbound
ARGUS: Hallucination and Omission Evaluation in Video-LLMs AuroraCap: Efficient, Performant Video Detailed Captioning and a New Benchmark
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3bf9bb16-7765-40f5-b839-9587f6f7bfcb · inbound
ToSA: Token Merging with Spatial Awareness AuroraCap: Efficient, Performant Video Detailed Captioning and a New Benchmark
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2e94a8c9-1420-4ad4-97ce-400db8aba0f9 · inbound
AVC-DPO: Aligned Video Captioning via Direct Preference Optimization AuroraCap: Efficient, Performant Video Detailed Captioning and a New Benchmark
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7b21e852-8f3b-4f2d-9372-9a4dce0f8077 · inbound
AuroraLong: Bringing RNNs Back to Efficient Open-Ended Video Understanding AuroraCap: Efficient, Performant Video Detailed Captioning and a New Benchmark
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 50b612f2-6f4e-4f4d-b273-6ace6548816b · inbound
Empowering Nanoscale Connectivity through Molecular Communication: A Case Study of Virus Infection AuroraCap: Efficient, Performant Video Detailed Captioning and a New Benchmark
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 307b969c-1c54-4663-83d5-2209560690f9 · inbound
Token Reduction via Local and Global Contexts Optimization for Efficient Video Large Language Models AuroraCap: Efficient, Performant Video Detailed Captioning and a New Benchmark
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation cdf41a37-fada-4535-901d-9ea0ad180cda · inbound
ForestPrune: High-ratio Visual Token Compression for Video Multimodal Large Language Models via Spatial-Temporal Forest Modeling AuroraCap: Efficient, Performant Video Detailed Captioning and a New Benchmark
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9fdf966d-545e-41f9-a686-acde1a8d4700 · inbound
Building a Precise Video Language with Human-AI Oversight AuroraCap: Efficient, Performant Video Detailed Captioning and a New Benchmark
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1f19ee89-936f-4124-b487-3852d4baad26 · inbound
MSD-Score: Multi-Scale Distributional Scoring for Reference-Free Image Caption Evaluation AuroraCap: Efficient, Performant Video Detailed Captioning and a New Benchmark
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5715f666-a5d1-451f-928e-f5e25c060ee4 · inbound
Auteur: Language-Driven Cinematographic Framing for Human-Centric Video Generation AuroraCap: Efficient, Performant Video Detailed Captioning and a New Benchmark
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9c1e0d0d-2b0d-4e9c-a200-53bc6e838f59 · inbound
Balancing Image Compression and Generation with Bootstrapped Tokenization AuroraCap: Efficient, Performant Video Detailed Captioning and a New Benchmark
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6098b20b-c083-409d-ae59-491209b36e95 · inbound
GOPAgen: Motion-Aware and Efficient Agentic Long-Video Understanding with Structural Memory and Hierarchical Reasoning AuroraCap: Efficient, Performant Video Detailed Captioning and a New Benchmark
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 00fc8319-181c-4837-ad2f-97b1f9e2570d · inbound
Watch, Remember, Reason: Human-View Video Understanding with MLLMs AuroraCap: Efficient, Performant Video Detailed Captioning and a New Benchmark
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7aff1fca-7e12-45ca-a8f0-9ba6eb33cb73 · inbound
CapRiCorn-1K: A Comprehensive Benchmark for Video Captioning and Subject Referential Consistency Across Temporal Scales AuroraCap: Efficient, Performant Video Detailed Captioning and a New Benchmark
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 48128038-df63-4ab8-b3bf-897687c62909 · inbound
From Structure to Synergy: A Survey of Vision-Language Perception Paradigm Evolution in Multimodal Large Language Models AuroraCap: Efficient, Performant Video Detailed Captioning and a New Benchmark
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4a2ab20c-32cd-4207-afa5-4298ace80d12 · inbound
PercepCap: Video Captioner with Structured Spatio-Temporal Perception AuroraCap: Efficient, Performant Video Detailed Captioning and a New Benchmark
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2f5f4e4f-4e40-44e3-8325-75d037f4915d · inbound
Visual Token Compression Enhances Robustness of MLLMs AuroraCap: Efficient, Performant Video Detailed Captioning and a New Benchmark
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.