Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 9 inbound Pith citation observations for arXiv:2403.18406.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-08T15:31:16.929272Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-23T05:45:28.275913Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 02d99ae4-fa34-4897-85c4-fb20e8603cbb · inbound
PLLaVA : Parameter-free LLaVA Extension from Images to Videos for Video Dense Captioning An Image Grid Can Be Worth a Video: Zero-shot Video Question Answering Using a VLM
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 1a86423a-5ea2-446c-a0fa-169954ba3fae · inbound
MotionBench: Benchmarking and Improving Fine-grained Video Motion Understanding for Vision Language Models An Image Grid Can Be Worth a Video: Zero-shot Video Question Answering Using a VLM
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 1d5c33d3-2e3b-476c-bf38-ed8bb295810e · inbound
CoS: Chain-of-Shot Prompting for Long Video Understanding An Image Grid Can Be Worth a Video: Zero-shot Video Question Answering Using a VLM
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1697f5ac-2c46-4d4c-be1a-f72ab08548b7 · inbound
RTime-QA: A Benchmark for Atomic Temporal Event Understanding in Large Multi-modal Models An Image Grid Can Be Worth a Video: Zero-shot Video Question Answering Using a VLM
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 40d3da12-d19b-4694-bdbe-5413a14da1da · inbound
LeAdQA: LLM-Driven Context-Aware Temporal Grounding for Video Question Answering An Image Grid Can Be Worth a Video: Zero-shot Video Question Answering Using a VLM
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 86d87ae0-451a-468d-ba72-f19c6f13db1c · inbound
DynImg: Key Frames with Visual Prompts are Good Representation for Multi-Modal Video Understanding An Image Grid Can Be Worth a Video: Zero-shot Video Question Answering Using a VLM
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2bea321f-524f-4942-b37b-b774c9ab31f1 · inbound
Progressive Video Condensation with MLLM Agent for Long-form Video Understanding An Image Grid Can Be Worth a Video: Zero-shot Video Question Answering Using a VLM
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 71b9b300-e04e-4249-bfa0-b5a8df08b993 · inbound
Video Active Perception: Effective Inference-Time Long-Form Video Understanding with Vision-Language Models An Image Grid Can Be Worth a Video: Zero-shot Video Question Answering Using a VLM
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 920c2d81-8894-4271-bb3e-3b84f6186c3c · inbound
Enhancing Visual Token Representations for Video Large Language Models via Training-Free Spatial-Temporal Pooling and Gridding An Image Grid Can Be Worth a Video: Zero-shot Video Question Answering Using a VLM
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.