Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 16 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 11 inbound Pith citation observations for arXiv:2305.13786.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-15T20:34:12.843409Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-24T05:03:55.426403Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 282c0876-504b-4e4d-aacd-3e1f4b66717f · inbound
Gemini: A Family of Highly Capable Multimodal Models Perception Test: A Diagnostic Benchmark for Multimodal Video Models
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 996a85db-de6c-4d09-933c-9919afc93569 · inbound
TempCompass: Do Video LLMs Really Understand Videos? Perception Test: A Diagnostic Benchmark for Multimodal Video Models
Reference 113
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 52d23f46-1c7e-4ab5-ada8-1cdc7367bec6 · inbound
MAmmoTH-VL: Eliciting Multimodal Reasoning with Instruction Tuning at Scale Perception Test: A Diagnostic Benchmark for Multimodal Video Models
Reference 2012
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a9b6ac0b-271d-4707-9b02-efa9b456f11e · inbound
Movie2Story: A framework for understanding videos and telling stories in the form of novel text Perception Test: A Diagnostic Benchmark for Multimodal Video Models
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3630f9cd-e051-4639-8f7b-db90f12e0eb9 · inbound
J-EDI QA: Benchmark for deep-sea organism-specific multimodal LLM Perception Test: A Diagnostic Benchmark for Multimodal Video Models
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 14ccbf6e-e7e4-41ea-8b8d-4b0709e9d7bf · inbound
Correspondence of high-dimensional emotion structures elicited by video clips between humans and Multimodal LLMs Perception Test: A Diagnostic Benchmark for Multimodal Video Models
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4a4ee60c-492e-4d88-a047-99ce02a09dfd · inbound
CausalVQA: A Physically Grounded Causal Reasoning Benchmark for Video Models Perception Test: A Diagnostic Benchmark for Multimodal Video Models
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bdd09533-6b82-46f0-921e-33d77498e756 · inbound
LaCo: Efficient Layer-wise Compression of Visual Tokens for Multimodal Large Language Models Perception Test: A Diagnostic Benchmark for Multimodal Video Models
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 66a69bad-3da5-4e39-a2cc-3e7d8433c4d8 · inbound
ReGATE: Learning Faster and Better with Fewer Tokens in MLLMs Perception Test: A Diagnostic Benchmark for Multimodal Video Models
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 7f192378-6656-46d9-82d4-cce3ea3690ab · inbound
Seeing More, Saying More: Lightweight Language Experts are Dynamic Video Token Compressors Perception Test: A Diagnostic Benchmark for Multimodal Video Models
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3c68a8dd-bae2-424e-9ef6-57b524686eb7 · inbound
HERMES: KV Cache as Hierarchical Memory for Efficient Streaming Video Understanding Perception Test: A Diagnostic Benchmark for Multimodal Video Models
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.