Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 10 inbound Pith citation observations for arXiv:2308.16890.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T18:23:49.558553Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-03T17:18:44.036310Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 504ac152-8454-4025-b9ba-0b8070a74063 · inbound
TempCompass: Do Video LLMs Really Understand Videos? TouchStone: Evaluating Vision-Language Models by Language Models
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 5c644f71-cdfb-41ec-a154-953b0256558e · inbound
MME-RealWorld: Could Your Multimodal LLM Challenge High-Resolution Real-World Scenarios that are Difficult for Humans? TouchStone: Evaluating Vision-Language Models by Language Models
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 2dcc7400-3811-4bfa-9905-b9f83fd47c1f · inbound
MM-RLHF: The Next Step Forward in Multimodal LLM Alignment TouchStone: Evaluating Vision-Language Models by Language Models
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f66346ef-ac07-4053-9dfd-1d66b8b07ad0 · inbound
P2P: Automated Paper-to-Poster Generation and Fine-Grained Benchmark TouchStone: Evaluating Vision-Language Models by Language Models
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cc05ce37-688a-45b6-a477-0b980139db4f · inbound
MINT: Multimodal Instruction Tuning with Multimodal Interaction Grouping TouchStone: Evaluating Vision-Language Models by Language Models
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8208d6aa-a0b3-4ca8-8279-123b8ef4fb3c · inbound
From Image Captioning to Visual Storytelling TouchStone: Evaluating Vision-Language Models by Language Models
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 767be77d-c9a4-44d1-87ef-859039bb6f60 · inbound
OxyEcomBench: Benchmarking Multimodal Foundation Models across E-Commerce Ecosystems TouchStone: Evaluating Vision-Language Models by Language Models
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 2cd4afe7-caa7-4bbb-97ad-3e8574ee3dcc · inbound
DisasterBench: A Multimodal Benchmark for UAV-Based Disaster Response in Complex Environments TouchStone: Evaluating Vision-Language Models by Language Models
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d0d7280d-7980-4aac-9987-55276a82775b · inbound
Qwen-RobotWorld Technical Report: Unifying Embodied World Modeling through Language-Conditioned Video Generation TouchStone: Evaluating Vision-Language Models by Language Models
Reference 94
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation eb30ef8c-b841-49a5-8ddb-3203d059bf3a · inbound
Qwen-Audio-VAE Technical Report TouchStone: Evaluating Vision-Language Models by Language Models
Reference 109
Source-reported events for the cited work
Unavailable: canonical work link unavailable.