Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 14 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 10 inbound Pith citation observations for arXiv:2310.00653.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-11T20:24:58.819520Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
4
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation 217ca226-394d-49d1-82e7-fb462604f334 · inbound
MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models Reformulating Vision-Language Foundation Models and Datasets Towards Universal Multimodal Assistants
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 6aa60318-4518-4baf-b743-feb9d0bdf296 · inbound
Hallucination of Multimodal Large Language Models: A Survey Reformulating Vision-Language Foundation Models and Datasets Towards Universal Multimodal Assistants
Reference 197
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation b5641bb9-40a0-4262-8502-9a954cfc4b04 · inbound
MiniCPM-V: A GPT-4V Level MLLM on Your Phone Reformulating Vision-Language Foundation Models and Datasets Towards Universal Multimodal Assistants
Reference 110
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 0b33969a-66ec-4006-8894-d29737ff13b1 · inbound
NVILA: Efficient Frontier Visual Language Models Reformulating Vision-Language Foundation Models and Datasets Towards Universal Multimodal Assistants
Reference 130
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation c745772d-88be-4c47-9f00-6d7516d1e0ac · inbound
Learning to Correction: Explainable Feedback Generation for Visual Commonsense Reasoning Distractor Reformulating Vision-Language Foundation Models and Datasets Towards Universal Multimodal Assistants
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1edcaefa-fd54-4104-8ec9-14bc52363ac6 · inbound
CHiP: Cross-modal Hierarchical Direct Preference Optimization for Multimodal LLMs Reformulating Vision-Language Foundation Models and Datasets Towards Universal Multimodal Assistants
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 82484be8-65d5-475b-8c38-226399ffc1f5 · inbound
From Visuals to Vocabulary: Establishing Equivalence Between Image and Text Token Through Autoregressive Pre-training in MLLMs Reformulating Vision-Language Foundation Models and Datasets Towards Universal Multimodal Assistants
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d24988dc-67bf-42b4-b1df-8f20cb1da434 · inbound
ASPO: Adaptive Sentence-Level Preference Optimization for Fine-Grained Multimodal Reasoning Reformulating Vision-Language Foundation Models and Datasets Towards Universal Multimodal Assistants
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f30a3137-a4df-42e0-871f-31751ce3591c · inbound
Deep Pre-Alignment for VLMs Reformulating Vision-Language Foundation Models and Datasets Towards Universal Multimodal Assistants
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 965a76ad-644f-4ffc-87e6-c643293715e6 · inbound
Task-Aware Structured Memory for Dynamic Multi-modal In-Context Learning Reformulating Vision-Language Foundation Models and Datasets Towards Universal Multimodal Assistants
Reference 168
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.