Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 22 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 31 inbound Pith citation observations for arXiv:2309.11419.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-12T12:52:10.280261Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-07-10T13:27:05.552868Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation d71fdbcf-0d36-4319-b658-ea545a8cba30 · inbound
How Far Are We to GPT-4V? Closing the Gap to Commercial Multimodal Models with Open-Source Suites KOSMOS-2.5: A Multimodal Literate Model
Reference 77
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 01e0e38b-26fe-4082-b818-ccec1f9638a5 · inbound
MinerU: An Open-Source Solution for Precise Document Content Extraction KOSMOS-2.5: A Multimodal Literate Model
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 227ab008-0fa0-484c-aa81-4b0e5b2a84ea · inbound
SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE KOSMOS-2.5: A Multimodal Literate Model
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 55824c0c-9478-463f-8219-ebe694609b52 · inbound
DOGR: Towards Versatile Visual Document Grounding and Referring KOSMOS-2.5: A Multimodal Literate Model
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ec05580b-55f9-47b3-8f33-4751a25e6efa · inbound
ChatRex: Taming Multimodal LLM for Joint Perception and Understanding KOSMOS-2.5: A Multimodal Literate Model
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 92df3bc4-0195-4ce9-bdc4-f6412729c6ed · inbound
CC-OCR: A Comprehensive and Challenging OCR Benchmark for Evaluating Large Multimodal Models in Literacy KOSMOS-2.5: A Multimodal Literate Model
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d14e1eef-9b97-4868-8462-b5c5073baecc · inbound
AV-Odyssey Bench: Can Your Multimodal LLMs Really Understand Audio-Visual Information? KOSMOS-2.5: A Multimodal Literate Model
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6c523022-cb1e-44f8-b80a-60d4e97f94e9 · inbound
V2PE: Improving Multimodal Long-Context Capability of Vision-Language Models with Variable Visual Position Encoding KOSMOS-2.5: A Multimodal Literate Model
Reference 77
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ee1b0aa8-656e-4374-b20a-32bffae4b191 · inbound
InstructOCR: Instruction Boosting Scene Text Spotting KOSMOS-2.5: A Multimodal Literate Model
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9eca4f22-337d-46d1-8a28-0836cec76d99 · inbound
Next Token Prediction Towards Multimodal Intelligence: A Comprehensive Survey KOSMOS-2.5: A Multimodal Literate Model
Reference 285
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 46f1f4dd-0b04-4f43-867f-71dad31aed6a · inbound
Survey on Question Answering over Visually Rich Documents: Methods, Challenges, and Trends KOSMOS-2.5: A Multimodal Literate Model
Reference 69
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f092e221-6002-4759-95c0-d245e880bfbc · inbound
\'Eclair -- Extracting Content and Layout with Integrated Reading Order for Documents KOSMOS-2.5: A Multimodal Literate Model
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c6d922dc-292b-4119-a33a-29d6dcdd10b1 · inbound
MuDoC: An Interactive Multimodal Document-grounded Conversational AI System KOSMOS-2.5: A Multimodal Literate Model
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 875dff2a-34fb-4356-b4d9-97215fd81879 · inbound
Dolphin: Document Image Parsing via Heterogeneous Anchor Prompting KOSMOS-2.5: A Multimodal Literate Model
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e759238f-dab3-4c82-8923-2bcb3f9e021a · inbound
Remote Sensing Large Vision-Language Model: Semantic-augmented Multi-level Alignment and Semantic-aware Expert Modeling KOSMOS-2.5: A Multimodal Literate Model
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 195cbc1e-4303-4036-aab0-dde0d903b056 · inbound
DREAM: Document Reconstruction via End-to-end Autoregressive Model KOSMOS-2.5: A Multimodal Literate Model
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 89cb16e3-c619-4a7a-a2b7-0f6800ab48c9 · inbound
A Survey on MLLM-based Visually Rich Document Understanding: Methods, Challenges, and Emerging Trends KOSMOS-2.5: A Multimodal Literate Model
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 93b1e719-603b-4d90-b458-ff54a1b5a326 · inbound
From Plausibility to Verifiability: Risk-Controlled Generative OCR with Vision-Language Models KOSMOS-2.5: A Multimodal Literate Model
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation ac398883-1f0d-41be-9a61-1671c3d694c7 · inbound
ShredBench: Evaluating the Semantic Reasoning Capabilities of Multimodal LLMs in Document Reconstruction KOSMOS-2.5: A Multimodal Literate Model
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 5e2fee03-ac49-4edd-807d-a17ef55f27ec · inbound
CC-OCR V2: Benchmarking Large Multimodal Models for Literacy in Real-world Document Processing KOSMOS-2.5: A Multimodal Literate Model
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation ccfa463d-4a18-4714-8688-5414e90469d6 · inbound
Towards Characterizing Scientific Image Utility and Upgradability KOSMOS-2.5: A Multimodal Literate Model
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation c7a1dff8-f4d8-44f0-9374-2abae61392f0 · inbound
Vision Language Model Helps Private Information De-Identification in Vision Data KOSMOS-2.5: A Multimodal Literate Model
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 7ef6317b-b9f2-48b7-93ea-d1db3b4d17dc · inbound
Cross-Temporal Sinhala OCR: Page-Level Adaptation and Diachronic Analysis KOSMOS-2.5: A Multimodal Literate Model
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation aa2516aa-7cdf-4590-b65a-4475a4517604 · inbound
Semantic-Guided Reading Order Reconstruction in Historical Armenian Newspapers with LLMs KOSMOS-2.5: A Multimodal Literate Model
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation a1da6840-be94-461b-babc-ab510dd557ab · inbound
MORE: A Multilingual Document Parsing Benchmark and Evaluation KOSMOS-2.5: A Multimodal Literate Model
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9127f1fe-4245-49f1-9952-7e6a7468c280 · inbound
Rethinking Small VLM Quantization: From Component-Wise Analysis to Hardware-Aware Edge Deployment KOSMOS-2.5: A Multimodal Literate Model
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation ad987fdd-86ef-40e7-a4a4-e14abcc0413e · inbound
Multi-Expert Routing for Multi-Domain Low-Resource OCR: A Manchu Case Study KOSMOS-2.5: A Multimodal Literate Model
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation af234a1e-faf5-40b5-8d18-930e6cf4014f · inbound
Stop Thinking, Start Looking: Efficient Post-Training for Multimodal Document Question Answering via Reasoning-Free Alignment KOSMOS-2.5: A Multimodal Literate Model
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6fcf668d-273b-4f2b-a57d-2c8e85e97bd3 · inbound
DocOCR-Eval: A Correction-Based Framework for OCR Tool Selection Without Ground Truth KOSMOS-2.5: A Multimodal Literate Model
Reference 174
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation daa62925-ab67-491f-b611-66f55496c011 · inbound
DocPO: Advancing Document Policy Optimization via Tailored Step-Aware Rewards KOSMOS-2.5: A Multimodal Literate Model
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9f2100cf-6c2e-4745-baf1-9640a5a0ec1d · inbound
DocPO: Advancing Document Policy Optimization via Tailored Step-Aware Rewards KOSMOS-2.5: A Multimodal Literate Model
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.