Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 26 inbound Pith citation observations for arXiv:2403.12895.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T06:02:30.234729Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-07-10T17:07:25.743722Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 9636f990-2707-4fc7-ac5b-fcd8b73293c8 · inbound
A Survey on Multimodal Large Language Models mPLUG-DocOwl 1.5: Unified Structure Learning for OCR-free Document Understanding
Reference 158
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 87856638-81ac-452d-8d4d-5731764e1050 · inbound
How Far Are We to GPT-4V? Closing the Gap to Commercial Multimodal Models with Open-Source Suites mPLUG-DocOwl 1.5: Unified Structure Learning for OCR-free Document Understanding
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 24372ced-89e2-4eb8-9ceb-ec8101d89e10 · inbound
InternLM-XComposer-2.5: A Versatile Large Vision Language Model Supporting Long-Contextual Input and Output mPLUG-DocOwl 1.5: Unified Structure Learning for OCR-free Document Understanding
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 20d6da08-e96d-4de7-90c8-56bcb0da7189 · inbound
MME-RealWorld: Could Your Multimodal LLM Challenge High-Resolution Real-World Scenarios that are Difficult for Humans? mPLUG-DocOwl 1.5: Unified Structure Learning for OCR-free Document Understanding
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 29680334-4140-4af1-ae06-56f51675dc11 · inbound
General OCR Theory: Towards OCR-2.0 via a Unified End-to-end Model mPLUG-DocOwl 1.5: Unified Structure Learning for OCR-free Document Understanding
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c2efe267-4983-441c-a96f-21bb893de652 · inbound
MinerU: An Open-Source Solution for Precise Document Content Extraction mPLUG-DocOwl 1.5: Unified Structure Learning for OCR-free Document Understanding
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 20b7ed55-3765-49f7-9512-ba4426200797 · inbound
PyramidDrop: Accelerating Your Large Vision-Language Models via Pyramid Visual Redundancy Reduction mPLUG-DocOwl 1.5: Unified Structure Learning for OCR-free Document Understanding
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0bb3dfa2-7fa4-495d-bb7d-4f3df067811d · inbound
Document Parsing Unveiled: Techniques, Challenges, and Prospects for Structured Information Extraction mPLUG-DocOwl 1.5: Unified Structure Learning for OCR-free Document Understanding
Reference 82
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 36eb7878-7b97-440b-86f7-8d44812fc120 · inbound
Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling mPLUG-DocOwl 1.5: Unified Structure Learning for OCR-free Document Understanding
Reference 93
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a94e4643-1970-4d16-8da9-39b3ed21684b · inbound
OCRBench v2: An Improved Benchmark for Evaluating Large Multimodal Models on Visual Text Localization and Reasoning mPLUG-DocOwl 1.5: Unified Structure Learning for OCR-free Document Understanding
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 13ce348c-6f4e-4ff0-a75c-36eed2ec4bee · inbound
CoMemo: LVLMs Need Image Context with Image Memory mPLUG-DocOwl 1.5: Unified Structure Learning for OCR-free Document Understanding
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aa7ec550-3dc2-4674-a241-f86be58475fe · inbound
GenRecal: Generation after Recalibration from Large to Small Vision-Language Models mPLUG-DocOwl 1.5: Unified Structure Learning for OCR-free Document Understanding
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 57b3f667-02d7-41d7-844d-fc169ac4be3f · inbound
Structured Attention Matters to Multimodal LLMs in Document Understanding mPLUG-DocOwl 1.5: Unified Structure Learning for OCR-free Document Understanding
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 29ef0fcd-2bc9-4672-916c-740cfec8368e · inbound
MusiXQA: Advancing Visual Music Understanding in Multimodal Large Language Models mPLUG-DocOwl 1.5: Unified Structure Learning for OCR-free Document Understanding
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 28ac8605-6fd0-46f2-bb60-72bcd742971a · inbound
ESTR-CoT: Towards Explainable and Accurate Event Stream based Scene Text Recognition with Chain-of-Thought Reasoning mPLUG-DocOwl 1.5: Unified Structure Learning for OCR-free Document Understanding
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 310cdb34-bc96-4d48-ab61-5808c1d8a2b8 · inbound
Single-to-mix Modality Alignment with Multimodal Large Language Model for Document Image Machine Translation mPLUG-DocOwl 1.5: Unified Structure Learning for OCR-free Document Understanding
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3cb95d74-63eb-4d1c-bd59-1439d6c85896 · inbound
A document is worth a structured record: Principled inductive bias design for document recognition mPLUG-DocOwl 1.5: Unified Structure Learning for OCR-free Document Understanding
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2921b240-0064-4586-9758-4014f25367d4 · inbound
HRSeg: High-Resolution Visual Perception and Enhancement for Reasoning Segmentation mPLUG-DocOwl 1.5: Unified Structure Learning for OCR-free Document Understanding
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7dd154c6-d455-4066-8564-f221c43666c4 · inbound
ChartAgent: A Multimodal Agent for Visually Grounded Reasoning in Complex Chart Question Answering mPLUG-DocOwl 1.5: Unified Structure Learning for OCR-free Document Understanding
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation acea7df4-fe20-4d75-9810-9ed67ee89d3c · inbound
CodeOCR: On the Effectiveness of Vision Language Models in Code Understanding mPLUG-DocOwl 1.5: Unified Structure Learning for OCR-free Document Understanding
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation dfca068a-ce1c-4348-87f7-f1a55734cadc · inbound
HART: High-Resolution Annotation-Free Reasoning Technique through a Closed-loop Framework mPLUG-DocOwl 1.5: Unified Structure Learning for OCR-free Document Understanding
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2ba7f8f2-2ea6-4e31-99fd-aa52ca51d8e0 · inbound
Q-Mask: Query-driven Causal Masks for Text Anchoring in OCR-Oriented Vision-Language Models mPLUG-DocOwl 1.5: Unified Structure Learning for OCR-free Document Understanding
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8ed5f5c4-6e72-483c-9ba1-909a80bff665 · inbound
SinkRouter: Sink-Aware Routing for Efficient Long-Context Decoding in Large Language and Multimodal Models mPLUG-DocOwl 1.5: Unified Structure Learning for OCR-free Document Understanding
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9c6b683e-1544-444d-9dd6-40278a803d47 · inbound
CiteVQA: Benchmarking Evidence Attribution for Trustworthy Document Intelligence mPLUG-DocOwl 1.5: Unified Structure Learning for OCR-free Document Understanding
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6492d638-4007-49ba-a24b-3196fe499c0d · inbound
Infinity-Parser2 Technical Report mPLUG-DocOwl 1.5: Unified Structure Learning for OCR-free Document Understanding
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ccfee072-f5e6-40a4-be92-bdac0ad074c8 · inbound
Infinity-Parser2 Technical Report mPLUG-DocOwl 1.5: Unified Structure Learning for OCR-free Document Understanding
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.