Pith. sign in

Paper Citation Record · LEDGER

mPLUG-DocOwl2: High-resolution Compressing for OCR-free Multi-page Document Understanding

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 16 inbound Pith citation observations for arXiv:2409.03420.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2409.03420 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 16 of 16 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 16 of 16 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T22:34:37.142434Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-30T13:24:40.403192Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 851bce59-32ba-48a6-b376-3f99419b21cb · inbound

MinerU: An Open-Source Solution for Precise Document Content Extraction cites this paper.

MinerU: An Open-Source Solution for Precise Document Content Extraction mPLUG-DocOwl2: High-resolution Compressing for OCR-free Multi-page Document Understanding

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-16T04:00:25.813195Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-16T04:00:25.624430Z digest=sha256:2754df715c05229cf8f4c5313454f2a0643365ed82d9bedeeb83077abd6e8df0

Observation a1bac248-8a5a-448f-9e6a-43dcbaa7a470 · inbound

VisRAG: Vision-based Retrieval-augmented Generation on Multi-modality Documents cites this paper.

VisRAG: Vision-based Retrieval-augmented Generation on Multi-modality Documents mPLUG-DocOwl2: High-resolution Compressing for OCR-free Multi-page Document Understanding

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-16T15:37:25.887847Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-16T15:37:25.781240Z digest=sha256:19b88ba618eb5aa7d32d31fc6b943ebbfd3da83912c1cb147623f669ad565c13

Observation eb5d99ca-3ca3-4ed2-ac21-55e664e21c4e · inbound

Document Parsing Unveiled: Techniques, Challenges, and Prospects for Structured Information Extraction cites this paper.

Document Parsing Unveiled: Techniques, Challenges, and Prospects for Structured Information Extraction mPLUG-DocOwl2: High-resolution Compressing for OCR-free Multi-page Document Understanding

Reference 83

Resolution
verified exact
arxiv_id, observed 2026-05-23T19:15:47.227275Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-23T19:15:21.695801Z digest=sha256:f2136d787240f437960616b684753280738155b31e02e180a5eaa2f2c5ccae69

Observation 639f9c5a-3974-4c55-a9ca-946c3a22b8e0 · inbound

OCRBench v2: An Improved Benchmark for Evaluating Large Multimodal Models on Visual Text Localization and Reasoning cites this paper.

OCRBench v2: An Improved Benchmark for Evaluating Large Multimodal Models on Visual Text Localization and Reasoning mPLUG-DocOwl2: High-resolution Compressing for OCR-free Multi-page Document Understanding

Reference 148

Resolution
verified exact
arxiv_id, observed 2026-05-17T20:33:26.917214Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-17T20:33:26.613927Z digest=sha256:0e244c11983c02ca3d6789ebb2aa822d6960c0a7b62ff0d9e8e4fceb001e8439

Observation d5c9bf1e-a4e8-4fb4-b9f3-a137212ed0c2 · inbound

DRISHTIKON: Visual Grounding at Multiple Granularities in Documents cites this paper.

DRISHTIKON: Visual Grounding at Multiple Granularities in Documents mPLUG-DocOwl2: High-resolution Compressing for OCR-free Multi-page Document Understanding

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T22:34:37.142434Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:34:37.142434Z digest=sha256:ac000ea1d1235372a2646e968be95997e13a8c31c7e39aec36ebcf1ea1124810

Observation 1811b313-2de0-4609-84e8-33bc76d5fa5e · inbound

Improving MLLM's Document Image Machine Translation via Synchronously Self-reviewing Its OCR Proficiency cites this paper.

Improving MLLM's Document Image Machine Translation via Synchronously Self-reviewing Its OCR Proficiency mPLUG-DocOwl2: High-resolution Compressing for OCR-free Multi-page Document Understanding

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T18:28:11.394510Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:28:11.394510Z digest=sha256:6c5cffa95a415c91c5569bc0f1b63299e0a94c500036c0eed83b6bd558b12d0d

Observation faf53257-4eb2-4980-a638-9bf324d05f69 · inbound

A Survey on MLLM-based Visually Rich Document Understanding: Methods, Challenges, and Emerging Trends cites this paper.

A Survey on MLLM-based Visually Rich Document Understanding: Methods, Challenges, and Emerging Trends mPLUG-DocOwl2: High-resolution Compressing for OCR-free Multi-page Document Understanding

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-19T04:42:04.341361Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-19T04:38:49.512293Z digest=sha256:2d85157f563825831ad252f0a538793742fdebe537a217e61894635f1471fc19

Observation 1c5fe10e-e0f1-413f-b1ee-b83d1ece2d6d · inbound

ExpliCIT-QA: Explainable Code-Based Image Table Question Answering cites this paper.

ExpliCIT-QA: Explainable Code-Based Image Table Question Answering mPLUG-DocOwl2: High-resolution Compressing for OCR-free Multi-page Document Understanding

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T17:08:15.333002Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:08:15.333002Z digest=sha256:004c1ab31922e998e1dd4d7fbf0ada5f9b3c0f39a1fef4254384d4c4a3ae37b2

Observation 1e4266db-48e4-4e6e-90ed-a8f540d82797 · inbound

Spatially Grounded Explanations in Vision Language Models for Document Visual Question Answering cites this paper.

Spatially Grounded Explanations in Vision Language Models for Document Visual Question Answering mPLUG-DocOwl2: High-resolution Compressing for OCR-free Multi-page Document Understanding

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T17:07:52.666510Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:07:52.666510Z digest=sha256:31eac5212c534a419fd4e2e661a6eb40b8583c30a0a82346954dd4970291cbfa

Observation e5814c65-4bc7-47c9-818f-92d8ba83cbed · inbound

Docopilot: Improving Multimodal Models for Document-Level Understanding cites this paper.

Docopilot: Improving Multimodal Models for Document-Level Understanding mPLUG-DocOwl2: High-resolution Compressing for OCR-free Multi-page Document Understanding

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T15:57:01.022352Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:57:01.022352Z digest=sha256:be20134ea04fb12684966abd2ccd18e9115ccbb667fe348f3794f86b95725cf6

Observation 2508b251-1ed1-4cf8-87d8-3812321e215f · inbound

Survey of Specialized Large Language Model cites this paper.

Survey of Specialized Large Language Model mPLUG-DocOwl2: High-resolution Compressing for OCR-free Multi-page Document Understanding

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-05T15:37:49.232031Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:37:49.232031Z digest=sha256:183632812d92dfe432cfec680f9958dbf081a06c8637a16ecbe261f3848bac10

Observation 8ffc5b73-2fb0-4973-9181-d92aa4eb664e · inbound

RPC-Bench: A Fine-grained Benchmark for Research Paper Comprehension cites this paper.

RPC-Bench: A Fine-grained Benchmark for Research Paper Comprehension mPLUG-DocOwl2: High-resolution Compressing for OCR-free Multi-page Document Understanding

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T14:48:00.804059Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-16T14:45:20.535179Z digest=sha256:9044b414094598dade45344844869323c3e5783a462c62c2113e041b3dec4008

Observation b95467a6-2bbb-4148-8252-87686a9f9a15 · inbound

MoDora: Tree-Based Semi-Structured Document Analysis System cites this paper.

MoDora: Tree-Based Semi-Structured Document Analysis System mPLUG-DocOwl2: High-resolution Compressing for OCR-free Multi-page Document Understanding

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-15T19:10:15.582956Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-15T19:09:41.482713Z digest=sha256:6f1eac92f04b6bf9f3e82186b512d4633b884713cf9870a4b5f0fa76e651b87e

Observation 13349e51-5bd2-4dc0-bb99-3977a5cabfc5 · inbound

Towards Real-World Document Parsing via Realistic Scene Synthesis and Document-Aware Training cites this paper.

Towards Real-World Document Parsing via Realistic Scene Synthesis and Document-Aware Training mPLUG-DocOwl2: High-resolution Compressing for OCR-free Multi-page Document Understanding

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-15T01:18:26.617773Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-15T01:15:26.757215Z digest=sha256:8510a408cd29af8b456cc8429c58799d74473bb05041858cedf744fb119d73e1

Observation b3fde4fd-881b-4d17-bee6-89b7105f20b6 · inbound

Unveil: Unified Visual-Textual Integration and Distillation for Multi-modal Document Retrieval cites this paper.

Unveil: Unified Visual-Textual Integration and Distillation for Multi-modal Document Retrieval mPLUG-DocOwl2: High-resolution Compressing for OCR-free Multi-page Document Understanding

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-06-30T13:24:40.404764Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-30T13:17:04.441743Z digest=sha256:af3417ce30fefa67d0169c18032b0724d3706836e668446ff8b7281d16dd7caa

Observation e796fad5-1ef1-43a3-ba26-0d6cd142a45b · inbound

BVS: Bayesian Visual Search with Multimodal Large Language Model for Fine-grained Perception cites this paper.

BVS: Bayesian Visual Search with Multimodal Large Language Model for Fine-grained Perception mPLUG-DocOwl2: High-resolution Compressing for OCR-free Multi-page Document Understanding

Reference 190

Resolution
unresolved
no resolver link, observed 2026-07-12T04:17:40.198357Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-12T04:17:40.198357Z digest=sha256:b1546b17160429b39ae6e7dfe80929273a07e48984a027d4cc25f34ae585cd7f