Pith. sign in

Paper Citation Record · LEDGER

MDocAgent: A Multi-Modal Multi-Agent Framework for Document Understanding

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 12 inbound Pith citation observations for arXiv:2503.13964.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2503.13964 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 12 of 12 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 12 of 12 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:57:37.280223Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T03:37:35.692691Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation b1074efd-badc-440a-a15d-ab1799f567cf · inbound

From EduVisBench to EduVisAgent: A Benchmark and Multi-Agent Framework for Reasoning-Driven Pedagogical Visualization cites this paper.

From EduVisBench to EduVisAgent: A Benchmark and Multi-Agent Framework for Reasoning-Driven Pedagogical Visualization MDocAgent: A Multi-Modal Multi-Agent Framework for Document Understanding

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:37.280223Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:37.280223Z digest=sha256:e38571d7dbcd97aa3f09f5358bdfa7280edc333bf91023a84b302784a883c436

Observation 412043ba-676f-4dd1-b790-7765516916aa · inbound

Structured Attention Matters to Multimodal LLMs in Document Understanding cites this paper.

Structured Attention Matters to Multimodal LLMs in Document Understanding MDocAgent: A Multi-Modal Multi-Agent Framework for Document Understanding

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T23:47:37.986765Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:47:37.986765Z digest=sha256:796eb92410cd1bd403dc4c95740d28de4bfc01e0f394fa4110bc8d202edbacbf

Observation 873d134c-a90f-426f-aeeb-ed8110160504 · inbound

A Survey on MLLM-based Visually Rich Document Understanding: Methods, Challenges, and Emerging Trends cites this paper.

A Survey on MLLM-based Visually Rich Document Understanding: Methods, Challenges, and Emerging Trends MDocAgent: A Multi-Modal Multi-Agent Framework for Document Understanding

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-19T04:42:04.383328Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-19T04:38:49.512293Z digest=sha256:c506ddc39e879eead6b530444a6d71775a4b3d433d1f314a6502dcd922d0089e

Observation aa78b0f3-1d76-4993-8e3f-a5cf439dfde3 · inbound

Dual Latent Memory for Visual Multi-agent System cites this paper.

Dual Latent Memory for Visual Multi-agent System MDocAgent: A Multi-Modal Multi-Agent Framework for Document Understanding

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-03T06:04:51.137581Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T06:04:51.137581Z digest=sha256:dfd850665651434bbd289017e3ebf90cb257ad17e45648d2dcd67ba9ace52055

Observation 7d76c752-4654-40a7-b42e-adb102541d21 · inbound

DocPrune:Efficient Document Question Answering via Background, Question, and Comprehension-aware Token Pruning cites this paper.

DocPrune:Efficient Document Question Answering via Background, Question, and Comprehension-aware Token Pruning MDocAgent: A Multi-Modal Multi-Agent Framework for Document Understanding

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-11T19:06:09.384776Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-08T12:42:39.231468Z digest=sha256:2f8225aea5c290d67261e0d12d876dfe0f68e3134be36b4bf19d5bc055ad8d70

Observation f7309a69-6b88-4852-be51-a920ee220759 · inbound

Hierarchical Attacks for Multi-Modal Multi-Agent Reasoning cites this paper.

Hierarchical Attacks for Multi-Modal Multi-Agent Reasoning MDocAgent: A Multi-Modal Multi-Agent Framework for Document Understanding

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-14T20:19:27.908626Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-14T20:16:30.070819Z digest=sha256:ad9c5fba9c8e9876d9d0aa5c4f42d8d68ed74167716a58c7a9977c7be258910d

Observation 6ac94978-921b-4170-9234-04f3dc3e9f38 · inbound

EviProp: Seeded Relevance Diffusion on Chunk-Page Graphs for Long Multimodal Document Retrieval cites this paper.

EviProp: Seeded Relevance Diffusion on Chunk-Page Graphs for Long Multimodal Document Retrieval MDocAgent: A Multi-Modal Multi-Agent Framework for Document Understanding

Reference 9

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T03:37:35.694129Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-27T15:04:30.757297Z digest=sha256:2a6cb6bbb26b174667103452f2c4f9124a8b98d3caa6b86898e027b0c3241651

Observation 8113f3ec-f0ba-4a70-b410-772ffe9f1138 · inbound

Hybrid Retriever Evolution for Multimodal Document Reasoning Agents cites this paper.

Hybrid Retriever Evolution for Multimodal Document Reasoning Agents MDocAgent: A Multi-Modal Multi-Agent Framework for Document Understanding

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T07:04:21.625144Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-30T06:58:56.930942Z digest=sha256:17f01a36395332ecb108baa085fed2fca936d45fad27145aafb31984c0e87429

Observation 8f4d180f-4297-4370-86a0-2ba6247cd0a9 · inbound

Enhancing Large Multimodal Models in Key Information Extraction via Scene-Aware Document Synthesis cites this paper.

Enhancing Large Multimodal Models in Key Information Extraction via Scene-Aware Document Synthesis MDocAgent: A Multi-Modal Multi-Agent Framework for Document Understanding

Reference 16

Resolution
unresolved
no resolver link, observed 2026-07-11T16:02:00.920066Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T16:02:00.920066Z digest=sha256:24b777d19fc6fa572264ff503600dfcfae56c6603b05cf6e265ec925153e3c68

Observation 9278ccfa-f191-4192-a14d-eda8ac272860 · inbound

FinSAgent: Corpus-Aligned Multi-Agent RAG Framework for Evidence-Grounded SEC Filing Question Answering cites this paper.

FinSAgent: Corpus-Aligned Multi-Agent RAG Framework for Evidence-Grounded SEC Filing Question Answering MDocAgent: A Multi-Modal Multi-Agent Framework for Document Understanding

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-01T16:10:06.364874Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T16:10:06.364874Z digest=sha256:8724a691986bc2dacdd00e231bb9a9c9a987721cc36c4819fa6e2d2b44041d98

Observation d2c24b7c-cd98-49a0-b79b-a9c0118b92af · inbound

HierDoc: Hierarchical Page-to-Region Evidence Routing for Long-Document Visual Question Answering cites this paper.

HierDoc: Hierarchical Page-to-Region Evidence Routing for Long-Document Visual Question Answering MDocAgent: A Multi-Modal Multi-Agent Framework for Document Understanding

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-03T02:57:48.542979Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T02:57:48.542979Z digest=sha256:2075a102ced829d290cfcaaf980cb242d17cea2f5a795478be16fbae74475ff2

Observation c2f26311-8657-4584-a05d-a8255acd2ffa · inbound

XL-DocBench: Benchmarking Evidence-Grounded Extra-Long Document Understanding cites this paper.

XL-DocBench: Benchmarking Evidence-Grounded Extra-Long Document Understanding MDocAgent: A Multi-Modal Multi-Agent Framework for Document Understanding

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-04T01:39:07.462177Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T01:39:07.462177Z digest=sha256:285b812d627e6f5569a111cec19db1a356956b5d1254cd5e4564ad650707da8e