Pith. sign in

Paper Citation Record · LEDGER

GDP.pdf: Benchmarking Grounded Multimodal Reasoning over Professional PDF Documents

As of 9 August 2026, this Paper Citation Record lists 40 of 40 outbound references and 0 inbound Pith citation observations for arXiv:2607.11192.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.11192 v3

Coverage vector

measured 40 of 40 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-02T06:59:50.917441Z

measured 40 of 40 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

40 of 40 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved40
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 1503f42d-2456-4b72-a690-f76d13f48eaf · outbound

This paper cites an unresolved cited work.

GDP.pdf: Benchmarking Grounded Multimodal Reasoning over Professional PDF Documents Unresolved cited work

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-02T06:59:47.512163Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:59:47.512163Z digest=sha256:09583f7e2c2aaee6690601e84fe3971b0f85189b0a49859980d8f8cee0e71b9c

Observation ce926125-01ae-4ea7-bec4-e6c04aaa7034 · outbound

This paper cites MEGA-Bench: Scaling multimodal evaluation to over 500 real-world tasks.

GDP.pdf: Benchmarking Grounded Multimodal Reasoning over Professional PDF Documents MEGA-Bench: Scaling multimodal evaluation to over 500 real-world tasks

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-02T06:59:47.618955Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:59:47.618955Z digest=sha256:56fecb3f9d19373143ad9d861a3e5a80c834ef1af017b6218cccdd48db284c76

Observation f42f38fe-2bdb-4d37-bee1-191651d23c39 · outbound

This paper cites HybridQA: A dataset of multi-hop question answering over tabular and textual data.

GDP.pdf: Benchmarking Grounded Multimodal Reasoning over Professional PDF Documents HybridQA: A dataset of multi-hop question answering over tabular and textual data

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-02T06:59:47.741062Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:59:47.741062Z digest=sha256:3497b5e597b7a11ea088897737834344481c0cf75a088d83ebe5f10ef3841ce7

Observation ae972ef0-121a-4b3a-97da-fa86effe503f · outbound

This paper cites RoDLA: Benchmarking the robustness of document layout analysis models.

GDP.pdf: Benchmarking Grounded Multimodal Reasoning over Professional PDF Documents RoDLA: Benchmarking the robustness of document layout analysis models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-02T06:59:47.818325Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:59:47.818325Z digest=sha256:c7e05a0a6e52dc796c2c005511db8175962fa073101de236e4e505204e83e650

Observation fa97481b-8b08-42f0-bdd4-ae2a715a85e0 · outbound

This paper cites FinQA: A dataset of numerical reasoning over financial data.

GDP.pdf: Benchmarking Grounded Multimodal Reasoning over Professional PDF Documents FinQA: A dataset of numerical reasoning over financial data

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-02T06:59:47.887088Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:59:47.887088Z digest=sha256:ae254e710b054f12df0c305d32cdfa75d32283db08390ce6a2147587326daa10

Observation 1758d779-b0e3-46b8-820f-2fa1926ddd01 · outbound

This paper cites M-LongDoc: A benchmark for multimodal super- long document understanding and a retrieval-aware tuning framework.

GDP.pdf: Benchmarking Grounded Multimodal Reasoning over Professional PDF Documents M-LongDoc: A benchmark for multimodal super- long document understanding and a retrieval-aware tuning framework

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-02T06:59:47.964650Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:59:47.964650Z digest=sha256:dd8d2e27fe4babc227c0e8500c043543fe6b649c1ff5f821e92c1dab57ecaace

Observation 0813694c-b113-495d-8769-8033b83f521c · outbound

This paper cites LongDocURL: a comprehensive multimodal long document benchmark integrating understanding, reason- ing, and locating.

GDP.pdf: Benchmarking Grounded Multimodal Reasoning over Professional PDF Documents LongDocURL: a comprehensive multimodal long document benchmark integrating understanding, reason- ing, and locating

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-02T06:59:48.084336Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:59:48.084336Z digest=sha256:9ac959bfabacaafea2a47bdf1fc02fd27e614c638292f7c0175e8c88d25111cd

Observation 87d17363-7126-457e-a2f9-3888a67aab7e · outbound

This paper cites Benchmarking retrieval- augmented multimodal generation for document question answering.

GDP.pdf: Benchmarking Grounded Multimodal Reasoning over Professional PDF Documents Benchmarking retrieval- augmented multimodal generation for document question answering

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-02T06:59:48.175604Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:59:48.175604Z digest=sha256:409b45a98803205035929bc4f20517c132d1e9e2fd9f09d9fa92d2a235c1e82b

Observation 649cf94f-ccb4-4c6c-acac-1efd58691928 · outbound

This paper cites OCRBench v2: An improved benchmark for evaluating large multimodal models on visual text localization and reason- ing.

GDP.pdf: Benchmarking Grounded Multimodal Reasoning over Professional PDF Documents OCRBench v2: An improved benchmark for evaluating large multimodal models on visual text localization and reason- ing

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-02T06:59:48.265432Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:59:48.265432Z digest=sha256:d027b6d9bcfaaf94c33a23d9d2eae3308cfe624a20ff3a2ee74aad231158e914

Observation bd9df8c1-19e7-4ae0-b385-874eff285fcd · outbound

This paper cites Ho, Christopher R ´e, Adam Chilton, Aditya Narayana, Alex Chohlas-Wood, Austin Peters, Brandon Waldon, Daniel N.

GDP.pdf: Benchmarking Grounded Multimodal Reasoning over Professional PDF Documents Ho, Christopher R ´e, Adam Chilton, Aditya Narayana, Alex Chohlas-Wood, Austin Peters, Brandon Waldon, Daniel N

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-02T06:59:48.354389Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:59:48.354389Z digest=sha256:80aa61a4788c17b9347b133a332963457693b46ab74cc6633e1204766d672749

Observation de12bad6-9928-4403-a343-baeb8b5b53dc · outbound

This paper cites FinanceBench: A New Benchmark for Financial Question Answering.

GDP.pdf: Benchmarking Grounded Multimodal Reasoning over Professional PDF Documents FinanceBench: A New Benchmark for Financial Question Answering

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-02T06:59:48.470437Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:59:48.470437Z digest=sha256:a8b4da6dd12e1a9d4b4520bc2cff9db76b7134bac28346641a3d430370bb1dc4

Observation aea17296-4621-43b4-896b-90b57445ca0e · outbound

This paper cites FUNSD: A dataset for form understanding in noisy scanned documents.

GDP.pdf: Benchmarking Grounded Multimodal Reasoning over Professional PDF Documents FUNSD: A dataset for form understanding in noisy scanned documents

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-02T06:59:48.563708Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:59:48.563708Z digest=sha256:4ff1d4430418c1dd05486c074be5575ba1509ca1c9ae3c2d4c0b1bf4807adb84

Observation 7a659499-5611-41bc-b05c-349e5c7a6e2a · outbound

This paper cites an unresolved cited work.

GDP.pdf: Benchmarking Grounded Multimodal Reasoning over Professional PDF Documents Unresolved cited work

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-02T06:59:48.656576Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:59:48.656576Z digest=sha256:fc11a85d26be2a09ca4307a48ce03255c58f131cfe7c6f2dccec2a2a41139c9a

Observation dcd29ea2-af0e-4f38-8116-a7ec9fcde986 · outbound

This paper cites OCRBench: On the hidden mystery of OCR in large multimodal models.Science China Information Sciences, 67(12):220102, 2024.

GDP.pdf: Benchmarking Grounded Multimodal Reasoning over Professional PDF Documents OCRBench: On the hidden mystery of OCR in large multimodal models.Science China Information Sciences, 67(12):220102, 2024

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-02T06:59:48.744101Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:59:48.744101Z digest=sha256:5dbf1b92569ef31219ee56582dd5d7a0fbf9d2118915372c2cd9882c11638fc2

Observation 0e8bb9bc-eeca-4549-a62e-76d8d84dc41d · outbound

This paper cites MathVista: Evaluating mathemat- ical reasoning of foundation models in visual contexts.

GDP.pdf: Benchmarking Grounded Multimodal Reasoning over Professional PDF Documents MathVista: Evaluating mathemat- ical reasoning of foundation models in visual contexts

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-02T06:59:48.834842Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:59:48.834842Z digest=sha256:82db293550259e8964fadf6ad728e339e1d7dabc551de3f369427c399c21b12c

Observation f61b70dc-e9b4-4801-bc6f-05934a7a1bdd · outbound

This paper cites MMLongBench-Doc: Bench- marking long-context document understanding with visualiza- tions.

GDP.pdf: Benchmarking Grounded Multimodal Reasoning over Professional PDF Documents MMLongBench-Doc: Bench- marking long-context document understanding with visualiza- tions

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-02T06:59:48.923992Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:59:48.923992Z digest=sha256:45fc57f8a53270ae3f093e7f1e62f86f4714ea8da7c5218e93cb540a3087b8ae

Observation 0af192d1-7a62-4dd6-b9fe-94530adde980 · outbound

This paper cites OK-VQA: A visual question answering benchmark requiring external knowledge.

GDP.pdf: Benchmarking Grounded Multimodal Reasoning over Professional PDF Documents OK-VQA: A visual question answering benchmark requiring external knowledge

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-02T06:59:48.992266Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:59:48.992266Z digest=sha256:f7ee9784c1034492760d3b3ca31563e13c5efde05db59a7c2523b8b127754ca8

Observation 8b872526-b1bf-4200-bc91-220acd6c05d1 · outbound

This paper cites ChartQA: A benchmark for question answering about charts with visual and logical reasoning.

GDP.pdf: Benchmarking Grounded Multimodal Reasoning over Professional PDF Documents ChartQA: A benchmark for question answering about charts with visual and logical reasoning

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-02T06:59:49.063092Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:59:49.063092Z digest=sha256:c754ed4cfaa75cc2b39e54260bdd48dbe1c61b6ad8c2c2d210a5d6c61ba0345d

Observation f486bcc6-a4bd-4e26-b4e0-0a7a3b1624b3 · outbound

This paper cites ChartQAPro: A more diverse and challenging benchmark for chart question answering.

GDP.pdf: Benchmarking Grounded Multimodal Reasoning over Professional PDF Documents ChartQAPro: A more diverse and challenging benchmark for chart question answering

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-02T06:59:49.124703Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:59:49.124703Z digest=sha256:1bf9de7e478f96598e7f72c2334930440509f69b25f1c1b06c63c298b27f5bb2

Observation 84d2dff2-ebdc-4450-8bc6-b2df960b5754 · outbound

This paper cites an unresolved cited work.

GDP.pdf: Benchmarking Grounded Multimodal Reasoning over Professional PDF Documents Unresolved cited work

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-02T06:59:49.229554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:59:49.229554Z digest=sha256:20a3b5b393dec0be646a13ef694abf92bccdc74937010eb35d113f663d519363

Observation 3d3e52b5-44f5-44ea-a4d3-1998f019f6d4 · outbound

This paper cites an unresolved cited work.

GDP.pdf: Benchmarking Grounded Multimodal Reasoning over Professional PDF Documents Unresolved cited work

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-02T06:59:49.315744Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:59:49.315744Z digest=sha256:f8270727755f57b9cfd445719bb3c4995e2874af148e5043b4211cd62f0294a7

Observation 92254d90-4566-446c-9797-8bb1fda97fae · outbound

This paper cites Khapra, and Pratyush Kumar.

GDP.pdf: Benchmarking Grounded Multimodal Reasoning over Professional PDF Documents Khapra, and Pratyush Kumar

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-02T06:59:49.400835Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:59:49.400835Z digest=sha256:3dff16552b5a7ba512005ea3f0a0b2947fce74d10cc9fe5eaf4e3ed9795f1b43

Observation a7d6c6ae-7b8a-4a9c-a0b9-07b8405d0de1 · outbound

This paper cites OmniDocBench: Benchmarking di- verse PDF document parsing with comprehensive annotations.

GDP.pdf: Benchmarking Grounded Multimodal Reasoning over Professional PDF Documents OmniDocBench: Benchmarking di- verse PDF document parsing with comprehensive annotations

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-02T06:59:49.493409Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:59:49.493409Z digest=sha256:ec73f9b73fb8fa8f8db30d323ad892e83f173c451156fba48e2ad5e0759ff721

Observation f42ddebf-350c-41b7-bf56-8a982b89b375 · outbound

This paper cites CORD: A consolidated receipt dataset for post-OCR parsing.

GDP.pdf: Benchmarking Grounded Multimodal Reasoning over Professional PDF Documents CORD: A consolidated receipt dataset for post-OCR parsing

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-02T06:59:49.593004Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:59:49.593004Z digest=sha256:1ce1936c263138dfaece29a3e655e81c513fe9e7c486d04886b23c7e32d4e7c2

Observation 24b8ed1b-db6d-42ad-a10d-5b3f8fb1cedd · outbound

This paper cites ANLS* -- A Universal Document Processing Metric for Generative Large Language Models.

GDP.pdf: Benchmarking Grounded Multimodal Reasoning over Professional PDF Documents ANLS* -- A Universal Document Processing Metric for Generative Large Language Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-02T06:59:49.672205Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:59:49.672205Z digest=sha256:92f8a82d9d14ba53df5a6e28e52091bf33ce062ccfecfd73e3ebabbdede426ca

Observation 80ecc4cf-06e7-4946-8282-1f8d7d04cfe9 · outbound

This paper cites Nassar, and Peter W.

GDP.pdf: Benchmarking Grounded Multimodal Reasoning over Professional PDF Documents Nassar, and Peter W

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-02T06:59:49.761158Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:59:49.761158Z digest=sha256:58cb56d84f9eca3e8ab11c33697f19aa83f13ea7d2218124417f394887971778

Observation 8c11d521-ab77-4be5-ad5f-bad25e686ef9 · outbound

This paper cites Rossi, and Franck Dernoncourt.

GDP.pdf: Benchmarking Grounded Multimodal Reasoning over Professional PDF Documents Rossi, and Franck Dernoncourt

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-02T06:59:49.823769Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:59:49.823769Z digest=sha256:2406dba538c0e4166cab79e6fdaf1757c2f8abd8f3166e73461a43646af844dc

Observation ede9e7f1-8e60-4b31-9358-c1f87da616a1 · outbound

This paper cites A-OKVQA: A benchmark for visual question answering using world knowl- edge.

GDP.pdf: Benchmarking Grounded Multimodal Reasoning over Professional PDF Documents A-OKVQA: A benchmark for visual question answering using world knowl- edge

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-02T06:59:49.901797Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:59:49.901797Z digest=sha256:29109b68cb7cfb1483729e03aa7dec789d3867a200301e75b1c019c970175c79

Observation 79f09951-8e71-4775-b55f-9c5e2e53a30f · outbound

This paper cites DocILE benchmark for document information localization and extraction.

GDP.pdf: Benchmarking Grounded Multimodal Reasoning over Professional PDF Documents DocILE benchmark for document information localization and extraction

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-02T06:59:50.020674Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:59:50.020674Z digest=sha256:6034f578c1547f181ed2c5ac0c6dc8cfb020dfcd38b957c6daacfedc5b89beee

Observation 92e3c209-609e-49f7-a35d-f78ef6f23148 · outbound

This paper cites Hi- erarchical multimodal transformers for Multipage DocVQA.

GDP.pdf: Benchmarking Grounded Multimodal Reasoning over Professional PDF Documents Hi- erarchical multimodal transformers for Multipage DocVQA

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-02T06:59:50.104411Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:59:50.104411Z digest=sha256:bbf8db36f1cdfe152e04f19b63de3dd8eaeb033dc0a89e3b08267613161a491d

Observation c82a8dbe-274e-4496-897b-efe72a965661 · outbound

This paper cites Document understanding dataset and evaluation (DUDE).

GDP.pdf: Benchmarking Grounded Multimodal Reasoning over Professional PDF Documents Document understanding dataset and evaluation (DUDE)

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-02T06:59:50.194909Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:59:50.194909Z digest=sha256:874f87f7bd976204e3e6f6066107116104994ce6b4afd2f1c0bb642e7c6ff2fe

Observation 8bebf2e7-3eaf-415c-ab46-f8c169c7e410 · outbound

This paper cites CharXiv: Charting gaps in realistic chart understand- ing in multimodal LLMs.

GDP.pdf: Benchmarking Grounded Multimodal Reasoning over Professional PDF Documents CharXiv: Charting gaps in realistic chart understand- ing in multimodal LLMs

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-02T06:59:50.290926Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:59:50.290926Z digest=sha256:3bdfb2b0e01843ac7feece46e928c769ac30ab1111f2ca9b63c89e2f0d4fa0d7

Observation d60f0562-17b8-4307-bc4e-7ce972cc611f · outbound

This paper cites MMMU: A massive multi-discipline multimodal understanding and reasoning benchmark for expert AGI.

GDP.pdf: Benchmarking Grounded Multimodal Reasoning over Professional PDF Documents MMMU: A massive multi-discipline multimodal understanding and reasoning benchmark for expert AGI

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-02T06:59:50.379963Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:59:50.379963Z digest=sha256:91ad6ba936182b9f58540df783a5c3b773944d68f04bba8fb4f34d98f9715c3a

Observation 040927e4-e03f-4c22-8e85-47ea0024c31b · outbound

This paper cites MMMU-Pro: A more robust multi-discipline multimodal understanding benchmark.

GDP.pdf: Benchmarking Grounded Multimodal Reasoning over Professional PDF Documents MMMU-Pro: A more robust multi-discipline multimodal understanding benchmark

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-02T06:59:50.498953Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:59:50.498953Z digest=sha256:079788ca6ca5753e5bf05af07e8418dd455c974d35419ffeb13895032be46fc6

Observation 22fbf20f-1d95-48c8-accb-db24db9354b6 · outbound

This paper cites Pub- LayNet: Largest dataset ever for document layout analysis.

GDP.pdf: Benchmarking Grounded Multimodal Reasoning over Professional PDF Documents Pub- LayNet: Largest dataset ever for document layout analysis

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-02T06:59:50.558777Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:59:50.558777Z digest=sha256:543593cfa6526b55aba92d042c3c9acd9350703584920d1effcb3df2fcf957d2

Observation 5918a48f-22b1-4db9-a739-69fd7559c5d8 · outbound

This paper cites Real5-OmniDocBench: A Full-Scale Physical Reconstruction Benchmark for Robust Document Parsing in the Wild.

GDP.pdf: Benchmarking Grounded Multimodal Reasoning over Professional PDF Documents Real5-OmniDocBench: A Full-Scale Physical Reconstruction Benchmark for Robust Document Parsing in the Wild

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-02T06:59:50.615006Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:59:50.615006Z digest=sha256:0b54736eff93d93c93098650ff477bf66bae57f31e36545ea2247850b2e298c3

Observation d0f1e0a6-8b21-41fc-8b51-2c919fd545f9 · outbound

This paper cites TAT-QA: A question answering benchmark on a hybrid of tabular and textual content in finance.

GDP.pdf: Benchmarking Grounded Multimodal Reasoning over Professional PDF Documents TAT-QA: A question answering benchmark on a hybrid of tabular and textual content in finance

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-02T06:59:50.686542Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:59:50.686542Z digest=sha256:a4eb729a60e22acc05d47651cd34e03d4c90ad927dd2c893a8978b54e4f8b20a

Observation 64126c30-e86a-46ef-8b1d-4e2f06a934b2 · outbound

This paper cites Towards complex document understanding by discrete reasoning.

GDP.pdf: Benchmarking Grounded Multimodal Reasoning over Professional PDF Documents Towards complex document understanding by discrete reasoning

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-02T06:59:50.761784Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:59:50.761784Z digest=sha256:7ef5756aad278e19816cfcd29b2cc17448301ce633e340901d403075c60b0189

Observation b6e7b5f7-4d36-451c-a7fc-3ee6d0384bc1 · outbound

This paper cites MMDocBench: Benchmarking large vision-language models for fine-grained visual document understanding and grounding.

GDP.pdf: Benchmarking Grounded Multimodal Reasoning over Professional PDF Documents MMDocBench: Benchmarking large vision-language models for fine-grained visual document understanding and grounding

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-02T06:59:50.862993Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:59:50.862993Z digest=sha256:737756eb56fca5d2fe1c3e06ec688356af11f0418f3b7c36805ae83f7bcb966b

Observation bfd442fb-a8db-4fb9-8d76-8b659cd90987 · outbound

This paper cites DOCBENCH: A Benchmark for Evaluating LLM-based Document Reading Systems.

GDP.pdf: Benchmarking Grounded Multimodal Reasoning over Professional PDF Documents DOCBENCH: A Benchmark for Evaluating LLM-based Document Reading Systems

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-02T06:59:50.917441Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:59:50.917441Z digest=sha256:8a47f5269db13175d8bb014c195ace383bbd5e2462aebf92d0b1e0709e84a1fd

Pith citing papers

No inbound Pith citation observations are available.