Pith. sign in

Paper Citation Record · LEDGER

DocVQA: A Dataset for VQA on Document Images

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 38 inbound Pith citation observations for arXiv:2007.00398.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2007.00398 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 38 of 38 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 38 of 38 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:53:24.790502Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T21:00:09.476068Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 9fc75294-2741-4978-a25d-8c0ab778eea9 · inbound

Long Context Transfer from Language to Vision cites this paper.

Long Context Transfer from Language to Vision DocVQA: A Dataset for VQA on Document Images

Reference 54

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T07:08:36.126572Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-12T07:08:35.946669Z digest=sha256:727653c2b2360201da92eda8b05f652224d2da14ceed1318ca2e36d4d1b403f2

Observation 50d9352d-b38e-467e-b124-af0163694f8b · inbound

PaliGemma: A versatile 3B VLM for transfer cites this paper.

PaliGemma: A versatile 3B VLM for transfer DocVQA: A Dataset for VQA on Document Images

Reference 94

Resolution
verified exact
arxiv_id, observed 2026-05-11T13:10:20.770367Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-11T13:10:19.972353Z digest=sha256:cdd21d743aaf60daecc2f23f6d7becaab72e39e7d1e27f381f601fe9434483ed

Observation db46a353-cc1a-4cdc-8314-d9c82893565f · inbound

PaliGemma 2: A Family of Versatile VLMs for Transfer cites this paper.

PaliGemma 2: A Family of Versatile VLMs for Transfer DocVQA: A Dataset for VQA on Document Images

Reference 64

Resolution
verified exact
arxiv_id, observed 2026-05-15T09:15:07.712195Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-15T09:15:07.523565Z digest=sha256:5e607928add17b48f9c33f676fff4b0430a5a580819937610b465237d68a0b5b

Observation 369fbbf5-48c1-4fab-94fc-9a7d3f52ddf4 · inbound

FLARE: Fully Integration of Vision-Language Representations for Deep Cross-Modal Understanding cites this paper.

FLARE: Fully Integration of Vision-Language Representations for Deep Cross-Modal Understanding DocVQA: A Dataset for VQA on Document Images

Reference 53

Resolution
verified exact
arxiv_id, observed 2026-05-22T19:52:01.796335Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T19:49:00.961388Z digest=sha256:0ab8f74d3d78ca634d7af23e7957c8e7aa85f82de32fcca9eb784354d415894e

Observation 99fdb9fa-25a6-4dd7-960c-7f37b6558ffc · inbound

Understand, Think, and Answer: Advancing Visual Reasoning with Large Multimodal Models cites this paper.

Understand, Think, and Answer: Advancing Visual Reasoning with Large Multimodal Models DocVQA: A Dataset for VQA on Document Images

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-07T13:53:24.790502Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:53:24.790502Z digest=sha256:11d91f15cac71b867247b7bc4c63718301aae4d97462784fe0af87c0d0b2131e

Observation eab244cd-062d-451b-91ab-aaaca92a01da · inbound

Spoken question answering for visual queries cites this paper.

Spoken question answering for visual queries DocVQA: A Dataset for VQA on Document Images

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:39.942212Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:52:39.942212Z digest=sha256:64d26e8b28bc574d7d7b1ec131e0a8db50334d7e1daec2439bccb2b22ee604ee

Observation 7e1f81cd-a1b2-4d68-b093-2d83107fdba0 · inbound

EffiVLM-BENCH: A Comprehensive Benchmark for Evaluating Training-Free Acceleration in Large Vision-Language Models cites this paper.

EffiVLM-BENCH: A Comprehensive Benchmark for Evaluating Training-Free Acceleration in Large Vision-Language Models DocVQA: A Dataset for VQA on Document Images

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T12:09:06.849342Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:09:06.849342Z digest=sha256:87658c1bdde4c6c77931b30bea0ae06c6f9fabf68635de5a22138f7b1f9181de

Observation 202d097c-9a67-4c76-85eb-19e2422e6c74 · inbound

On the Comprehensibility of Multi-structured Financial Documents using LLMs and Pre-processing Tools cites this paper.

On the Comprehensibility of Multi-structured Financial Documents using LLMs and Pre-processing Tools DocVQA: A Dataset for VQA on Document Images

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T10:27:36.789692Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:27:36.789692Z digest=sha256:ce785febea7ee8b11be496b149a136ab475d0df74a51fbcf70025db9533e3c5c

Observation 40e38e9f-cae0-47d4-9c06-2aae302fb3f7 · inbound

VDInstruct: Zero-Shot Key Information Extraction via Content-Aware Vision Tokenization cites this paper.

VDInstruct: Zero-Shot Key Information Extraction via Content-Aware Vision Tokenization DocVQA: A Dataset for VQA on Document Images

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T17:59:05.622333Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:59:05.622333Z digest=sha256:5f2de58572f98203205257f2544d4bb7d419c7e3780fa14ef2fb2f4a66f528de

Observation d6733d6f-d499-4144-9636-5486db821583 · inbound

Differential Multimodal Transformers cites this paper.

Differential Multimodal Transformers DocVQA: A Dataset for VQA on Document Images

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T16:39:20.935176Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:39:20.935176Z digest=sha256:68fe780e68c252ef47c15bb8fb13482e444e596e9856695e583bc3b2442ece1e

Observation b8386a33-20c2-4723-803b-f2ef52d0b13a · inbound

Analyze-Prompt-Reason: A Collaborative Agent-Based Framework for Multi-Image Vision-Language Reasoning cites this paper.

Analyze-Prompt-Reason: A Collaborative Agent-Based Framework for Multi-Image Vision-Language Reasoning DocVQA: A Dataset for VQA on Document Images

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T10:14:26.472521Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:14:26.472521Z digest=sha256:7e8a6ca9f4373f0fb918b7d6aa0d3a9df06ce54b6a01990c9724c92eedccfe09

Observation 2b5709b0-a04a-4bd6-87d9-71b46c381cc6 · inbound

VLMs-in-the-Wild: Bridging the Gap Between Academic Benchmarks and Enterprise Reality cites this paper.

VLMs-in-the-Wild: Bridging the Gap Between Academic Benchmarks and Enterprise Reality DocVQA: A Dataset for VQA on Document Images

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-05T11:11:20.832097Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:11:20.832097Z digest=sha256:b9b1e5c7df384c492f85f528ae2afb050b5eab857497ada66e6368554421cd52

Observation 054fa4f0-9302-40b0-aa7c-d26a758aa5bf · inbound

Visual-TableQA: Open-Domain Benchmark for Reasoning over Table Images cites this paper.

Visual-TableQA: Open-Domain Benchmark for Reasoning over Table Images DocVQA: A Dataset for VQA on Document Images

Reference 26

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T17:42:47.586020Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-18T17:37:31.837022Z digest=sha256:f8b81cde09cf150ddfa138c92945eff9dc98f2877c7a1a49c1bef201d0740c94

Observation ab61b841-c5db-45f1-afeb-b8d1dd3d97dd · inbound

Routing-Based Continual Learning for Multimodal Large Language Models cites this paper.

Routing-Based Continual Learning for Multimodal Large Language Models DocVQA: A Dataset for VQA on Document Images

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-05-18T00:55:35.336139Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-18T00:52:36.700027Z digest=sha256:c22191e4a6817f0c837df5960be61156119fee1c4ed1a96bfb660bef7383c22f

Observation 2fdfb66c-d33a-4093-b387-627443e15855 · inbound

FinCriticalED: A Visual Benchmark for Financial Fact-Level OCR cites this paper.

FinCriticalED: A Visual Benchmark for Financial Fact-Level OCR DocVQA: A Dataset for VQA on Document Images

Reference 20

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T20:20:11.510920Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-17T20:19:52.701262Z digest=sha256:eb70a753093af927ea76cbd3495054b37b2909dcb914ced1cd8b644e2545bed1

Observation 2d897ada-8f28-4f91-9800-5f8d6012ed55 · inbound

Reconstructing Content with Collaborative Attention for Universal Multimodal Representation Learning cites this paper.

Reconstructing Content with Collaborative Attention for Universal Multimodal Representation Learning DocVQA: A Dataset for VQA on Document Images

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-02T19:41:33.738459Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T19:41:33.738459Z digest=sha256:a0ccdeda9b360106287912cfc346c714a436f5e00fbcded7030218eb234b62d6

Observation e8010136-aa82-4ae3-8ded-81593f01f200 · inbound

FileGram: Grounding Agent Personalization in File-System Behavioral Traces cites this paper.

FileGram: Grounding Agent Personalization in File-System Behavioral Traces DocVQA: A Dataset for VQA on Document Images

Reference 17

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T22:45:51.276505Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T19:33:57.643549Z digest=sha256:efdb4dee63ac6fc7844fb6b66f1480f63822e2e90373c127548848623c182584

Observation 756994a1-db77-498a-90b5-6873aee623f7 · inbound

Q-Zoom: Query-Aware Adaptive Perception for Efficient Multimodal Large Language Models cites this paper.

Q-Zoom: Query-Aware Adaptive Perception for Efficient Multimodal Large Language Models DocVQA: A Dataset for VQA on Document Images

Reference 29

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T23:55:52.856348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T18:46:26.869644Z digest=sha256:54d182cf195b6b185382a27887f8dc945c9571bfc3de19f7769060dc1027d136

Observation 99a7eafb-0ef2-47b6-bbc3-23c10c5fed5c · inbound

Vision-Language Foundation Models for Comprehensive Automated Pavement Condition Assessment cites this paper.

Vision-Language Foundation Models for Comprehensive Automated Pavement Condition Assessment DocVQA: A Dataset for VQA on Document Images

Reference 50

Resolution
verified exact
arxiv_id, observed 2026-05-11T06:41:22.377691Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T17:30:39.410040Z digest=sha256:672401e4f3bc9dcc886e0f8f189c48f6375ee1b017210dfe61b851f7a04bd3a0

Observation f54d59ad-4dd6-4343-9bec-42260ac96d19 · inbound

Entropy-Gradient Grounding: Training-Free Evidence Retrieval in Vision-Language Models cites this paper.

Entropy-Gradient Grounding: Training-Free Evidence Retrieval in Vision-Language Models DocVQA: A Dataset for VQA on Document Images

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-11T07:16:11.146347Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T17:14:11.941977Z digest=sha256:10e851bdd1d27dc84fec2227fca2c0e96b7e23b2699406993fdb817196a676cf

Observation ed301277-68ed-493f-b006-79df3acf6942 · inbound

Knowing When Not to Answer: Evaluating Abstention in Multimodal Reasoning Systems cites this paper.

Knowing When Not to Answer: Evaluating Abstention in Multimodal Reasoning Systems DocVQA: A Dataset for VQA on Document Images

Reference 1

Resolution
malformed identifier
arxiv_id, observed 2026-05-10T11:50:20.766938Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T11:42:13.100463Z digest=sha256:6714fe5aca53332f020a700c1af500e8b6570888126175867d2180d9e6a7536f

Observation 1bb123da-ab2f-465d-98e3-8289bbacef79 · inbound

ReaLB: Real-Time Load Balancing for Multimodal MoE Inference cites this paper.

ReaLB: Real-Time Load Balancing for Multimodal MoE Inference DocVQA: A Dataset for VQA on Document Images

Reference 28

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T13:36:10.208047Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-10T01:18:27.597377Z digest=sha256:fb04ae7401fea012014cf8726250ad9f53513e6137af90164a4c0d9bce4d35f1

Observation fd9d38bd-1825-406a-be64-8868f184dab1 · inbound

Visual Reasoning through Tool-supervised Reinforcement Learning cites this paper.

Visual Reasoning through Tool-supervised Reinforcement Learning DocVQA: A Dataset for VQA on Document Images

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-11T12:46:03.231775Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T03:05:21.688216Z digest=sha256:ef0a9e86cab42848d1849b86eabdd84797185e3b410155cd90cf17e9e0606d7a

Observation 38b54dfb-70d2-4d82-9413-4a23a8d5eea4 · inbound

The category of Whittaker modules over the Cartan Type Lie algebra $\bar{S}_2$ cites this paper.

The category of Whittaker modules over the Cartan Type Lie algebra $\bar{S}_2$ DocVQA: A Dataset for VQA on Document Images

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-07-01T09:25:40.475504Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-07-01T09:08:06.592577Z digest=sha256:eb811e95ee4f74d9a9434a6cb4374dcf41a87f424ad1fe3b12575fd5599d589b

Observation 3b1e86e4-3282-4455-a195-5d9290b8e94a · inbound

FCMBench-Video: Benchmarking Document Video Intelligence cites this paper.

FCMBench-Video: Benchmarking Document Video Intelligence DocVQA: A Dataset for VQA on Document Images

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-11T23:21:41.346550Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-07T17:14:19.186123Z digest=sha256:aa46606a1efd7b381752d5e49738212f29598dcaab6c4e39516e65c3cefdf2fc

Observation fec6371b-9edf-4e6c-b000-4668c7c106e8 · inbound

RetentiveKV: State-Space Memory for Uncertainty-Aware Multimodal KV Cache Eviction cites this paper.

RetentiveKV: State-Space Memory for Uncertainty-Aware Multimodal KV Cache Eviction DocVQA: A Dataset for VQA on Document Images

Reference 45

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:26:02.150150Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-10T15:29:17.567557Z digest=sha256:aaa66547d6372d383ee1b7a41a19d883b2fb40024cfbf9f6c3515be6e289a704

Observation 430876f6-6afd-4c69-92ed-2e3c253887b0 · inbound

Focus-then-Context: Subject-Centric Progressive Visual Token Reduction for Vision-Language Models cites this paper.

Focus-then-Context: Subject-Centric Progressive Visual Token Reduction for Vision-Language Models DocVQA: A Dataset for VQA on Document Images

Reference 55

Resolution
verified exact
arxiv_id, observed 2026-05-21T05:23:58.581700Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-21T05:20:55.448430Z digest=sha256:f467575014e79287866b4cabdeea6d6afc7cf2f38508f99f59cf424d69051afb

Observation fdc13621-692c-4c1a-9d69-acb3c29209ca · inbound

Reinforcement Learning with Robust Rubric Rewards cites this paper.

Reinforcement Learning with Robust Rubric Rewards DocVQA: A Dataset for VQA on Document Images

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-06-29T07:43:13.742120Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-29T07:39:21.677389Z digest=sha256:17c2962462beb888790c85d6355c8824d1470f644a7a571ea960b87dd1cd3bed

Observation cbe0c011-0592-4df5-9f91-8eeaf596ab3f · inbound

Representation Forcing for Bottleneck-Free Unified Multimodal Models cites this paper.

Representation Forcing for Bottleneck-Free Unified Multimodal Models DocVQA: A Dataset for VQA on Document Images

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:16:00.977140Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T22:54:10.460872Z digest=sha256:ab086a93cfb8814475032ec0503fd53449954e565c9b99d46b562c7b14b8aba2

Observation 4c9a5c00-a365-4c3e-89fc-85f4d9e504a0 · inbound

Representation Forcing for Bottleneck-Free Unified Multimodal Models cites this paper.

Representation Forcing for Bottleneck-Free Unified Multimodal Models DocVQA: A Dataset for VQA on Document Images

Reference 35

Resolution
unresolved
no resolver link, observed 2026-07-12T15:31:57.426559Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T15:31:57.426559Z digest=sha256:23caae1703f6935f0c1bcdbea855b825230336774539acb4bfbef398a863528f

Observation 010cabf7-d743-4021-b8fc-fcd648a84a74 · inbound

Loss Landscape Poisoning: Targeted Extraction of Unseen Training Data from LLMs cites this paper.

Loss Landscape Poisoning: Targeted Extraction of Unseen Training Data from LLMs DocVQA: A Dataset for VQA on Document Images

Reference 48

Resolution
verified exact
arxiv_id, observed 2026-07-03T17:38:43.814213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T03:59:30.468854Z digest=sha256:96aaf6d3c9d9786ae70a3889956afa06c5ab5b1cb494756d2fd6f80a73fa3f45

Observation 1bf935f7-a473-45a5-873b-0012da53ba03 · inbound

How Robust is OCR-Reasoning? Evaluating OCR-Reasoning Robustness of Vision-Language Models under Visual Perturbations cites this paper.

How Robust is OCR-Reasoning? Evaluating OCR-Reasoning Robustness of Vision-Language Models under Visual Perturbations DocVQA: A Dataset for VQA on Document Images

Reference 38

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T21:00:09.477417Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-25T19:13:27.527971Z digest=sha256:49f31951d0fbb06c02df3562ac545cac178a4531067f64598b71b1ba88ae1c20

Observation d3a070c7-8575-4c1c-96f1-715e82e35265 · inbound

Combating Textual Noise and Redundancy: Entropy-Aware Dense Visual Token Pruning cites this paper.

Combating Textual Noise and Redundancy: Entropy-Aware Dense Visual Token Pruning DocVQA: A Dataset for VQA on Document Images

Reference 47

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T14:48:32.412815Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-07-03T14:47:35.377391Z digest=sha256:7390d61524c656f2b001ddb828ebd9db5094b4e933950ebe4874750da8783836

Observation d0a1ea24-90f5-47ea-a7ef-7e51dccc1a17 · inbound

RADIO1D: Elastic Representations for Condensed Vision Modeling cites this paper.

RADIO1D: Elastic Representations for Condensed Vision Modeling DocVQA: A Dataset for VQA on Document Images

Reference 40

Resolution
unresolved
no resolver link, observed 2026-07-12T01:07:20.766474Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:07:20.766474Z digest=sha256:e057774a4fe6d059ee0468ea21a6288aa20f9b6fd23ad91d68184ef72b774765

Observation 7e001a6e-0255-4ee9-93d8-2d2459c10395 · inbound

Data Pyramid for Embodied Manipulation cites this paper.

Data Pyramid for Embodied Manipulation DocVQA: A Dataset for VQA on Document Images

Reference 253

Resolution
unresolved
no resolver link, observed 2026-07-31T06:18:55.901623Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:18:55.901623Z digest=sha256:5c26ca5d6588779dee8883105cf11094b37d7dedbb30a0bd65a2ef023d99127d

Observation 2774c8d8-9c4f-4ac0-b7c5-6e6d4d973c7e · inbound

DocAnnot -- Accelerating the Creation of Key Information Extraction Datasets with GenAI-Powered Auto-annotation cites this paper.

DocAnnot -- Accelerating the Creation of Key Information Extraction Datasets with GenAI-Powered Auto-annotation DocVQA: A Dataset for VQA on Document Images

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-02T14:38:35.492534Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T14:38:35.492534Z digest=sha256:741b4e9e27ffd90c6e380d3d7e33f4fdd4ac5c5bdc7e5c46181d93e37bc74266

Observation 86b507e3-44ac-49ba-a61e-a3f591c852f8 · inbound

CLBench-V: Evaluating Multimodal Context Learning from Grounding to Knowledge Acquisition cites this paper.

CLBench-V: Evaluating Multimodal Context Learning from Grounding to Knowledge Acquisition DocVQA: A Dataset for VQA on Document Images

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-01T02:55:23.403485Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:55:23.403485Z digest=sha256:f4685d9742e29031bdf27372f723a89df09fb07b7be600b6d80a223581c96a41

Observation d129ac94-b538-4ea4-b655-27c8cecd55c1 · inbound

Rethinking Self-Evolving Agent Skills: Feedback Dynamics over Multiple Rounds cites this paper.

Rethinking Self-Evolving Agent Skills: Feedback Dynamics over Multiple Rounds DocVQA: A Dataset for VQA on Document Images

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-05T04:26:29.325114Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T04:26:29.325114Z digest=sha256:73f5e861eff5a840a4183142f55e57be4bee37cde36695549ae95483954ee5fd