Pith. sign in

Paper Citation Record · LEDGER

Spatial Dual-Modality Graph Reasoning for Key Information Extraction

As of 23 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 11 inbound Pith citation observations for arXiv:2103.14470.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2103.14470 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 11 of 11 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 11 of 11 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T00:09:18.780309Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-08T02:04:26.390365Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation a81de05e-3496-4a19-b7b3-a975e5135dfd · inbound

BlueLM-V-3B: Algorithm and System Co-Design for Multimodal Large Language Models on Mobile Devices cites this paper.

BlueLM-V-3B: Algorithm and System Co-Design for Multimodal Large Language Models on Mobile Devices Spatial Dual-Modality Graph Reasoning for Key Information Extraction

Reference 111

Resolution
unresolved
no resolver link, observed 2026-08-12T19:33:00.952235Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:33:00.952235Z digest=sha256:bdffe578c0eb4ffb76ac5bce5993cc026eac1fee598375973f5511c5556e3adc

Observation f64efca0-3d6f-4ff7-aca3-99c6fa340b25 · inbound

Adaptive Markup Language Generation for Contextually-Grounded Visual Document Understanding cites this paper.

Adaptive Markup Language Generation for Contextually-Grounded Visual Document Understanding Spatial Dual-Modality Graph Reasoning for Key Information Extraction

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-15T23:08:51.809034Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:08:51.809034Z digest=sha256:6069294873a15072ebe3cc682ce8b1d6498d03e17af4070fea515e5a8ddc6314

Observation 8073ba43-4ed4-47ee-81e1-5da74344ec25 · inbound

OCR-Reasoning Benchmark: Unveiling the True Capabilities of MLLMs in Complex Text-Rich Image Reasoning cites this paper.

OCR-Reasoning Benchmark: Unveiling the True Capabilities of MLLMs in Complex Text-Rich Image Reasoning Spatial Dual-Modality Graph Reasoning for Key Information Extraction

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:20.180732Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:57:20.180732Z digest=sha256:108db7671e3d4311aa5d3d01f982eb349b3041710cc14e52190384577425d45a

Observation 7927cf82-2d59-4f37-a184-cef2e104c169 · inbound

VDInstruct: Zero-Shot Key Information Extraction via Content-Aware Vision Tokenization cites this paper.

VDInstruct: Zero-Shot Key Information Extraction via Content-Aware Vision Tokenization Spatial Dual-Modality Graph Reasoning for Key Information Extraction

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-06T17:59:06.590430Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:59:06.590430Z digest=sha256:981cfa20b439232a46ea04db10a6f16caeb37dc1a3e665e4a2cabeb94cb1f74d

Observation 88fd108d-58f0-4c8a-8dc5-dbec2ca4b345 · inbound

Noncrossing Duality and the Geometry of Positive Tropical Linear Spaces cites this paper.

Noncrossing Duality and the Geometry of Positive Tropical Linear Spaces Spatial Dual-Modality Graph Reasoning for Key Information Extraction

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-07-01T09:25:39.933815Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-07-01T09:23:20.044291Z digest=sha256:9e1e60a8d97a4b9435a95cfa9ebc755e6323252638699c4c9cd013e04c4b359b

Observation 1d2d7d9c-673c-437d-94e3-9d0c2ebdcbf7 · inbound

When the Forger Is the Judge: GPT-Image-2 Cannot Recognize Its Own Faked Documents cites this paper.

When the Forger Is the Judge: GPT-Image-2 Cannot Recognize Its Own Faked Documents Spatial Dual-Modality Graph Reasoning for Key Information Extraction

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-11T23:31:15.041032Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-07T16:54:17.319933Z digest=sha256:ba4590ba20167cd0b809afd197fe2bdc433be107304bb86a675b5b00fcebbbaa

Observation f5ac9e35-cfa7-4522-98f5-dc6b98b0a828 · inbound

Enhancing Large Multimodal Models in Key Information Extraction via Scene-Aware Document Synthesis cites this paper.

Enhancing Large Multimodal Models in Key Information Extraction via Scene-Aware Document Synthesis Spatial Dual-Modality Graph Reasoning for Key Information Extraction

Reference 6

Resolution
unresolved
no resolver link, observed 2026-07-11T16:02:00.920066Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T16:02:00.920066Z digest=sha256:5b64567505a038c0146f558ea5bc287a39cba9fea41130b66d2f0a63d87b0b85

Observation 03cb6a75-c915-4770-96a9-9d6c9f69018e · inbound

Vision as Unified Multimodal Generation cites this paper.

Vision as Unified Multimodal Generation Spatial Dual-Modality Graph Reasoning for Key Information Extraction

Reference 160

Resolution
verified exact
local_arxiv, observed 2026-07-08T02:04:26.392042Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-07-08T01:54:30.649092Z digest=sha256:f5307792db68763f9a4c53658b0906f8004a1f69b8c47b73613fcf589ed227bb

Observation 327a0edc-4c53-434a-9e33-80f92e42df3c · inbound

StepX-Edge: An On-Device UI Vision-Language Model via Architecture-Training-Deployment Co-Design cites this paper.

StepX-Edge: An On-Device UI Vision-Language Model via Architecture-Training-Deployment Co-Design Spatial Dual-Modality Graph Reasoning for Key Information Extraction

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-01T16:31:11.264506Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T16:31:11.264506Z digest=sha256:672948ee227b5d7aab001cdd199fd402ba6ccc082d16f379af98885cc4465475

Observation 72097a27-1519-433d-ade6-d4797b0fdc45 · inbound

AI-generated Images Challenge Visual Trust in High-risk Scenarios cites this paper.

AI-generated Images Challenge Visual Trust in High-risk Scenarios Spatial Dual-Modality Graph Reasoning for Key Information Extraction

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-01T08:52:44.665708Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T08:52:44.665708Z digest=sha256:ce5c84b395cddf53f1515c4b6dbcdffddd53425df057e27deda24b44f89d8c5a

Observation c73a9685-0999-496c-a167-62e30349fabb · inbound

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding cites this paper.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Spatial Dual-Modality Graph Reasoning for Key Information Extraction

Reference 182

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.780309Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.780309Z digest=sha256:d30282031b2824022c315e7587b1fac0affb0241fe4177244cf2dae13e60b678