Pith. sign in

Paper Citation Record · LEDGER

A Token-level Text Image Foundation Model for Document Understanding

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:2503.02304.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2503.02304 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 5 of 5 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-02T14:55:43.470957Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-20T14:43:22.323230Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 97f94665-d118-465f-8682-43be64fe1d70 · inbound

CodePercept: Code-Grounded Visual STEM Perception for MLLMs cites this paper.

CodePercept: Code-Grounded Visual STEM Perception for MLLMs A Token-level Text Image Foundation Model for Document Understanding

Reference 17

Resolution
unresolved
no resolver link, observed 2026-07-14T23:22:13.847876Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T23:22:13.847876Z digest=sha256:3ed7de70655981df1d6f1c28098c0f3a07b37a2752894ed249a27a2cf25e562f

Observation 3a06ee62-7ad7-44a4-b8b4-d0b1bf4a9e40 · inbound

StyleTextGen: Style-Conditioned Multilingual Scene Text Generation cites this paper.

StyleTextGen: Style-Conditioned Multilingual Scene Text Generation A Token-level Text Image Foundation Model for Document Understanding

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-15T04:55:03.129477Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-15T04:54:52.729257Z digest=sha256:06502bfd08b9c3d7f55b24b0883a1260a7c5b8deac54d3118028b93f6b47f61e

Observation aa67edf7-4fb4-44c2-b300-89d8c139ed19 · inbound

Beyond Detection: A Structure-Aware Framework for Scene Text Tracking cites this paper.

Beyond Detection: A Structure-Aware Framework for Scene Text Tracking A Token-level Text Image Foundation Model for Document Understanding

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-20T14:43:22.333828Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-20T14:42:38.961728Z digest=sha256:30de19db675fbe0d7503c64b0a792ee427b90e8c0c25ec5883ca59a9ad6f7500

Observation ed34658a-2942-4f20-8c00-76aefe254146 · inbound

DocOCR-Eval: A Correction-Based Framework for OCR Tool Selection Without Ground Truth cites this paper.

DocOCR-Eval: A Correction-Based Framework for OCR Tool Selection Without Ground Truth A Token-level Text Image Foundation Model for Document Understanding

Reference 227

Resolution
unresolved
no resolver link, observed 2026-08-02T14:55:43.470957Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T14:55:43.470957Z digest=sha256:d5ae2e040380b15f93e9998d141f220a91ed28c46c0e047494cba310a6ea6e51

Observation 21de2478-3a78-4df6-af50-ec66cee72b28 · inbound

SCOPE and SCION: A Benchmark and an Auditable Reference Pipeline for Schema Induction and Fusion from Text cites this paper.

SCOPE and SCION: A Benchmark and an Auditable Reference Pipeline for Schema Induction and Fusion from Text A Token-level Text Image Foundation Model for Document Understanding

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-02T13:36:56.239734Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T13:36:56.239734Z digest=sha256:b3d99c73134f82c5d3978b98af61d2915ee904644cdcd2ddf9d3d94e7c570acb