Pith. sign in

Paper Citation Record · LEDGER

LLaVA-Read: Enhancing Reading Ability of Multimodal Language Models

As of 23 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2407.19185.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2407.19185 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T23:08:52.203082Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 03acbac3-de00-4c32-926f-a2c60e18e971 · inbound

EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering cites this paper.

EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering LLaVA-Read: Enhancing Reading Ability of Multimodal Language Models

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-08T12:54:56.375610Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T12:54:56.375610Z digest=sha256:7cdd3df95a46f6b7078cd5d5885d671bbdbe04dd3d17e3c1c3169354b0c27c2b

Observation 97b5a255-71c8-4c62-9a18-46fd0f6bdf04 · inbound

Adaptive Markup Language Generation for Contextually-Grounded Visual Document Understanding cites this paper.

Adaptive Markup Language Generation for Contextually-Grounded Visual Document Understanding LLaVA-Read: Enhancing Reading Ability of Multimodal Language Models

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-15T23:08:52.203082Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:08:52.203082Z digest=sha256:0d4162a6a2baa7ac76f827b8f6ec7e55c853e2e271c70a995f29857005998dd2

Observation 97993c1c-9a6d-4154-9aff-ae7d5edad3c3 · inbound

MusiXQA: Advancing Visual Music Understanding in Multimodal Large Language Models cites this paper.

MusiXQA: Advancing Visual Music Understanding in Multimodal Large Language Models LLaVA-Read: Enhancing Reading Ability of Multimodal Language Models

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T21:57:07.053168Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:57:07.053168Z digest=sha256:b565fd2d70e58b1f7d97e9a46c236b70c3b7081cc0c61a8d50acce523e6caa95

Observation 2cb5e786-46b7-4cdb-aa97-67a2c5b36239 · inbound

A Survey on MLLM-based Visually Rich Document Understanding: Methods, Challenges, and Emerging Trends cites this paper.

A Survey on MLLM-based Visually Rich Document Understanding: Methods, Challenges, and Emerging Trends LLaVA-Read: Enhancing Reading Ability of Multimodal Language Models

Reference 72

Resolution
verified exact
arxiv_id, observed 2026-05-19T04:42:04.077947Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-19T04:38:49.512293Z digest=sha256:a757b1433da175da2a08b447f9177aadc94ba106c41dec96be9e5fff7f095d12

Observation 994391db-86c6-4abe-b1a3-4383ced3ff08 · inbound

ReforMe: Re-Shaping Documents with Contextual Prompting and Layout-Aware Propagation cites this paper.

ReforMe: Re-Shaping Documents with Contextual Prompting and Layout-Aware Propagation LLaVA-Read: Enhancing Reading Ability of Multimodal Language Models

Reference 55

Resolution
verified exact
arxiv_id, observed 2026-07-02T04:56:38.956998Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-28T08:43:27.671508Z digest=sha256:3de4452ad9e10de56c6eed2d9676e3ea18c2049b24a356ab5d61827bfcab34fd

Observation 94de8d19-2ad0-4711-b9bc-c59ded39dc22 · inbound

DocOCR-Eval: A Correction-Based Framework for OCR Tool Selection Without Ground Truth cites this paper.

DocOCR-Eval: A Correction-Based Framework for OCR Tool Selection Without Ground Truth LLaVA-Read: Enhancing Reading Ability of Multimodal Language Models

Reference 161

Resolution
unresolved
no resolver link, observed 2026-08-02T14:55:37.457570Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T14:55:37.457570Z digest=sha256:edd2bc62e828da8e12d61953f33693d0b87e9321dfa639724c0a5bc40b3b4acb