Pith. sign in

Paper Citation Record · LEDGER

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models

As of 9 August 2026, this Paper Citation Record lists 57 of 57 outbound references and 0 inbound Pith citation observations for arXiv:2607.21617.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.21617 v1

Coverage vector

measured 57 of 57 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-02T12:46:12.113877Z

measured 57 of 57 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

57 of 57 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved57
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 2506e149-4c5d-4c23-b8c7-5e1bec7c55c7 · outbound

This paper cites an unresolved cited work.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models Unresolved cited work

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:05.563971Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:05.563971Z digest=sha256:0e12b6b5617c1c2d778055dda25de5943e997a4e21f495ef833e069e94072376

Observation 12ce943d-fa6f-4641-bbe1-d67a27e3d30d · outbound

This paper cites an unresolved cited work.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models Unresolved cited work

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:05.652444Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:05.652444Z digest=sha256:4cd9f82773b621137651733fddead03ba0d91dfc68bdf1a376a41c9e40652c50

Observation aa56fc46-2ecc-4925-83b7-d8274caf0537 · outbound

This paper cites Journal of Foo , volume = 13, number = 1, pages =.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models Journal of Foo , volume = 13, number = 1, pages =

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:05.817702Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:05.817702Z digest=sha256:570b1aae130e196c7caca098b255057e2669eef38792231bc71b0663270e9866

Observation 3feaa8fe-548c-4cb2-ae68-fd73b28f048a · outbound

This paper cites Journal of Foo , volume = 14, number = 1, pages =.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models Journal of Foo , volume = 14, number = 1, pages =

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:05.895645Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:05.895645Z digest=sha256:096ece2577f8b01bb4540fed80ab2f2916e8442bd3fb5b98a6865d9b91b6e575

Observation 4d39162e-ad9d-4bcd-a501-f0a62e46ae3c · outbound

This paper cites Trends in Cognitive Sciences , volume =.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models Trends in Cognitive Sciences , volume =

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:06.004527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:06.004527Z digest=sha256:db72b1c26ca00b95c4999c38461de1a406822f4ee3d7aeef60381bd6456763b3

Observation e75dceea-c1bf-4d75-be06-48461e49af6c · outbound

This paper cites an unresolved cited work.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models Unresolved cited work

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:06.101706Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:06.101706Z digest=sha256:368612b659a1572213d759fb800b0a40ab3fc709c0c5e8bb634e1efe41c339b2

Observation 81ad02b1-c20c-4302-9581-236f3c276028 · outbound

This paper cites arXiv preprint arXiv:2601.20552 , year=.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models arXiv preprint arXiv:2601.20552 , year=

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:06.210404Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:06.210404Z digest=sha256:94a80c72d7f063c158820e0ab967eebdc400b5faea945063f53dc6a4addfcac0

Observation 495cdd24-5398-410a-bef2-65f8dd430fe7 · outbound

This paper cites an unresolved cited work.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models Unresolved cited work

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:06.415503Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:06.415503Z digest=sha256:ab7f1c34b617385c2a8d4b5f29a64e7ffb2bf31a47ae388e334f69f880cc09bb

Observation e5209757-c56d-42c5-b2ce-b6d55e9b39a2 · outbound

This paper cites 5: A decoupled vision-language model for efficient high-resolution document parsing , author=.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models 5: A decoupled vision-language model for efficient high-resolution document parsing , author=

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:06.591204Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:06.591204Z digest=sha256:7045c0c587b038d9ada7bfd3b496fb6d6af4922a775a60fd57d64aa2cb96be30

Observation 00aff762-07d1-4c8a-b231-06e30cb5a17e · outbound

This paper cites READoc: A Unified Benchmark for Realistic Document Structured Extraction.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models READoc: A Unified Benchmark for Realistic Document Structured Extraction

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:06.801715Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:06.801715Z digest=sha256:c4796b2e4f6cf807d8acdac8854a662ffefe99d9e4b73b0a0911f28c344dad1f

Observation e8fb02da-0a11-4c58-b63e-8f2d9080f8f1 · outbound

This paper cites Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:06.931296Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:06.931296Z digest=sha256:4b9a493c53e719410857833cb48980e1465d3936136af3ce22642d5d78a79e37

Observation b1c0290a-9c00-43fa-9366-08f455c5770f · outbound

This paper cites Qwen3-VL Technical Report.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models Qwen3-VL Technical Report

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:07.069203Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:07.069203Z digest=sha256:c70fd4ae12980ae843059415e3c4f98762cb9913b0cf99704651a0a6e9d2c954

Observation 58f3bfe7-80bf-4b5e-ad40-49f0033ef59f · outbound

This paper cites an unresolved cited work.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models Unresolved cited work

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:07.238033Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:07.238033Z digest=sha256:d007976a37ca463e46507f05c4b32c217e0bd2b6e116968fa962f7b594afa814

Observation f62387f8-0529-431a-a407-c6c2a6b700bb · outbound

This paper cites 2026 , eprint=.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models 2026 , eprint=

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:07.417593Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:07.417593Z digest=sha256:b97aac580a55eab590d72e9bb6de36c0cce7e216e25c3965f3af39c774314368

Observation a3473b54-a9c2-43a3-8d74-158a70e4d2fe · outbound

This paper cites Proceedings of the IEEE international conference on computer vision , pages=.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models Proceedings of the IEEE international conference on computer vision , pages=

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:07.559805Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:07.559805Z digest=sha256:926e05aff890b57ab4da4e237165553cd284a67e7e8fbf40596c87108e1c98a2

Observation 117d116e-3141-421d-96f4-800bea718c1f · outbound

This paper cites 2021 , eprint=.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models 2021 , eprint=

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:07.675539Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:07.675539Z digest=sha256:421f491a935a6f64fb8a4e98ba6bfcc2bbbf6fa0e99321d23506f0c4e895f7c1

Observation 9fa0babf-6708-4460-93d5-4728f1b8b621 · outbound

This paper cites 2024 , eprint=.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models 2024 , eprint=

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:07.737098Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:07.737098Z digest=sha256:2dae9467e17e80445a96214cd8c5903c6535e89af0b9edb8de5e0372ebe8eab1

Observation d095c590-24f6-470f-a039-d5219e9a0152 · outbound

This paper cites InternVL3.5: Advancing Open-Source Multimodal Models in Versatility, Reasoning, and Efficiency.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models InternVL3.5: Advancing Open-Source Multimodal Models in Versatility, Reasoning, and Efficiency

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:07.800275Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:07.800275Z digest=sha256:1cb4a777e897ae0516e5d6902bce7bd915f58303ce306ff5b686e0fa10544790

Observation f1f44dd4-73b9-4bac-9200-c35f04826d3f · outbound

This paper cites Visual Instruction Tuning , url =.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models Visual Instruction Tuning , url =

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:07.865094Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:07.865094Z digest=sha256:67672e9cca8a45368c8c7db4f2ff226d7a846f7d400aa50e39517c401b0c7d3f

Observation a88e607f-c268-407b-b433-a6ce145eb330 · outbound

This paper cites GPT-4 Technical Report.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models GPT-4 Technical Report

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:07.918581Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:07.918581Z digest=sha256:a812f45ac1ddbe3e5e5b70b4fb293afc73b1a964dc3422fc8d67c53c8deec1ad

Observation 94eab9b6-08c6-4345-8d31-fe76ff7847b2 · outbound

This paper cites 2023 , eprint=.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models 2023 , eprint=

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:08.056765Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:08.056765Z digest=sha256:df00084f3ecd727e32085afce5fee8968e899b3e913dcf603ab78e5973aa4711

Observation cf9bbd9e-5b13-4e88-852f-78888a90c54a · outbound

This paper cites OCRBench: on the hidden mystery of OCR in large multimodal models , volume=.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models OCRBench: on the hidden mystery of OCR in large multimodal models , volume=

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:08.181936Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:08.181936Z digest=sha256:650c3d125f5f309a8127da1af55aa500805035690617eb1e4a58cf1151412f34

Observation 3266d655-2118-45e4-8058-452e9f17999f · outbound

This paper cites 2025 , eprint=.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models 2025 , eprint=

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:08.377781Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:08.377781Z digest=sha256:41529e388416ee1b99cbd42c5fc7c32e0049acdbb26d9be01c42d7135463a207

Observation 1b2d3719-4388-4112-93f2-d4aed75c1894 · outbound

This paper cites 2021 , eprint=.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models 2021 , eprint=

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:08.539837Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:08.539837Z digest=sha256:73461e603a7ed91bbc6337fb6fb2c1d331c0270e0c7e97b9562948fbf66804f8

Observation 96228a5e-dbda-4b47-b000-a1a800e4b1d1 · outbound

This paper cites 2024 , eprint=.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models 2024 , eprint=

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:08.733743Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:08.733743Z digest=sha256:61592e73e6c5f966a4028fb98997147f59604e54747a9581c15529cd626592f7

Observation acdba914-2f72-4139-b604-3dcb96c3bd33 · outbound

This paper cites 2020 , eprint=.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models 2020 , eprint=

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:08.880444Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:08.880444Z digest=sha256:dd5703ed99474849982aa2be6e83c44778c4befff9d2287bd4846aa844ce38f0

Observation 83be50c9-17ef-4e33-92d1-383f440ca962 · outbound

This paper cites Survey of Hallucination in Natural Language Generation , volume=.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models Survey of Hallucination in Natural Language Generation , volume=

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:09.009817Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:09.009817Z digest=sha256:25f01c3b70aa765f3afe30f1e743a2505802fb8ac49a67d4b82fd25edfd05ceb

Observation cc7a3d58-72f4-4f7a-9a39-c66ac86c1eb2 · outbound

This paper cites 2023 , eprint=.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models 2023 , eprint=

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:09.137858Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:09.137858Z digest=sha256:a6a108371c95dd49b88ee99f4215a4701970b62dc242e8b87d29f6cfdde016f6

Observation 7a509ac8-3a29-45a8-88a2-9077eb0490e1 · outbound

This paper cites Forty-second International Conference on Machine Learning , year=.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models Forty-second International Conference on Machine Learning , year=

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:09.278951Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:09.278951Z digest=sha256:ec9328fbafd7e3870ca24127a23d9258eb7ba15f54ba6dcf5259870af4709143

Observation 6fa70d81-a0e3-4978-a9d0-b120f9b03f9d · outbound

This paper cites Linux Journal , volume=.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models Linux Journal , volume=

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:09.381601Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:09.381601Z digest=sha256:88b125780a38ae9c624cb812255054b4b4039af762ecbf893eaf10f8ab6bb830

Observation e67333de-d374-4cc8-bede-b4727713c7b8 · outbound

This paper cites National Science Review , volume=.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models National Science Review , volume=

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:09.504605Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:09.504605Z digest=sha256:13cfc9f1e9cb988713fdd34c7aaab36bb69ae8876d2142f752c098ab388260c5

Observation ccaf5634-fbf2-4422-9539-f4bee3b06dcf · outbound

This paper cites IEEE transactions on pattern analysis and machine intelligence , volume=.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models IEEE transactions on pattern analysis and machine intelligence , volume=

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:09.576749Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:09.576749Z digest=sha256:3110eadb15aa1102c4810eeb5e1cf0beb2e44af2415129be9de192d6a75b1e9d

Observation f0a64ca0-7550-4def-924a-95cf1737f7f5 · outbound

This paper cites LayoutLM: Pre-training of Text and Layout for Document Image Understanding , url=.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models LayoutLM: Pre-training of Text and Layout for Document Image Understanding , url=

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:09.648762Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:09.648762Z digest=sha256:41270bcde66978b067ee88e5b8f162afa9b3e59f3506bd58f8880545cb231040

Observation 350f618f-059a-41b7-ba55-954416cb2379 · outbound

This paper cites Ocean-OCR: Towards General OCR Application via a Vision-Language Model.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models Ocean-OCR: Towards General OCR Application via a Vision-Language Model

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:09.712264Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:09.712264Z digest=sha256:5cec426ae25805cfe614051761ef140f0a8ac0088b52e83d8f3a952dc2f70e24

Observation f5dcf2d0-eb38-4d89-90a1-76b8b8a6d2bd · outbound

This paper cites arXiv preprint arXiv:2502.18443 , year=.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models arXiv preprint arXiv:2502.18443 , year=

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:09.786583Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:09.786583Z digest=sha256:ba63eef3a9acaf71cf6c8a0680290bdd5c119376350eb0a7fb4067e0dfae0b66

Observation a1d801ed-00cc-44b4-a674-a62160a0b3f4 · outbound

This paper cites MinerU: An Open-Source Solution for Precise Document Content Extraction.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models MinerU: An Open-Source Solution for Precise Document Content Extraction

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:09.834605Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:09.834605Z digest=sha256:059e719d502b1b785e9e783c3498b8f04f0b80e0c94051767a77b6af6815c1fa

Observation 6cf654f6-b8e2-46b7-bfd4-dc8472285628 · outbound

This paper cites PP-OCR: A Practical Ultra Lightweight OCR System.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models PP-OCR: A Practical Ultra Lightweight OCR System

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:09.903700Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:09.903700Z digest=sha256:e113b663f9719a2b9c8441059113ba7135e310f212bc879e1114ff08e22b58c5

Observation 3208d8a3-ba0b-4477-a22a-a040458141d6 · outbound

This paper cites Ninth international conference on document analysis and recognition (ICDAR 2007) , volume=.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models Ninth international conference on document analysis and recognition (ICDAR 2007) , volume=

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:09.982617Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:09.982617Z digest=sha256:3aee30eca249789260c0229762998f8b478c3358b6a4efcfe2e7f72eba6de748

Observation c6a80f1e-fd10-4168-98e4-4b78c9ac3e91 · outbound

This paper cites Proceedings of the 33rd ACM International Conference on Multimedia , pages=.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models Proceedings of the 33rd ACM International Conference on Multimedia , pages=

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:10.083752Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:10.083752Z digest=sha256:52a867c9bd5ba1209e353632c2559c6f3b4aedc3f64bf8755f24d97928f2d14a

Observation 7fbb042b-3d26-450e-a81c-df8ea532cd38 · outbound

This paper cites Journal of machine learning research , volume=.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models Journal of machine learning research , volume=

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:10.148535Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:10.148535Z digest=sha256:d0ce347f3d55d95f0c0ecde53543c7c4a5e0ce5e7cf247300b5eb716204cd2e7

Observation 1ff74ba2-0d7b-4154-9a18-52ad8d5f86df · outbound

This paper cites 2021 , publisher =.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models 2021 , publisher =

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:10.215369Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:10.215369Z digest=sha256:6bfe95cb787b559f13f94ba5015c5c6b11fa8c96eb9948a9e654c62f3555d3ee

Observation d790d216-96c4-4b5a-a2b9-a2aa1327bccd · outbound

This paper cites 2025 , howpublished =.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models 2025 , howpublished =

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:10.467072Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:10.467072Z digest=sha256:a21d56a492d9ddead1751e7f9b586b820001f6049733e0541cbbb0e248e6582d

Observation 4ff9148a-d2d0-41b5-8e2f-77fde8d1b489 · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models Gemini: A Family of Highly Capable Multimodal Models

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:10.631021Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:10.631021Z digest=sha256:15bd75792a2ef904a5c57d5eba92a9436a7cc9312f40b0b19d21af8aa32552d4

Observation 4be17acc-ce43-4ce5-a1d3-191c13f5a822 · outbound

This paper cites ArXiv , year=.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models ArXiv , year=

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:10.833031Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:10.833031Z digest=sha256:bedd1fbd113c7a139fb0061fe2476de8012d49d785e4cedbdafd6daaa590c905

Observation 7e7b6a7c-e5c9-4bd9-b3dd-fc45d1c41fbb · outbound

This paper cites Synthetic and Natural Noise Both Break Neural Machine Translation.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models Synthetic and Natural Noise Both Break Neural Machine Translation

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:10.997571Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:10.997571Z digest=sha256:b691925bf2aefe183c57681fd4160e8ddd3a5bdd1d8b0111eaf97e4a9b0b2add

Observation 830efd6d-2cf9-4606-8a6b-4d1427ef2019 · outbound

This paper cites Proceedings of the 56th Annual Meeting of the Association for Computational Linguistics (Volume 2: Short Papers) , pages=.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models Proceedings of the 56th Annual Meeting of the Association for Computational Linguistics (Volume 2: Short Papers) , pages=

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:11.160831Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:11.160831Z digest=sha256:a4804fb3478c45b45d8f830f4a4d54c59261d175cade15906c546b251c37463a

Observation 06c4e05f-6f1c-4428-84b8-3618773331e7 · outbound

This paper cites Proceedings of the 57th Annual Meeting of the Association for Computational Linguistics , pages=.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models Proceedings of the 57th Annual Meeting of the Association for Computational Linguistics , pages=

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:11.255823Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:11.255823Z digest=sha256:418d72264f8c36cfb2c8bd34922db4fa664d423d90f326ffbfab4393edb0275c

Observation 3bc4cac9-7383-4f29-b767-246a0b6046e6 · outbound

This paper cites Proceedings of the 2023 Conference on Empirical Methods in Natural Language Processing , pages=.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models Proceedings of the 2023 Conference on Empirical Methods in Natural Language Processing , pages=

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:11.345358Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:11.345358Z digest=sha256:1d6bc8ce615851401ddb29d458be036627260424a09a87763eba2ccda90f72b9

Observation 35c88fa4-4807-4b80-a29a-539699e3a66e · outbound

This paper cites Qwen3.5-Omni Technical Report.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models Qwen3.5-Omni Technical Report

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:11.456166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:11.456166Z digest=sha256:6027b86863ae5186f99e33969bdb2ccfd84eece5f349f720eaf51f3e1e87cc82

Observation bdbc8cab-a8f3-4595-baf1-9cc026a839ad · outbound

This paper cites arXiv preprint arXiv:2510.17771 , year=.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models arXiv preprint arXiv:2510.17771 , year=

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:11.588458Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:11.588458Z digest=sha256:167f25d48237af3febc6925232234e4d3150d8525168863baf74651a12b1e4a5

Observation 90c8079e-6d5d-48f6-a6c6-372bf49b2598 · outbound

This paper cites Vision Language Models are Biased.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models Vision Language Models are Biased

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:11.638518Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:11.638518Z digest=sha256:8454e7bbccb7113f7d728feadf0ebfdb0cce6317c52d1ffdc81c63eaf6c2aa04

Observation d01c5297-61e0-4901-bdd7-ef70a5c2b097 · outbound

This paper cites Findings of the Association for Computational Linguistics: NAACL 2025 , pages=.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models Findings of the Association for Computational Linguistics: NAACL 2025 , pages=

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:11.740750Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:11.740750Z digest=sha256:a21ecfe969433109fe28ca0bf3e70d4a4c23961f47a900575317a212b53a20d1

Observation 02b638c6-b96b-4c5b-b6dd-8f5c1690261e · outbound

This paper cites Proceedings of the 2023 conference on empirical methods in natural language processing , pages=.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models Proceedings of the 2023 conference on empirical methods in natural language processing , pages=

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:11.787638Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:11.787638Z digest=sha256:16133666ee3181dff31a856cf98b848263407bf52f5aaeb0c4356a72127d19cc

Observation 922e6cd1-b60e-4253-ad2c-c0c94c0dc56c · outbound

This paper cites Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:11.856663Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:11.856663Z digest=sha256:9b50cedd6bdb2d364668c28646da768b7e40142e7cd4dc2fd6609aff5cdfc32e

Observation 8f392fc2-ebd7-4c9d-b058-17b934684485 · outbound

This paper cites AMBER: An LLM-free Multi-dimensional Benchmark for MLLMs Hallucination Evaluation.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models AMBER: An LLM-free Multi-dimensional Benchmark for MLLMs Hallucination Evaluation

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:11.926310Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:11.926310Z digest=sha256:fe5d61b2273a768d55cb2a64edbd2b0af3a8956d58c76eb41e09a28242ab5b7f

Observation 29d1489b-8bfe-4abc-8a9d-3c27e3256531 · outbound

This paper cites 2026 , howpublished=.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models 2026 , howpublished=

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:11.997842Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:11.997842Z digest=sha256:063a57a2139a5ff80c34bf8b5b531bc93d8d2a9def6c84e32ee40d1a23e98b76

Observation f9e2377e-e515-40a3-92e3-f7617384f504 · outbound

This paper cites OpenAI GPT-5 System Card.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models OpenAI GPT-5 System Card

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:12.113877Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:12.113877Z digest=sha256:6578d307ce183949b2d2961b11ce078458680e3e26ffbf11fe6b0ed8e2a97c6d

Pith citing papers

No inbound Pith citation observations are available.