Pith. sign in

Paper Citation Record · LEDGER

OvisOCR2 Technical Report

As of 21 August 2026, this Paper Citation Record lists 53 of 53 outbound references and 1 inbound Pith citation observation for arXiv:2607.13639.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.13639 v1

Coverage vector

measured 53 of 53 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-02T04:40:07.276016Z

measured 54 of 54 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T21:04:30.112023Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-15T21:04:30.892547Z

Reference resolution

53 of 53 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved53
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a601b3fb-4a16-47c1-a3db-41e0e7da2bda · outbound

This paper cites OmniDocBench: Benchmarking diverse PDF document parsing with comprehensive annotations.

OvisOCR2 Technical Report OmniDocBench: Benchmarking diverse PDF document parsing with comprehensive annotations

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-02T04:40:01.259977Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T04:40:01.259977Z digest=sha256:28a9e5b6665f367d61b34ae3d03b86272166f43c3b96d7ba557277a25927a732

Observation 342ac919-ed8a-48c5-8684-8183f6ef5aff · outbound

This paper cites How Far Is Document Parsing from Solved? PureDocBench: A Source-TraceableBenchmark across Clean, Degraded, and Real-World Settings.

OvisOCR2 Technical Report How Far Is Document Parsing from Solved? PureDocBench: A Source-TraceableBenchmark across Clean, Degraded, and Real-World Settings

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-02T04:40:01.433462Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T04:40:01.433462Z digest=sha256:6e3bad002d26e3e633799530be043ff08c094edbd290fe152b5b67e59c8fd7f2

Observation 561d56b7-659a-4b49-9a46-dfa2faf8daea · outbound

This paper cites PaddleOCR-VL-1.6: Expanding the Frontier of Document Parsing with Under-Optimized Region Refinement and Progressive Post-Training.

OvisOCR2 Technical Report PaddleOCR-VL-1.6: Expanding the Frontier of Document Parsing with Under-Optimized Region Refinement and Progressive Post-Training

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-02T04:40:01.558740Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T04:40:01.558740Z digest=sha256:44c7525fe8cd4726cff2642806e3e754d24410f73eeeb07b5496223a8bac824c

Observation 1c70908e-5519-4314-8101-32f3acdb3f36 · outbound

This paper cites MinerU2.5: A decoupled vision-language model for efficient high-resolution document parsing.

OvisOCR2 Technical Report MinerU2.5: A decoupled vision-language model for efficient high-resolution document parsing

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-02T04:40:01.734149Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T04:40:01.734149Z digest=sha256:b1f9d1e1ddba2abaf29862e6ae577e0ae74a7a2571c217cd38ec44526e993645

Observation fecbc486-1339-429d-ad8e-1e0aca303b73 · outbound

This paper cites GLM-OCR technical report.arXiv preprint arXiv:2603.10910, 2026.

OvisOCR2 Technical Report GLM-OCR technical report.arXiv preprint arXiv:2603.10910, 2026

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-02T04:40:01.896669Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T04:40:01.896669Z digest=sha256:d79d1ff61a76a88be033eb3a636160104da2bdbc277a38fb9bc28df7e48fe5ae

Observation 0edfedcf-0652-405e-ab59-c633a6c43a0f · outbound

This paper cites MinerU2.5-Pro: Pushing the Limits of Data-Centric Document Parsing at Scale.

OvisOCR2 Technical Report MinerU2.5-Pro: Pushing the Limits of Data-Centric Document Parsing at Scale

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-02T04:40:02.010695Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T04:40:02.010695Z digest=sha256:525c5b063c873c49f27a55fbee6e57e902039bccb940f1a03d18669cca372969

Observation 77a061e0-16cb-45fc-a999-06746ea3692d · outbound

This paper cites OmniDocBench public leaderboard.

OvisOCR2 Technical Report OmniDocBench public leaderboard

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-02T04:40:02.123125Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T04:40:02.123125Z digest=sha256:f51d895faed4f50d3ad76b26e16116bf75eb84ed1206b6eb97e73f8e5ebc9f27

Observation 9cb51a65-45cf-4769-9420-086fa4d67eb4 · outbound

This paper cites HunyuanOCR-1.5: Making Lightweight OCR VLMs Faster and Better.

OvisOCR2 Technical Report HunyuanOCR-1.5: Making Lightweight OCR VLMs Faster and Better

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-02T04:40:02.221409Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T04:40:02.221409Z digest=sha256:f648b10978501d32cfa1f03551b21e500da2d99ebe1217e6b6330fc4d3c12869

Observation 7593d479-f94b-48ec-8988-8284cee30d9b · outbound

This paper cites Unlimited OCR Works.

OvisOCR2 Technical Report Unlimited OCR Works

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-02T04:40:02.330789Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T04:40:02.330789Z digest=sha256:394e49b955703b2a50bfe037aadb853b63d454688830db1a8500b2f1383f3681

Observation 858832ca-3502-4c2c-9b30-0a015bc9b0d6 · outbound

This paper cites OvisOCR: End-to-end document parsing via aligning specialized perception with general reasoning.

OvisOCR2 Technical Report OvisOCR: End-to-end document parsing via aligning specialized perception with general reasoning

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-02T04:40:02.464627Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T04:40:02.464627Z digest=sha256:e7d1cda696353e6d09fb1baf194e096030a3f86fec65d163589e6acb785ef02e

Observation ba14643c-53d9-40ad-b957-c8acfc5d6304 · outbound

This paper cites Qwen3.5: Towards native multimodal agents.

OvisOCR2 Technical Report Qwen3.5: Towards native multimodal agents

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-02T04:40:02.618627Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T04:40:02.618627Z digest=sha256:aca2783d4bd3809db2025a2987d6f3f7d24f0b399384ef33a6fc3e2564a6b621

Observation 0ea104eb-fecf-4cb6-a45f-072b90cd4c53 · outbound

This paper cites PaddleOCR-VL-1.5: Towards a Multi-Task 0.9B VLM for Robust In-the-Wild Document Parsing.

OvisOCR2 Technical Report PaddleOCR-VL-1.5: Towards a Multi-Task 0.9B VLM for Robust In-the-Wild Document Parsing

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-02T04:40:02.774648Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T04:40:02.774648Z digest=sha256:eb05d6d3f11cf21cda8451eb5704f5807e325eb4fcbf8db4c0f61dae9a214c09

Observation f8ff5d12-506c-48f1-a8dc-51d597de6758 · outbound

This paper cites Playwright: Fast and reliable end-to-end testing for modern web apps.

OvisOCR2 Technical Report Playwright: Fast and reliable end-to-end testing for modern web apps

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-02T04:40:02.900278Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T04:40:02.900278Z digest=sha256:e1521728769c613f5da76c62311467dcb090eb6ce9335ebeb10681820d593f34

Observation a6e11d44-f663-432e-9c0a-d75b0d536d3d · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

OvisOCR2 Technical Report DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-02T04:40:03.001768Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T04:40:03.001768Z digest=sha256:8aa7c1be2ededf65f97a673f2a8a741ecc9107ad0cac970038f3599db8c03be0

Observation ace0e5dd-05b5-4f59-9cb7-cd94be8003cb · outbound

This paper cites Image over text: Transforming formula recognition evaluation with character detection matching.

OvisOCR2 Technical Report Image over text: Transforming formula recognition evaluation with character detection matching

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-02T04:40:03.067995Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T04:40:03.067995Z digest=sha256:bd5dd538b8aa91f793fa13bc89c683ca78f84d1d1eb4ae5db93e8f2a5e479e95

Observation 977261d8-12af-4eee-83ab-9979444636dd · outbound

This paper cites Image-based table recognition: Data, model, and evaluation.

OvisOCR2 Technical Report Image-based table recognition: Data, model, and evaluation

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-02T04:40:03.115928Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T04:40:03.115928Z digest=sha256:e3fe825cbab9cc759e0867016415be1905e7ea999a03bcc84f5a88b7051ceec9

Observation 0ad0383f-6286-4056-94de-c93cbd1709cf · outbound

This paper cites Distilling the Knowledge in a Neural Network.

OvisOCR2 Technical Report Distilling the Knowledge in a Neural Network

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-02T04:40:03.183618Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T04:40:03.183618Z digest=sha256:11c7d90ee0173590041efa4a3d64e37e08c30d7eacefbe0f5ebfd7ea23e648b1

Observation de4679eb-4997-46e7-9403-808700220e05 · outbound

This paper cites On-policy distillation of language models: Learning from self-generated mistakes.

OvisOCR2 Technical Report On-policy distillation of language models: Learning from self-generated mistakes

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-02T04:40:03.245742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T04:40:03.245742Z digest=sha256:50cb2a8224026048c43d6993869233e48a650371e2a2e4c885cae1764e727d44

Observation aae4d24f-98ac-4a3a-9715-83769c4ec3f0 · outbound

This paper cites Rethinking On-Policy Distillation of Large Language Models: Phenomenology, Mechanism, and Recipe.

OvisOCR2 Technical Report Rethinking On-Policy Distillation of Large Language Models: Phenomenology, Mechanism, and Recipe

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-02T04:40:03.309100Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T04:40:03.309100Z digest=sha256:36855d2532e56154ba3c0ca354bae40dcd4de7636f5415f6122f2f5dfe5b176f

Observation fa331f93-7164-4e3e-b2d8-ba7077e705bb · outbound

This paper cites Morcos, Hongseok Namkoong, Ali Farhadi, Yair Carmon, Simon Kornblith, and Ludwig Schmidt.

OvisOCR2 Technical Report Morcos, Hongseok Namkoong, Ali Farhadi, Yair Carmon, Simon Kornblith, and Ludwig Schmidt

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-02T04:40:03.357692Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T04:40:03.357692Z digest=sha256:26bca289531d62132674994940805b0d96206d34f6ac466c409de1c1b7e5c678

Observation 8b0ec144-a422-4278-8c43-062b89666d18 · outbound

This paper cites InternVL3.5: Advancing Open-Source Multimodal Models in Versatility, Reasoning, and Efficiency.

OvisOCR2 Technical Report InternVL3.5: Advancing Open-Source Multimodal Models in Versatility, Reasoning, and Efficiency

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-02T04:40:03.411990Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T04:40:03.411990Z digest=sha256:c492b4ba3dbd48e2c85b797f54cced8b96149a57a3fc1ba0dd3d60eb304717d8

Observation b7e2f639-bb24-4df3-b406-2c94a2b196f0 · outbound

This paper cites Kimi K2.5: Visual Agentic Intelligence.

OvisOCR2 Technical Report Kimi K2.5: Visual Agentic Intelligence

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-02T04:40:03.486812Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T04:40:03.486812Z digest=sha256:a9cb7ed19e305c4d6cb1d501b63b531ffc4fad8c18f883586c3edf1cfbe45d46

Observation 080620d8-886d-4a74-ba32-347ede3f2c43 · outbound

This paper cites Update to GPT-5 system card: GPT-5.2.

OvisOCR2 Technical Report Update to GPT-5 system card: GPT-5.2

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-02T04:40:03.564849Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T04:40:03.564849Z digest=sha256:7c7f402b1869e6f692bbfd0722e1ad8264bb15ab6c66b7d668d060912294862d

Observation 93be6811-84da-4b13-83e2-3cc77de566b1 · outbound

This paper cites Qwen3-VL Technical Report.

OvisOCR2 Technical Report Qwen3-VL Technical Report

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-02T04:40:03.683731Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T04:40:03.683731Z digest=sha256:e080b729fc932f6a01bfd7da7dfac2a321a3bcf85e4e02aadc1cb349963882f5

Observation 12806a7b-92a2-4696-b13a-e28489aaaca0 · outbound

This paper cites Gemini 3 Flash model card.

OvisOCR2 Technical Report Gemini 3 Flash model card

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-02T04:40:03.799418Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T04:40:03.799418Z digest=sha256:69a73740b7167d3fd29cc8d6c60a6f1e56679ad2aeb97c8850fd78ad02f1e864

Observation e28a9df1-47a8-4f59-b92c-f11694bae833 · outbound

This paper cites Gemini 3 Pro model card.

OvisOCR2 Technical Report Gemini 3 Pro model card

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-02T04:40:03.936431Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T04:40:03.936431Z digest=sha256:8fa59d3a82e2dc1d8e3b70d12deaf7c1673970482416e021068b0e2c1fe6406b

Observation e4602724-477f-43ad-9c91-a27c0c116c1d · outbound

This paper cites Ovis2.5 Technical Report.

OvisOCR2 Technical Report Ovis2.5 Technical Report

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-02T04:40:04.052317Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T04:40:04.052317Z digest=sha256:899e6574e438579d54dea6d1e60a886b1d3236b0512a148f37f2b9ab01621451

Observation 9347515a-44e9-40d3-8dd6-0c289bc36821 · outbound

This paper cites Dolphin: Document image parsing via heterogeneous anchor prompting.

OvisOCR2 Technical Report Dolphin: Document image parsing via heterogeneous anchor prompting

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-02T04:40:04.223297Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T04:40:04.223297Z digest=sha256:56e1ce154619373607fb1ec709750967938614b26f0374382e7dd2bb7246b7f1

Observation f5fab770-0088-4f78-ac28-052a9ea45d80 · outbound

This paper cites MonkeyOCR: Document parsing with a structure-recognition-relation triplet paradigm.arXiv preprint arXiv:2506.05218, 2025.

OvisOCR2 Technical Report MonkeyOCR: Document parsing with a structure-recognition-relation triplet paradigm.arXiv preprint arXiv:2506.05218, 2025

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-02T04:40:04.397723Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T04:40:04.397723Z digest=sha256:e9ffa475f142982fa049425356ae63ee646d3313ffb6301e71196be0cc8d6e62

Observation 647786e6-5233-4b77-b337-d1e8c995a3c0 · outbound

This paper cites Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding.

OvisOCR2 Technical Report Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-02T04:40:04.566211Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T04:40:04.566211Z digest=sha256:84bf0333f0210b20e44cd6ad655a77f4e3e3ce061fbf6dccf40b3b404a2f66ca

Observation 5f57399f-a3b8-4903-a3a1-5cf0e119f0e8 · outbound

This paper cites PaddleOCR-VL: Boosting multilingual document parsing via a 0.9B ultra-compact vision-language model.arXiv preprint arXiv:2510.14528, 2025.

OvisOCR2 Technical Report PaddleOCR-VL: Boosting multilingual document parsing via a 0.9B ultra-compact vision-language model.arXiv preprint arXiv:2510.14528, 2025

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-02T04:40:04.742421Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T04:40:04.742421Z digest=sha256:3e756577be475bfa24aec35e554e2cd24d29a14a363d101dd6d988f5e088e59b

Observation 8aa7ac39-52df-463f-9b88-476e36bc0fc0 · outbound

This paper cites POINTS-Reader: Distillation-free adapta- tion of vision-language models for document conversion.

OvisOCR2 Technical Report POINTS-Reader: Distillation-free adapta- tion of vision-language models for document conversion

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-02T04:40:04.916438Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T04:40:04.916438Z digest=sha256:c98df052b501d6b23b8a304ed6ce35cf852626212195993b60abe8215b648540

Observation 2ac95d85-434a-41b1-8e21-c4ea689c2c5e · outbound

This paper cites Nanonets-OCR-S: A model for transforming documents into structured Markdown with intelligent con- tent recognition and semantic tagging.

OvisOCR2 Technical Report Nanonets-OCR-S: A model for transforming documents into structured Markdown with intelligent con- tent recognition and semantic tagging

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-02T04:40:05.060855Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T04:40:05.060855Z digest=sha256:a1e92be0726c575e9cae02ff0ac86949ff80ceb8cab5ee3c55c4c3215af44d4b

Observation 5acb6a12-f423-4a81-af96-64f0d31c315f · outbound

This paper cites olmOCR: Unlocking trillions of tokens in PDFs with vision language models.arXiv preprint arXiv:2502.18443, 2025.

OvisOCR2 Technical Report olmOCR: Unlocking trillions of tokens in PDFs with vision language models.arXiv preprint arXiv:2502.18443, 2025

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-02T04:40:05.227277Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T04:40:05.227277Z digest=sha256:955de8f36fc9307cd1d388e8941138ff3fd9ca218d7021d6264a73be268a66c6

Observation 786c03fe-5a27-4cf0-ab43-bfd09bd5caef · outbound

This paper cites OCRVerse: Towards holistic OCR in end-to- end vision-language models.arXiv preprint arXiv:2601.21639, 2026.

OvisOCR2 Technical Report OCRVerse: Towards holistic OCR in end-to- end vision-language models.arXiv preprint arXiv:2601.21639, 2026

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-02T04:40:05.348245Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T04:40:05.348245Z digest=sha256:22564fa1a22b395f24320e6b65c75e29284a48b877076dd159e674bb7404aca5

Observation 8c5459fb-1557-4d43-8702-278a9f0d210e · outbound

This paper cites HunyuanOCR technical report.arXiv preprint arXiv:2511.19575, 2025.

OvisOCR2 Technical Report HunyuanOCR technical report.arXiv preprint arXiv:2511.19575, 2025

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-02T04:40:05.529028Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T04:40:05.529028Z digest=sha256:847f3e8c4e449d78de4bf1d8d37d108ba5c6705f73ce05c3be00fc34e2b3942a

Observation 4d548933-cc80-464b-b5c0-4e713bd8fbf8 · outbound

This paper cites DeepSeek-OCR 2: Visual causal flow.arXiv preprint arXiv:2601.20552, 2026.

OvisOCR2 Technical Report DeepSeek-OCR 2: Visual causal flow.arXiv preprint arXiv:2601.20552, 2026

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-02T04:40:05.705502Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T04:40:05.705502Z digest=sha256:7ee29c09fe7d147a3881e113561b4a7df26fcab2f52384c47eed51e20a06dde2

Observation 5f496906-8501-48a0-a702-47aab7217bf7 · outbound

This paper cites UniRec-0.1B: Unified Text and Formula Recognition with 0.1B Parameters.

OvisOCR2 Technical Report UniRec-0.1B: Unified Text and Formula Recognition with 0.1B Parameters

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-02T04:40:05.856964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T04:40:05.856964Z digest=sha256:16e2875ca1de32c232cfb111ad8149f64107372eb3d5105e4233db73e1c4c431

Observation 18758040-9c4b-45ab-9e63-90309e27d03a · outbound

This paper cites dots.ocr: Multilingual document layout parsing in a single vision-language model.arXiv preprint arXiv:2512.02498, 2025.

OvisOCR2 Technical Report dots.ocr: Multilingual document layout parsing in a single vision-language model.arXiv preprint arXiv:2512.02498, 2025

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-02T04:40:05.945266Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T04:40:05.945266Z digest=sha256:86101d1d90ff1c446bbca92fac4b98fc8d9c3b9a901c7a1d77cb8f60b5f57ea4

Observation ec373b6c-f67e-4a77-a479-7c2616ecea43 · outbound

This paper cites FireRed-OCR technical report.arXiv preprint arXiv:2603.01840, 2026.

OvisOCR2 Technical Report FireRed-OCR technical report.arXiv preprint arXiv:2603.01840, 2026

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-02T04:40:06.002608Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T04:40:06.002608Z digest=sha256:41162f9561c17c19a05db8af1dadd377d42e6f7fdcf717379aa90f5c46905e47

Observation 76cdc04c-1bd3-4f8e-af27-be721ce3e435 · outbound

This paper cites ABot-OCR Technical Report.

OvisOCR2 Technical Report ABot-OCR Technical Report

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-02T04:40:06.065397Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T04:40:06.065397Z digest=sha256:d57211d33f2450014613ad391f93f4d2f56b75d415b190b2f3d889c233b3d3dc

Observation 7d3e9152-7027-4445-88ba-2bcb13dd1e5f · outbound

This paper cites Logics-Parsing technical report.arXiv preprint arXiv:2509.19760, 2025.

OvisOCR2 Technical Report Logics-Parsing technical report.arXiv preprint arXiv:2509.19760, 2025

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-02T04:40:06.184917Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T04:40:06.184917Z digest=sha256:a1ab0f74e266b9d282b7901d9aa0095ce72a3cf4025f6868a3d03c43ed905c5f

Observation cb94db28-c423-4db1-a266-3a990723e9e4 · outbound

This paper cites Qianfan-OCR: A unified end-to-end model for document intelligence.arXiv preprint arXiv:2603.13398, 2026.

OvisOCR2 Technical Report Qianfan-OCR: A unified end-to-end model for document intelligence.arXiv preprint arXiv:2603.13398, 2026

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-02T04:40:06.257726Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T04:40:06.257726Z digest=sha256:9b8916439c02e99924f19523e983a33bd4e29628d95d3f30d07443355f1c8025

Observation 2759cf15-214a-42dc-9d09-6894885d6a31 · outbound

This paper cites MiniCPM-V 4.5: Cooking Efficient MLLMs via Architecture, Data, and Training Recipe.

OvisOCR2 Technical Report MiniCPM-V 4.5: Cooking Efficient MLLMs via Architecture, Data, and Training Recipe

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-02T04:40:06.407090Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T04:40:06.407090Z digest=sha256:eb8b4b120bdefff1fc5fc88f26ffcfe4299cf286a0113aa6b0648cc29c6f55db

Observation d935928b-fbec-4d0c-97f1-b07a249d100f · outbound

This paper cites STEP3-VL-10B technical report.arXiv preprint arXiv:2601.09668, 2026.

OvisOCR2 Technical Report STEP3-VL-10B technical report.arXiv preprint arXiv:2601.09668, 2026

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-02T04:40:06.539789Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T04:40:06.539789Z digest=sha256:1a3d5f783e626ed770b57aceacba53976116cf2a6c7bd1060a80e036c873d0b6

Observation 67fe7ca1-a60a-4d6a-ac35-e008dd911e99 · outbound

This paper cites Gemini 3.1 Pro model card.

OvisOCR2 Technical Report Gemini 3.1 Pro model card

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-02T04:40:06.648330Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T04:40:06.648330Z digest=sha256:8fa29fa1d5ccdab6ddd92c03084db0bfbd6e98705e761037d12f2f543b2a6a44

Observation 4fe4d002-db3b-44ca-815a-34b5e7931e2e · outbound

This paper cites SVTRv2: CTC beats encoder-decoder models in scene text recognition.

OvisOCR2 Technical Report SVTRv2: CTC beats encoder-decoder models in scene text recognition

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-02T04:40:06.738106Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T04:40:06.738106Z digest=sha256:86d4f092f684bb65187e5d1ca99563f79754f9111d4f1a7a0985782054caf1cd

Observation 232b8b5b-3feb-4af5-a67b-b03839b48943 · outbound

This paper cites Multimodal OCR: Parse anything from documents.

OvisOCR2 Technical Report Multimodal OCR: Parse anything from documents

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-02T04:40:06.845336Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T04:40:06.845336Z digest=sha256:e2a97714f89e9ff1d411a066b5c99a07d0e38f387d84003b9aaaef7fbfa6b406

Observation b4f74ad1-7196-4ab8-ab47-1c68f1bb7e53 · outbound

This paper cites OCRFlux: A multimodal toolkit for converting documents into Markdown.

OvisOCR2 Technical Report OCRFlux: A multimodal toolkit for converting documents into Markdown

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-02T04:40:06.915445Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T04:40:06.915445Z digest=sha256:a8b988b28b1bf95eed9430d4d06624712995b10d08a5f3eebc3fde175bc2ad30

Observation 517f7cb8-4770-4624-9eb5-35f7c24ab57f · outbound

This paper cites DeepSeek-OCR: Contexts Optical Compression.

OvisOCR2 Technical Report DeepSeek-OCR: Contexts Optical Compression

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-02T04:40:07.008683Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T04:40:07.008683Z digest=sha256:f726c39f9bd359c449a13ba39c483b87b6aabf10e4e6028626805b6b4a2c65a1

Observation 0a2c2953-2cda-40e5-850c-a3c349e9c48f · outbound

This paper cites Nanonets-OCR2: A model for transforming documents into structured Markdown with intelligent content recognition and semantic tagging.

OvisOCR2 Technical Report Nanonets-OCR2: A model for transforming documents into structured Markdown with intelligent content recognition and semantic tagging

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-02T04:40:07.106147Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T04:40:07.106147Z digest=sha256:2457f6d87a9c1531792295dd2fdec86b5e1477fc2bfc83ae4079b5ee442b935a

Observation d1d0d39a-5d06-4e20-b9f0-d9c746f9188a · outbound

This paper cites olmOCR 2: Unit test rewards for document OCR.arXiv preprint arXiv:2510.19817, 2025.

OvisOCR2 Technical Report olmOCR 2: Unit test rewards for document OCR.arXiv preprint arXiv:2510.19817, 2025

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-02T04:40:07.183985Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T04:40:07.183985Z digest=sha256:e1017a5b4492b32911f311ee3f4091e52c5d4a980235cceb7f9a0624be60d391

Observation b3a049d1-e6e6-4066-8131-aba159c19f83 · outbound

This paper cites Reading or reasoning? format decoupled reinforcement learning for document OCR.

OvisOCR2 Technical Report Reading or reasoning? format decoupled reinforcement learning for document OCR

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-02T04:40:07.276016Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T04:40:07.276016Z digest=sha256:d660c022367ce7a9f65c51d71bf583fd2a4f18eb7dab79406a9b1dd2fdc110da

Pith citing papers

Observation d9fd1a58-12db-42cd-9077-a4924ef5a80a · inbound

NaviDC-OCR: Navigating Document Parsing Across Digital and Camera-Captured Documents cites this paper.

NaviDC-OCR: Navigating Document Parsing Across Digital and Camera-Captured Documents OvisOCR2 Technical Report

Reference 62

Resolution
metadata mismatch
local_arxiv, observed 2026-08-15T21:04:30.897654Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-15T21:04:30.112023Z digest=sha256:f4c73ece73daef39fa9a9532821dc6ead269912ddbef5b338fd1c77353e23820