Pith. sign in

Paper Citation Record · LEDGER

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding

As of 23 August 2026, this Paper Citation Record lists 59 of 59 outbound references and 11 inbound Pith citation observations for arXiv:2601.20430.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2601.20430 v2

Coverage vector

measured 59 of 59 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T15:44:17.303248Z

measured 70 of 70 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 11 of 11 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T21:04:30.042864Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T16:49:57.063058Z

Reference resolution

59 of 59 outbound references displayed

  • verified exact0
  • verified fuzzy10
  • unresolved49
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c3c7fcfe-c610-41ab-8f3a-4820dd6cf184 · outbound

This paper cites Document Parsing Unveiled: Techniques, Challenges, and Prospects for Structured Information Extraction.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding Document Parsing Unveiled: Techniques, Challenges, and Prospects for Structured Information Extraction

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:16.966655Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:16.966655Z digest=sha256:5d52f70bcfeba57a619997527c44b497690abbab92a8513f905720dd1e81c064

Observation 28aecddb-10c8-4960-a858-eb33f6df49bb · outbound

This paper cites Dolphin: Document Image Parsing via Heterogeneous Anchor Prompting.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding Dolphin: Document Image Parsing via Heterogeneous Anchor Prompting

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:16.973324Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:16.973324Z digest=sha256:35caa71ee4469bad823bf9e16db009058fec6c5839251c5c0a179ac1a4eb0cac

Observation 2b1f6253-d19b-407b-aaa4-221515c6d48a · outbound

This paper cites Monkeyocr: Document parsing with a structure-recognition-relation triplet paradigm.arXiv preprint arXiv:2506.05218, 2025.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding Monkeyocr: Document parsing with a structure-recognition-relation triplet paradigm.arXiv preprint arXiv:2506.05218, 2025

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:16.978948Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:16.978948Z digest=sha256:2396d26d9386ede7431666d4baa1ea3518aca45d5d73bb57eb067d89dcb75645

Observation 38bc556d-76f8-4df6-bfcf-fa4f878dfe24 · outbound

This paper cites MinerU: An Open-Source Solution for Precise Document Content Extraction.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding MinerU: An Open-Source Solution for Precise Document Content Extraction

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:16.984659Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:16.984659Z digest=sha256:c95e5eaa0dd7bfe73cc85ce30d25381834367c31cfa4dfb5860680e22d4545da

Observation fbd59e9a-c7e3-48f3-87a6-bd58f72aaca8 · outbound

This paper cites olmocr: Unlocking trillions of tokens in pdfs with vision language models.arXiv preprint arXiv:2502.18443, 2025.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding olmocr: Unlocking trillions of tokens in pdfs with vision language models.arXiv preprint arXiv:2502.18443, 2025

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:16.990144Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:16.990144Z digest=sha256:2d014b8989a696fed01eac748371f4b2ba8dad375d60112ebef5f8b13b75d6c5

Observation 4fb521d5-2558-4cef-b695-bda64cd4e81d · outbound

This paper cites an unresolved cited work.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding Unresolved cited work

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:16.995462Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:16.995462Z digest=sha256:9c8825812d44e9a520b9c755ac7f9227a476f78e54b968488c0cc09ed6c9f44f

Observation 65178962-49ea-4845-b8ee-059fc51e11d0 · outbound

This paper cites MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:17.001936Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:17.001936Z digest=sha256:4014a3e636f848f6489606630cce1562c077c06c5db6133001129b2db7b65395

Observation 81b97e13-f7cb-4139-a7dd-3ea4ed262be7 · outbound

This paper cites Paddleocr-vl: Boosting multilingual document parsing via a 0.9 b ultra-compact vision-language model.arXiv preprint arXiv:2510.14528, 2025.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding Paddleocr-vl: Boosting multilingual document parsing via a 0.9 b ultra-compact vision-language model.arXiv preprint arXiv:2510.14528, 2025

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:17.008593Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:17.008593Z digest=sha256:8cc93f649c546964de09b30badbf39a56856eb37227083a30ad3a3dd6e5926f7

Observation 51db00ed-ed1d-40a9-8942-9d9547d83dc7 · outbound

This paper cites Language models are few-shot learners.Advances in neural information processing systems, 33:1877–1901, 2020.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding Language models are few-shot learners.Advances in neural information processing systems, 33:1877–1901, 2020

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:17.013892Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:17.013892Z digest=sha256:8c82b8041497a0a244b3e62662c862cbb3eceb021249da5661fa43199cdc4761

Observation 6407f825-1915-4392-a1de-1b8e9630142b · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding LLaMA: Open and Efficient Foundation Language Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:17.018994Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:17.018994Z digest=sha256:698b0741c56fc334ff9b450b0feed088926a205dc3ed05111a3e926b90234fd6

Observation f82d6265-c768-4184-a7f5-7700df801663 · outbound

This paper cites GPT-4 Technical Report.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding GPT-4 Technical Report

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:17.025501Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:17.025501Z digest=sha256:d27a49389e56a126a4552d776de56e4f4d9ab7a3c5bcc5e8fcbd70a2c5ac0987

Observation c5043f4b-ba09-4296-bdae-802a4c47b9cc · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding Gemini: A Family of Highly Capable Multimodal Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:17.031620Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:17.031620Z digest=sha256:9e90ad42f276c091b166ab722d02fc94ebbb949851904693b6a41ca838e61cfa

Observation 800a5e75-8673-4c26-9d98-870939a307e0 · outbound

This paper cites Mistral 7B.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding Mistral 7B

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:17.038573Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:17.038573Z digest=sha256:64ad21664cb293a8966cb3065e55e05a4870608f980cce1e3e6a0ca48a7f79a4

Observation e0258661-845a-4506-8560-49e6f5284cef · outbound

This paper cites Learning transferable visual models from natural language supervision.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding Learning transferable visual models from natural language supervision

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:17.044069Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:17.044069Z digest=sha256:44fd378a8ba47debebcb3cf847125679975eb06a103673b2925c416484e9b59e

Observation ca900daa-7c5c-4724-b9db-1b983b302a36 · outbound

This paper cites Visual instruction tuning.Advances in neural information processing systems, 36:34892–34916, 2023.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding Visual instruction tuning.Advances in neural information processing systems, 36:34892–34916, 2023

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:17.049410Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:17.049410Z digest=sha256:347e60585abf78d868811c8582d18c5a033e9af2a896d9ac5967ac1b7a84fabb

Observation 1c583a0e-1b4d-4f34-9e07-a2e25a7241fd · outbound

This paper cites MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:17.056152Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:17.056152Z digest=sha256:c3fda3860b2ce870625d466c5a3356a03b6761dd5d4fa11187bbd679c2a41ec7

Observation 14265ea9-2609-4393-8036-c816c782a320 · outbound

This paper cites Instructblip: Towards general-purpose vision-language models with instruction tuning.Advances in neural information processing systems, 36:49250–49267, 2023.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding Instructblip: Towards general-purpose vision-language models with instruction tuning.Advances in neural information processing systems, 36:49250–49267, 2023

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:17.062657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:17.062657Z digest=sha256:e4fe5c9bf7dd0229477e35142f7c22c7b3c01047002766350ac66858ab44faab

Observation 275bff2b-0528-4d99-a769-5016c2a45a4c · outbound

This paper cites SigLIP 2: Multilingual Vision-Language Encoders with Improved Semantic Understanding, Localization, and Dense Features.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding SigLIP 2: Multilingual Vision-Language Encoders with Improved Semantic Understanding, Localization, and Dense Features

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:17.068839Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:17.068839Z digest=sha256:fdad0ca9d3960b11cc7f43e6ec46f188e28f8977ca9e8558f802eaedcc2a246e

Observation 8fdd8404-4467-4ca7-9efb-94bc36cf75f3 · outbound

This paper cites PaddleOCR 3.0 Technical Report.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding PaddleOCR 3.0 Technical Report

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:17.075783Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:17.075783Z digest=sha256:da6e4eaf754b7ff0285a5b41351c2fd51b4b4be0d54e720539463173d2349b85

Observation abf9538b-1101-4abd-96e1-a9471f9bf8e9 · outbound

This paper cites Docling: An Efficient Open-Source Toolkit for AI-driven Document Conversion.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding Docling: An Efficient Open-Source Toolkit for AI-driven Document Conversion

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:17.083690Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:17.083690Z digest=sha256:9c3af69477893647bd6ea2254203febfc3dbeec054a2c00dfb1f0520552a94cd

Observation 021ed39c-3c1f-4c42-8cc1-66a02c9bd9d1 · outbound

This paper cites DeepSeek-OCR: Contexts Optical Compression.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding DeepSeek-OCR: Contexts Optical Compression

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:17.088989Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:17.088989Z digest=sha256:fef16d7a2de703a3fa4df9b1d8b239a6fe6a4819459a7d4d0a42386f971b1bb6

Observation 265ca862-3f02-4b68-841b-abaf3cf10b11 · outbound

This paper cites Learning to Recover from Multi-Modality Errors for Non-Autoregressive Neural Machine Translation.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding Learning to Recover from Multi-Modality Errors for Non-Autoregressive Neural Machine Translation

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:17.094994Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:17.094994Z digest=sha256:6c18f47cf7a2949d84e8413dd7b44d33fc3d41d84cbb983bbd512922b86990e5

Observation 00b3ea6c-2fb6-4c2a-b774-495dd3199179 · outbound

This paper cites Maskgit: Masked generative image transformer.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding Maskgit: Masked generative image transformer

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:44:18.902835Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T15:44:17.099950Z digest=sha256:8a4ae282bdb4572d43aefb0167a1b49f74d81e4e435c5b34f513575fc4b76072

Observation 350c4962-09c2-4acf-8802-8d1c8442f4bd · outbound

This paper cites Medusa: Simple LLM Inference Acceleration Framework with Multiple Decoding Heads.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding Medusa: Simple LLM Inference Acceleration Framework with Multiple Decoding Heads

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:17.104779Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:17.104779Z digest=sha256:da8812b549a5b99947d4a8b05bfd37b861da96d4048022ece4949b2bb4819078

Observation 9057952b-921e-4f8e-ac46-f468fe359c36 · outbound

This paper cites Youtu-vl: Unleashing visual potential via unified vision-language supervision.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding Youtu-vl: Unleashing visual potential via unified vision-language supervision

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:17.109878Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:17.109878Z digest=sha256:15220415263c1c938862d42b1caddc507b3d6130668d436b940ed36268cf916e

Observation 5c4cca28-a230-4bb3-a035-9d3ad42778c6 · outbound

This paper cites Youtu-llm: Unlocking the native agentic potential for lightweight large language models.arXiv preprint arXiv:2512.24618, 2025.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding Youtu-llm: Unlocking the native agentic potential for lightweight large language models.arXiv preprint arXiv:2512.24618, 2025

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:17.114650Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:17.114650Z digest=sha256:9b05088c61eef7f0e67d8bbadcd72846c132640bed913fbe85a1e52011c23209

Observation 0ea370fd-202a-4d56-b225-e05a196ee111 · outbound

This paper cites Optimized table tokenization for table structure recognition.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding Optimized table tokenization for table structure recognition

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:17.119434Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:17.119434Z digest=sha256:a690500566b8c5665f854972175a08b9f129ecce492cef49b5c5d97d71b654f1

Observation 7db0fe3b-5fcf-4d01-aa20-ee8c4d0f55e0 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:17.124389Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:17.124389Z digest=sha256:9779c4749a0af4b4f33d5147b343a39c7fd20f821b223628619cc8924410e513

Observation a62de1df-af5b-440c-b49a-a0d43f451c13 · outbound

This paper cites Omnidocbench: Benchmarking diverse pdf document parsing with comprehensive annotations.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding Omnidocbench: Benchmarking diverse pdf document parsing with comprehensive annotations

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:17.129841Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:17.129841Z digest=sha256:aba2815ed7e68b3034e01ed9b3b06f84b2dc86f920780f0e4cb93c54b0654958

Observation 398feac7-ff3d-45e3-8f67-2da3619ec8de · outbound

This paper cites Marker.https://github.com/datalab-to/marker, 2025.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding Marker.https://github.com/datalab-to/marker, 2025

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:44:18.846582Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T15:44:17.134934Z digest=sha256:bbb28a623a45d2b0de63b13c0758f737df63fe8cee448813eb8346d0091d4144

Observation cfb6dc25-7455-4cb2-b3a5-8a114ff93372 · outbound

This paper cites GPT-4o System Card.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding GPT-4o System Card

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:17.140395Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:17.140395Z digest=sha256:bab4e1e70f60d573bdedc961d5abc6038cc6e95c2f358dfa7c32da4e445e9dc9

Observation 913246a2-824c-438a-867a-78f96aa42f9b · outbound

This paper cites InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:17.145667Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:17.145667Z digest=sha256:d8cd90fea5cbb2bd6582a8e3bfa09fb59a7d74139fe0f610ea956c239705152f

Observation 5c7a5446-4ce2-4536-99ff-060339556d2b · outbound

This paper cites Qwen2.5-VL Technical Report.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding Qwen2.5-VL Technical Report

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:17.156679Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:17.156679Z digest=sha256:2eb4354587f812e8f4eeff77525cf9553a5e6bfe0c5481ab0d2dbfbd449f5d6b

Observation 68e6045e-c5eb-48e3-8263-934d9cfab531 · outbound

This paper cites Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:17.162018Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:17.162018Z digest=sha256:15d718beb148cb524d0de8a5b8d1e4bdcb7927b0da20c96abfefe9ba5bf6ddb0

Observation f2c7bcfa-27b1-4a28-9681-d38efc63ea09 · outbound

This paper cites Ocrflux.https://github.com/chatdoc- com/OCRFlux, 2025.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding Ocrflux.https://github.com/chatdoc- com/OCRFlux, 2025

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:44:18.828056Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T15:44:17.167479Z digest=sha256:e8066cb5873afe5522f5d6feaccbed24e57a2ecca656cfb49a2a99b367545661

Observation 95e24531-f5e0-49de-b1ab-ac01fc966a01 · outbound

This paper cites Mistral-ocr.https://mistral.ai/news/mistral-ocr?utm source=ai-bot.cn, 2025.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding Mistral-ocr.https://mistral.ai/news/mistral-ocr?utm source=ai-bot.cn, 2025

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:44:18.806663Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T15:44:17.175157Z digest=sha256:d064574e2c10115052b37db1dbb474b2433f395fb71dff3594f8d70f105bd8fc

Observation b413d709-e95d-4c19-8996-3c5e53901baa · outbound

This paper cites Points-reader: Distillation-free adaptation of vision-language models for document conversion.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding Points-reader: Distillation-free adaptation of vision-language models for document conversion

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:44:18.788674Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T15:44:17.182344Z digest=sha256:aed55eaa97dfcb0d8386c65ac3ea6b78238dfd4035ec93537dce5026dda8f7f7

Observation 8f6593be-c8f6-4088-a3b5-7443e6954bbb · outbound

This paper cites Nanonets-ocr-s: A model for transforming documents into structured markdown with intelligent content recognition and semantic tagging, 2025.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding Nanonets-ocr-s: A model for transforming documents into structured markdown with intelligent content recognition and semantic tagging, 2025

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:17.187713Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:17.187713Z digest=sha256:62305c93375643439ab0e869029ab93e7073795f81f9baedd2f7629fa146d9ab

Observation 584ac40b-b000-43d9-9c34-e88a801f4b7b · outbound

This paper cites Image over text: Transforming formula recognition evaluation with character detection matching.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding Image over text: Transforming formula recognition evaluation with character detection matching

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:17.192599Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:17.192599Z digest=sha256:2f54e39cfcb79789254f03857106ce8ebc9954eb3f3d1a2833d612f477574e53

Observation 8374201a-4916-4d89-8531-ceb6b2fd6130 · outbound

This paper cites General OCR Theory: Towards OCR-2.0 via a Unified End-to-end Model.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding General OCR Theory: Towards OCR-2.0 via a Unified End-to-end Model

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:17.198748Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:17.198748Z digest=sha256:40bc318d317578c0cc5ff05c59da4f08b47cda0e438411024565674ceab0b5bf

Observation bfa9869d-5c7f-479b-ab76-594a6deb129d · outbound

This paper cites Google deepmind.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding Google deepmind

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:44:18.743820Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T15:44:17.204307Z digest=sha256:0794fc13294b400a9150cc4469768d91da3a8d297bedc6f0db1790cc080e5311

Observation f6f2c98a-57ea-4267-bbb0-7efbdebe247d · outbound

This paper cites Cc-ocr: A comprehensive and challenging ocr benchmark for evaluating large multimodal models in literacy.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding Cc-ocr: A comprehensive and challenging ocr benchmark for evaluating large multimodal models in literacy

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:17.209052Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:17.209052Z digest=sha256:31f74e6d2d86958365423ba048d79603c19d64289ad30402bdf9826c5dde4966

Observation 4099a557-3283-4f1b-a3ea-7496250d29b4 · outbound

This paper cites OCRBench v2: An Improved Benchmark for Evaluating Large Multimodal Models on Visual Text Localization and Reasoning.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding OCRBench v2: An Improved Benchmark for Evaluating Large Multimodal Models on Visual Text Localization and Reasoning

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:17.214122Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:17.214122Z digest=sha256:87e722a8867200faff15ced4d9e965a52845ceea51a6ea9f7d491158f2478109

Observation a96d7866-464b-4bde-99bc-dc0f3ea305eb · outbound

This paper cites Qwen3 Technical Report.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding Qwen3 Technical Report

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:17.219209Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:17.219209Z digest=sha256:5e1b07a3ea0052046a48aa640bdfd25c68e8d912614a6cbb6fb953c71b8f0d5f

Observation f1107f11-8a7f-4d10-bf80-4c163b796c7f · outbound

This paper cites Onechart: Purify the chart structural extraction via one auxiliary token.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding Onechart: Purify the chart structural extraction via one auxiliary token

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:44:18.712573Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T15:44:17.225007Z digest=sha256:c2c3c8f42486faecb5e4a49bbcf93033d7b54d93e443543b0d15f4c6318ba3ea

Observation 30f9e1a3-1ad4-458e-a7e2-ab5c0f8f3d14 · outbound

This paper cites Deplot: One-shot visual language reasoning by plot-to-table translation.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding Deplot: One-shot visual language reasoning by plot-to-table translation

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:17.230016Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:17.230016Z digest=sha256:0159df20e4f6a3ad0ea6bf9cc07060e014dcbaa7320251fb74ec71958bab03e2

Observation 8eea3c0f-5258-4300-ad5b-07682ca6ffc8 · outbound

This paper cites DocLayout-YOLO: Enhancing Document Layout Analysis through Diverse Synthetic Data and Global-to-Local Adaptive Perception.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding DocLayout-YOLO: Enhancing Document Layout Analysis through Diverse Synthetic Data and Global-to-Local Adaptive Perception

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:17.234986Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:17.234986Z digest=sha256:5a227150cf235dbd5b3cf6fd9cfa28f55caa103cf4e5922b9d2ebf22f8e02228

Observation a14a9b06-33ef-441e-83ad-40e11c6b45b9 · outbound

This paper cites PP-DocLayout: A Unified Document Layout Detection Model to Accelerate Large-Scale Data Construction.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding PP-DocLayout: A Unified Document Layout Detection Model to Accelerate Large-Scale Data Construction

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:17.241605Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:17.241605Z digest=sha256:9fd91a744733d06c8276d5f10558bc63cdbf82d3b4c1a60877a52f61446c4ab9

Observation 14057ee5-a041-4c7e-9f4f-bd9396fddfaa · outbound

This paper cites An end-to-end formula recognition method integrated attention mechanism.Mathematics, 11(1):177, 2022.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding An end-to-end formula recognition method integrated attention mechanism.Mathematics, 11(1):177, 2022

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:44:18.677116Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T15:44:17.247926Z digest=sha256:d33f4fe0c6160346b427c7e4cd5d1b065d71beedb0fee84be94ed2ce69a83819

Observation 212bb3d0-dab3-411c-8a86-64b1792ab6bb · outbound

This paper cites Deeptabstr: Deep learning based table structure recognition.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding Deeptabstr: Deep learning based table structure recognition

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:44:18.655210Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T15:44:17.252678Z digest=sha256:780d43f015d7590e8451c7de9568c99209fdb0df3832d7c1e8707724c41f6f14

Observation bd2b6cff-7e5e-4652-98ed-4fea655bfc79 · outbound

This paper cites Tsrformer: Table structure recognition with transformers.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding Tsrformer: Table structure recognition with transformers

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:44:18.636070Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T15:44:17.257609Z digest=sha256:0f4b37780ef06ea96fdb15d36f7e7f16cd8bed8967de018ea7f21cf817ff4d8b

Observation b595229f-2acc-4f99-9f13-5109d8aec9c5 · outbound

This paper cites LayoutReader: Pre-training of Text and Layout for Reading Order Detection.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding LayoutReader: Pre-training of Text and Layout for Reading Order Detection

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:17.262939Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:17.262939Z digest=sha256:2ed7304dfa7afc89003f669a472c0cfc24467ea26c9ad98b4aaf4739a0ff41d6

Observation 4fcce5de-9d68-4abf-8115-01f9643393c4 · outbound

This paper cites Internvl: Scaling up vision foundation models and aligning for generic visual-linguistic tasks.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding Internvl: Scaling up vision foundation models and aligning for generic visual-linguistic tasks

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:17.268807Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:17.268807Z digest=sha256:ccd9b20f5d865e41ec3eaa18891d6c9a49415d785b3c963a569e67233d1b8251

Observation 2e234018-d416-4158-b7d6-81971bb6d02a · outbound

This paper cites How Far Are We to GPT-4V? Closing the Gap to Commercial Multimodal Models with Open-Source Suites.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding How Far Are We to GPT-4V? Closing the Gap to Commercial Multimodal Models with Open-Source Suites

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:17.273912Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:17.273912Z digest=sha256:02cb4f4484f02139a5b658628239d08bd29e07ab737e6fd2d7f46a59084fbf92

Observation f5e87724-d7ac-4577-a0f1-a91f5910cb38 · outbound

This paper cites Enhancing the Reasoning Ability of Multimodal Large Language Models via Mixed Preference Optimization.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding Enhancing the Reasoning Ability of Multimodal Large Language Models via Mixed Preference Optimization

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:17.281082Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:17.281082Z digest=sha256:9316dc11a66f2a1c8e39af4433be84d900dd68b6ece590750a4a02aa7e536e18

Observation 750da50b-33ec-435e-af67-dd38fc66f501 · outbound

This paper cites Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:17.286792Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:17.286792Z digest=sha256:e6d4d84c2b340fd39469a9cd0b8847d7c5ed2119aba4dae83b147552d24ce198

Observation 00200826-4b3e-483b-b981-c59b066324db · outbound

This paper cites InternVL3.5: Advancing Open-Source Multimodal Models in Versatility, Reasoning, and Efficiency.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding InternVL3.5: Advancing Open-Source Multimodal Models in Versatility, Reasoning, and Efficiency

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:17.292277Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:17.292277Z digest=sha256:2a06bba16f6826cae32b83ac785025b29bac42dc6b188b55a7b56724d6ccf647

Observation 083a24fa-adc3-4c80-b1a8-0e3dde525d6b · outbound

This paper cites Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:17.297182Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:17.297182Z digest=sha256:04a658e180abda1519933048e9f4c6f283518ca7da0784dea058aefa846211c2

Observation 08c702be-99a7-495c-a07a-1c7fdcfdcdce · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:17.303248Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:17.303248Z digest=sha256:8bcf6021c0a35ea56065ac83c422430e0d6730a36e543a26708a966b893cc081

Pith citing papers

Observation 61e47448-70da-49d5-b4cd-2472b57d0468 · inbound

Parser-Oriented Structural Refinement for a Stable Layout Interface in Document Parsing cites this paper.

Parser-Oriented Structural Refinement for a Stable Layout Interface in Document Parsing Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-07-14T03:21:27.323268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-13T21:07:44.073353Z digest=sha256:bac8d442c6c10f303a455daffb2e3a7b1b8cc547915dc6a3e90ee459a5432fe9

Observation 4d1387c1-c444-482a-916d-cd38102058a3 · inbound

MinerU2.5-Pro: Pushing the Limits of Data-Centric Document Parsing at Scale cites this paper.

MinerU2.5-Pro: Pushing the Limits of Data-Centric Document Parsing at Scale Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-07-14T03:21:27.323268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-10T18:58:41.377996Z digest=sha256:bd0e6647ba558133c14f72898614f67d8dc76f551044756c926992b7fc5b7d77

Observation 14021d59-1781-42f0-a013-106d9dd5b04b · inbound

How Far Is Document Parsing from Solved? PureDocBench: A Source-TraceableBenchmark across Clean, Degraded, and Real-World Settings cites this paper.

How Far Is Document Parsing from Solved? PureDocBench: A Source-TraceableBenchmark across Clean, Degraded, and Real-World Settings Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-07-14T03:21:27.323268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-11T02:28:02.152600Z digest=sha256:059115331e5c3d60f55809790413dbea32e697775588130e9c155c61eba8aeb3

Observation 612e5ec8-f733-4f38-9cd7-72c37d730f45 · inbound

MPDocBench-Parse: Benchmarking Practical Multi-page Document Parsing cites this paper.

MPDocBench-Parse: Benchmarking Practical Multi-page Document Parsing Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding

Reference 59

Resolution
verified exact
arxiv_id, observed 2026-07-14T03:21:27.323268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-22T05:58:04.055855Z digest=sha256:87b5361586693421e9c0050d45c663fadd117876d4a503157a99f5ccd3eb4f5c

Observation 37772d44-a626-4a8f-a5ea-4be6664445b9 · inbound

MPDocBench-Parse: Benchmarking Practical Multi-page Document Parsing cites this paper.

MPDocBench-Parse: Benchmarking Practical Multi-page Document Parsing Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding

Reference 59

Resolution
verified exact
arxiv_id, observed 2026-07-14T03:21:27.323268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-30T17:37:33.750306Z digest=sha256:68abfafd1a9d1d07245d807572c6fe4d76932b851761504fc60a768aba3c77b7

Observation 762ba601-85d9-4e7e-b9f2-a7bf8e3e97f4 · inbound

ABot-OCR Technical Report cites this paper.

ABot-OCR Technical Report Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-07-14T03:21:27.323268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-29T13:29:17.221676Z digest=sha256:c79d5bf669d7cf7caec409e5259f7adebfc74928188331f411990aeb5703d60b

Observation 4de339ab-37c2-4da1-a298-ef31081c387f · inbound

PaddleOCR-VL-1.6: Expanding the Frontier of Document Parsing with Under-Optimized Region Refinement and Progressive Post-Training cites this paper.

PaddleOCR-VL-1.6: Expanding the Frontier of Document Parsing with Under-Optimized Region Refinement and Progressive Post-Training Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-07-14T03:21:27.323268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-28T10:39:14.444486Z digest=sha256:3eb6dc46dd7d451fe97ead3101a98dc4c68ab4b94c36b690a08d727f73e5cf43

Observation deaf8128-b58e-4a3a-8ac9-94642f1e2005 · inbound

P-MTP: Efficient Document Parsing via Multi-Token Prediction with Progressive Depth Scaling cites this paper.

P-MTP: Efficient Document Parsing via Multi-Token Prediction with Progressive Depth Scaling Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-07-14T03:21:27.323268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-26T00:18:07.278889Z digest=sha256:ad11aa449f289a87d1526e6a2ddd06293f641abcd9f21c5e668edd15f88b0ee0

Observation 647786e6-5233-4b77-b337-d1e8c995a3c0 · inbound

OvisOCR2 Technical Report cites this paper.

OvisOCR2 Technical Report Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-02T04:40:04.566211Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T04:40:04.566211Z digest=sha256:753820a5767fe1e7ad04380a19f5448092aaaeb8f6e5c56444a9c96945389dbe

Observation b96afe7f-78f9-4bf8-98cb-7c7211102b1c · inbound

HPD-Parsing: Hierarchical Parallel Document Parsing cites this paper.

HPD-Parsing: Hierarchical Parallel Document Parsing Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-01T14:16:08.107522Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:16:08.107522Z digest=sha256:23110efe30a1844b4abe09b7207c86cd7747479005f3b6aec515bbfa7dfc7a16

Observation df24d726-cb6d-4f2f-a3ab-c6b587dd6f62 · inbound

NaviDC-OCR: Navigating Document Parsing Across Digital and Camera-Captured Documents cites this paper.

NaviDC-OCR: Navigating Document Parsing Across Digital and Camera-Captured Documents Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-15T21:04:30.042864Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T21:04:30.042864Z digest=sha256:9128414161d58906bbbbf2e72fda3b63a332ae0589296b3626244837848c4cc9