Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-17T13:25:31.884175Z
Paper Citation Record · LEDGER
As of 5 August 2026, this Paper Citation Record lists 63 of 63 outbound references and 36 inbound Pith citation observations for arXiv:2509.22186.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-17T13:25:31.884175Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-05T00:46:40.906153Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-07-10T17:07:25.732459Z
63 of 63 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 0eb17dcd-cb96-42cf-bbf4-6cc66047925f · outbound
MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing GPT-4 Technical Report
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 573fdc67-6c36-4bd9-8f60-ebeefafb04cb · outbound
MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing Wukong-Reader: Multi-modal Pre-training for Fine-grained Visual Document Understanding
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation ba8d191b-7eb0-4511-8ae0-dc7bc2157634 · outbound
MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing Qwen2.5-VL Technical Report
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation b6b4b551-4d46-423c-9d97-1092ce276f30 · outbound
MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing Nougat: Neural Optical Understanding for Academic Documents
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 093bba68-43e4-4f8e-96c6-192b8f2bf9cc · outbound
MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing Ocrflux.https://github.com/chatdoc-com/OCRFlux
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 04421ea2-b829-40fd-b508-4093dac1d46d · outbound
MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing Ocean-OCR: Towards General OCR Application via a Vision-Language Model
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 61d5631e-b4a9-418b-8c94-bb801dfea1f4 · outbound
MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation cb70f3ae-da33-464a-b1c6-7fe38e5f921b · outbound
MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing PaddleOCR 3.0 Technical Report
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 03980eac-ab61-4aaa-ac6c-c7734cf21115 · outbound
MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing Vision grid transformer for document layout analysis
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation d12d51b4-6981-43e9-8205-ede07f0f853a · outbound
MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing Patch n’pack: Navit, a vision transformer for any aspect ratio and resolution.Advances in Neural Information Processing Systems, 36: 2252–2274
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 4f1bfba0-f5f8-4ac1-a8d4-0c2ccefa47ee · outbound
MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing Dolphin: Document Image Parsing via Heterogeneous Anchor Prompting
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 52d0a6e8-8ef2-4742-ab23-e6dc1ea29c06 · outbound
MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing OCRBench v2: An Improved Benchmark for Evaluating Large Multimodal Models on Visual Text Localization and Reasoning
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 4e95f10a-35a2-499a-9591-3c343fe030d8 · outbound
MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing Seed1.5-VL Technical Report
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 9a30f1f2-85b5-438b-b26f-f390800d78d4 · outbound
MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing Layoutlmv3: Pre-training for document ai with unified text and image masking
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation f0fc35e4-6a7a-4ce7-9598-8220d09b31e4 · outbound
MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing Ocr-free document understanding transformer
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 4ef1a096-d4b0-416e-9151-791e8128886c · outbound
MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing Gon- zalez, Hao Zhang, and Ion Stoica
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation bc1d039a-6fc1-438f-8669-2cf20ed0aabd · outbound
MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing arXiv preprint arXiv:2506.05218 , year=
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation a57bb7ce-7adb-4547-b85c-afe3fb10108d · outbound
MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing Doctr: Document transformer for structured information extraction in documents
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation ec8a855e-69e8-422c-933d-891fe3663a54 · outbound
MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing Revolutionizing Retrieval-Augmented Generation with Enhanced PDF Structure Recognition
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 4fe99761-6a07-45a8-8934-3a22c03b774e · outbound
MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing Hrvda: High-resolution visual document assistant
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 08762d7b-06fd-4937-88ce-1226f17b3ba2 · outbound
MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing PP-FormulaNet: Bridging Accuracy and Efficiency in Advanced Formula Recognition
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 58d19e35-81c9-4ed9-8fd8-5295c7f7a36d · outbound
MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing POINTS-Reader: Distillation-Free Adaptation of Vision-Language Models for Document Conversion
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation b1e09e38-907f-449a-b072-d1c13ad8f18f · outbound
MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing TextMonkey: An OCR-Free Large Multimodal Model for Understanding Document
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 5792a4b9-125a-4cde-b6e3-c5956b76c387 · outbound
MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing Docling: An Efficient Open-Source Toolkit for AI-driven Document Conversion
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 09fd7087-0dfa-44e7-961e-d63ffb8bf3c4 · outbound
MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing Optimized table tokenization for table structure recognition
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 41e4b1c1-372c-463b-9a14-04b5ad310093 · outbound
MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing Nanonets-ocr-s
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation b0937d3c-f884-40f6-949d-b11cc6de362d · outbound
MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing Mathpix.https://mathpix.com/
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 2243c99a-0e45-49f1-8fff-ddf6d83300a6 · outbound
MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing SmolDocling: An ultra-compact vision-language model for end-to-end multi-modal document conversion
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation e7ab0eda-b5a6-4a65-9a70-0eda9239920e · outbound
MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing Native Visual Understanding: Resolving Resolution Dilemmas in Vision-Language Models
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation d91cfd00-34e8-4325-932e-e93d5f482180 · outbound
MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing Pdf-extract-kit
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation daea5a8c-2d61-4ee5-a13f-614a0f586a9b · outbound
MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing Omnidocbench: Benchmarking diverse pdf document parsing with comprehensive annotations
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation e8807e89-2831-4c1e-b2e3-022393b65b0c · outbound
MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing Marker.https://github.com/datalab-to/marker
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 502ee242-ec00-4bfe-9ecf-fa1f03ab2649 · outbound
MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing Surya: A lightweight document ocr and analysis toolkit
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 99dc6d66-502b-470d-823d-2deba20c07b5 · outbound
MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing Doclaynet: A large human-annotated dataset for document-layout segmentation
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 42e4aeed-1dbf-42ab-93de-92416fc341fd · outbound
MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing olmocr: Unlocking trillions of tokens in pdfs with vi- sion language models.arXiv preprint arXiv:2502.18443, 2025a
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 81bd3dd6-80a7-4f3e-82db-f6e7427823d2 · outbound
MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing Rapid table.https://github.com/RapidAI/RapidTable
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 17d7cbcd-681d-4272-8def-c566e71e047f · outbound
MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing dots.ocr: Multilingual document layout parsing in a single vision-language model
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation ae2b91bf-a4b2-4910-ad9e-61ba3ad00edc · outbound
MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing Real-time single image and video super-resolution using an efficient sub-pixel convolu- tional neural network
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation a093c681-9128-4104-b413-19487e3d56fa · outbound
MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing Roformer: Enhanced transformer with rotary position embedding.Neurocomputing, 568:127063
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation f70ad65b-3350-45c5-b78d-415902720509 · outbound
MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing Unifying vision, text, and layout for universal document processing
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 17af5ca7-490f-4b47-a1d1-61f233288be0 · outbound
MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing Mistral-ocr
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 93f37dde-91f7-4890-874a-25a952bcf41d · outbound
MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing Qwen2 Technical Report
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation d51660c9-a1b7-4d7d-8f49-009a4f899987 · outbound
MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing Omniparser: A unified framework for text spotting key information extraction and table recognition
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation e3691aa4-84fc-4532-ae02-70616618af65 · outbound
MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing Yolov10: Real-time end-to-end object detection.Advances in Neural Information Processing Systems, 37:107984–108011
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 1004e3a8-1653-4f9e-a5dc-a5db11d1fcb3 · outbound
MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing UniMERNet: A Universal Network for Real-World Mathematical Expression Recognition
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation c7931339-2df4-41a4-abbe-cc18a883d718 · outbound
MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing MinerU: An Open-Source Solution for Precise Document Content Extraction
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 8e7153a0-26d5-4d20-a281-ed8fa869d40a · outbound
MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing Image over text: Transforming formula recognition evaluation with character detection matching
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 05f17076-2410-4fd1-9208-3759224807ea · outbound
MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 994d66c7-f5d7-4c88-8908-910037e8a3b0 · outbound
MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing InternVL3.5: Advancing Open-Source Multimodal Models in Versatility, Reasoning, and Efficiency
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 95ac9905-f8ce-43f7-aba8-c0db34cf4bc0 · outbound
MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing LayoutReader: Pre-training of Text and Layout for Reading Order Detection
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 30b891ca-a3eb-4600-8eee-a1c01e80bee7 · outbound
MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing Vrdu: A benchmark for visually-rich document understanding
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 765ad383-e252-4db4-9a90-bec6d7788331 · outbound
MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing General OCR Theory: Towards OCR-2.0 via a Unified End-to-end Model
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 5740e854-0131-47fa-9469-3996b278d76e · outbound
MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing CC-OCR: A Comprehensive and Challenging OCR Benchmark for Evaluating Large Multimodal Models in Literacy
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 20685813-7daa-4a34-9c0f-8b578341645e · outbound
MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing MiniCPM-V: A GPT-4V Level MLLM on Your Phone
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 52a366ce-7778-499b-b0a9-3e1c069b89e5 · outbound
MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing MiniCPM-V 4.5: Cooking Efficient MLLMs via Architecture, Data, and Training Recipe
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 8c64bdec-bfed-4055-8992-c5e8dc0bd9fa · outbound
MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing OCR Hinders RAG: Evaluating the Cascading Impact of OCR on Retrieval-Augmented Generation
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 07e0b3f9-3385-4af6-90fe-ea0beeff32d0 · outbound
MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing Document Parsing Unveiled: Techniques, Challenges, and Prospects for Structured Information Extraction
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 3b32a8be-a118-49cc-96d6-abb1e6f9bed1 · outbound
MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing Retrieval-Augmented Generation for AI-Generated Content: A Survey
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 942e003c-d139-401d-b390-bb6c219ad5df · outbound
MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing DocLayout-YOLO: Enhancing Document Layout Analysis through Diverse Synthetic Data and Global-to-Local Adaptive Perception
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation c727ceae-3690-4abf-b417-d1f1dd3bfd21 · outbound
MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing Sglang: Efficient execution of structured language model programs.Advances in neural information processing systems, 37:62557–62583
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 15bbe352-60a6-4854-abcf-3d9247b7b49e · outbound
MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing Global table extractor (gte): A framework for joint table identification and cell structure recognition using visual context
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 7bf75f20-0a1b-4ae1-89c0-a39023d19af6 · outbound
MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing Image-based table recognition: data, model, and evaluation
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 89880a42-32d6-42e5-ae30-76fd55f6602a · outbound
MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 5307e98b-c8c7-4a47-bb0d-32f89cf36d49 · inbound
FinCriticalED: A Visual Benchmark for Financial Fact-Level OCR MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation fa00f7a0-fc2c-4669-8566-93e8311d8c49 · inbound
Low-Resolution Editing is All You Need for High-Resolution Editing MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7b66e0d8-457f-4a8e-9507-8d98c1bbc67c · inbound
UniRec-0.1B: Unified Text and Formula Recognition with 0.1B Parameters MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3dcf2ae4-de54-485b-8b49-3ff70fbcc97c · inbound
ChartVerse: Scaling Chart Reasoning via Reliable Programmatic Synthesis from Scratch MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 752379fa-e195-4dd4-aaa4-32406bf79e66 · inbound
PaddleOCR-VL-1.5: Towards a Multi-Task 0.9B VLM for Robust In-the-Wild Document Parsing MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation c18ac8ea-f245-4968-b645-40226da52518 · inbound
HSD: Training-Free Acceleration for Document Parsing Vision-Language Models with Hierarchical Speculative Decoding MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7542957f-7bc6-487a-9840-64a2f94a9087 · inbound
DECKBench: Benchmarking Multi-Agent Frameworks for Academic Slide Generation and Editing MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9fb10128-ac5d-4d04-b332-6cae22609192 · inbound
HVR-Met: A Hypothesis-Verification-Replanning Agentic System for Extreme Weather Diagnosis MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 65aa4065-9a5e-4434-97c5-ff115bf9e586 · inbound
Real5-OmniDocBench: A Full-Scale Physical Reconstruction Benchmark for Robust Document Parsing in the Wild MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 28f09890-f0a7-4b23-bd7f-6bc39be97dee · inbound
Logics-Parsing-Omni Technical Report MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation dba10c65-cbe3-4939-b45d-da9ad6c94229 · inbound
Visual-ERM: Reward Modeling for Visual Equivalence MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 7e5ca173-117e-4cbd-be28-634907b1b85e · inbound
Boosting Document Parsing Efficiency and Performance with Coarse-to-Fine Visual Processing MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation eb39dfcf-ab99-401f-b733-d8c5ea767c99 · inbound
Q-Mask: Query-driven Causal Masks for Text Anchoring in OCR-Oriented Vision-Language Models MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 5e2c53d9-5a58-4515-a350-7428cee67402 · inbound
Parser-Oriented Structural Refinement for a Stable Layout Interface in Document Parsing MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 8f56fab7-0eb4-41c2-8843-e7566bff37e8 · inbound
CharTool: Tool-Integrated Visual Reasoning for Chart Understanding MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 41be4938-f2d1-4e54-b221-7dbd782dd1e7 · inbound
InstructTable: Improving Table Structure Recognition Through Instructions MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 0cef38e2-9ff6-4cf5-9a26-6595a41f6fd1 · inbound
MinerU2.5-Pro: Pushing the Limits of Data-Centric Document Parsing at Scale MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 2733da71-3ba4-4760-9c90-0576bfba7102 · inbound
Visual Late Chunking: An Empirical Study of Contextual Chunking for Efficient Visual Document Retrieval MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 8c0534f0-97b8-4363-bd27-665910ca53f7 · inbound
The Perceptual Bandwidth Bottleneck in Vision-Language Models: Active Visual Reasoning via Sequential Experimental Design MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 22de78c7-d998-443e-bf55-9c5167dae9d5 · inbound
The Perceptual Bandwidth Bottleneck in Vision-Language Models: Active Visual Reasoning via Sequential Experimental Design MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 4f03c4df-cb9e-4ae1-92ef-4c05d49cc088 · inbound
Is It Novel and Why? Fine-Grained Patent Novelty Prediction Based on Passage Retrieval MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 7c18a55d-5e74-4ea7-acbf-fe93182b4c01 · inbound
How Far Is Document Parsing from Solved? PureDocBench: A Source-TraceableBenchmark across Clean, Degraded, and Real-World Settings MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation cb55df40-c513-486e-9e4b-86746d7b6be5 · inbound
Information Extraction of Nested Complex Structure of Quantum Cascade Lasers via Large Language Models MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 8a5459e4-6c2b-43ee-89df-a22adc55de57 · inbound
CiteVQA: Benchmarking Evidence Attribution for Trustworthy Document Intelligence MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 29fb8c45-ea74-4d90-837b-d875804041a2 · inbound
UniPPTBench: A Unified Benchmark for Presentation Generation Across Diverse Input Settings MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 6c7d8c8c-3cab-46f4-95f8-798f47a7e05e · inbound
AiraXiv: An AI-Driven Open-Access Platform for Human and AI Scientists MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation e8c0c729-5609-4b36-9035-54bf283284c9 · inbound
MPDocBench-Parse: Benchmarking Practical Multi-page Document Parsing MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 67fa30c4-b48b-47a9-a754-8e52982c1b40 · inbound
MPDocBench-Parse: Benchmarking Practical Multi-page Document Parsing MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation aed0d763-fb88-4e84-b5cf-0af1f25bea24 · inbound
ABot-OCR Technical Report MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation b12a3b30-40e9-4781-86b7-51ab969ef015 · inbound
ToolGate: Token-Efficient Pre-Call Control for Tool-Augmented Vision-Language Agents MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation a994715f-010f-4de8-9c72-e7859120a4ad · inbound
PaddleOCR-VL-1.6: Expanding the Frontier of Document Parsing with Under-Optimized Region Refinement and Progressive Post-Training MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 2834b552-aa6c-49cf-907d-485eda7aab09 · inbound
RT-DocLayout: Real-Time End-to-End Document Layout Analysis with Reading Order in the Wild MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation ad8863fc-922d-48e5-b117-06ef60504f21 · inbound
StrucTab: A Structured Optimization Framework for Table Parsing MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 37d3c6b0-0e50-44ae-bfc2-b56fce723207 · inbound
Infinity-Parser2 Technical Report MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation a5f84112-a861-4163-80f5-bffd6f2b7a51 · inbound
Infinity-Parser2 Technical Report MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 191f36d1-a339-466d-a0fd-6d10bd7dc35b · inbound
DocPO: Advancing Document Policy Optimization via Tailored Step-Aware Rewards MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.