Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-07-14T04:45:32.682508Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 100 of 137 outbound references and 0 inbound Pith citation observations for arXiv:2607.11562.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-07-14T04:45:32.682508Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
100 of 137 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 956487cb-9bff-4207-8694-19f36768c8f5 · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Qwen3-VL Technical Report
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation db94fda7-7ff5-482a-8bfb-7ca78c7587a6 · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Qwen2.5-VL Technical Report
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c8714cae-c14e-4450-8b26-9b73b2f205d1 · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI BEiT: BERT pre-training of image transformers
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9fe81157-476a-4f73-9f90-ef93006201a7 · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Scene text recognition with permuted autoregressive sequence models
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f0b88761-7919-4e83-bbda-a6ca5d329c8f · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Nougat: Neu- ral optical understanding for academic documents
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 63566e54-418f-4025-9ebd-a50ec41a9218 · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Emerging properties in self-supervised vision transformers
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5570681b-8f6e-405a-8024-858bdeaa5c01 · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Encoder-decoder with atrous separable convolution for semantic image segmentation
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7d445a3c-6f95-4497-96f7-92168485031f · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Enhancing tampered text detection through frequency feature fusion and decomposition
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c69d1d07-5d04-4bb0-82ae-e9562d2c703f · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Masked-attention mask transformer for universal image segmentation
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9e94e238-eb65-4307-aa30-b300977122f7 · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Per-pixel classification is not all you need for semantic segmentation.Advances in Neural Information Processing Systems, 34:17864–17875, 2021
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 15c99d84-fa77-4cec-9adf-5b88bdbe8e62 · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI M6doc: a large-scale multi-format, multi-type, multi-layout, multi-language, multi-annotation category dataset for modern document layout analysis
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0765eb9d-c493-48ba-94a6-1fe8a336c462 · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Reproducible scaling laws for contrastive language-image learning
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cf0b5cba-23cb-4d6e-90b0-c82d2a309265 · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Total-text: A comprehensive dataset for scene text detection and recognition
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4de5f807-5022-456f-a1e0-95b72d9cd0e7 · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Icdar2019 robust reading challenge on arbitrary-shaped text-rrc-art
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8786798c-eef4-4a2e-93b1-922796038491 · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI PaddleOCR-VL-1.5: Towards a Multi-Task 0.9B VLM for Robust In-the-Wild Document Parsing
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5483938f-14b9-42b2-a50f-e0a1d955a917 · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Boosting document parsing efficiency and performance with coarse-to-fine visual processing
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b8adad7e-7a38-497c-b2d9-b59a41ea8126 · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Vision grid transformer for document layout analysis
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9b558c00-1ee2-4efc-adc4-cb9060208712 · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Imagenet: A large- scale hierarchical image database
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ea56ecce-d9ee-4c57-92ec-f3911336f702 · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Decaf: A deep convolutional activation feature for generic visual recognition
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d4f1d11a-8712-47a1-b29b-a56f85f0bd87 · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI An image is worth 16x16 words: Transformers for image recognition at scale
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6b5cdcfa-9fb6-498c-b13c-db94b6b22c5a · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Out of length text recognition with sub-string matching
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5e982cc3-5d97-4f1b-87a8-852587c337ed · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Context perception parallel decoder for scene text recognition.IEEE Transactions on Pattern Analysis and Machine Intelligence, 2025
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d4194c98-d26d-4048-95c1-a77fd25bef09 · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Instruction-guided scene text recognition.IEEE Transactions on Pattern Analysis and Machine Intelligence, 47(4):2723–2738, 2025
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2cf2c263-9db2-4b6b-99c9-7cacee222f4a · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Svtrv2: Ctc beats encoder-decoder models in scene text recognition
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 59b938e5-6407-4d77-9935-3a041565c732 · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI UniRec-0.1B: Unified Text and Formula Recognition with 0.1B Parameters
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f145dc73-0f90-4780-8f54-b3f2ef79eeb1 · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Glm-ocr technical report.arXiv preprint arXiv:2603.10910, 2026
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fff5c9ab-7ed2-48b1-967a-b8c2f5dde944 · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Read like humans: Autonomous, bidirectional and iterative language modeling for scene text recognition
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 47b6db9a-c553-4f67-9059-d31968fd052c · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Mathwriting: A dataset for handwrit- ten mathematical expression recognition
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7b74a901-5379-4562-94b5-d6ad3911972e · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI White, Silvia C
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 084acf49-3b43-4494-9446-2e30697605a2 · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Rich feature hierarchies for accurate object detection and semantic segmentation
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation af881e44-1710-45ab-88b6-af1e529bfda2 · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Unresolved cited work
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3ba7009f-7e4a-48f5-802e-b42fa9396dca · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Connectionist temporal classification: labelling unsegmented sequence data with recurrent neural networks
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 079ba129-1e5f-476b-9197-4c523e07b507 · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Speech recognition with deep recurrent neural networks
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 924b39b7-45ae-4723-9350-9dbe46eeb697 · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Unimernet: A universal network for real-world mathematical expression recognition
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 208a16f9-688c-4720-bdb5-cafe636234f4 · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Deep residual learning for image recognition
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c11203c2-00a3-4f30-a964-7f5266998a88 · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Icpr2018 contest on robust reading for multi-type web images
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5444bede-46bd-447c-a75a-c5d8af7bf177 · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Radiov2.5: Improved baselines for agglomerative vision founda- tion models
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 412ae612-9dc8-40b0-a639-0999e1c5fa16 · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI mplug-docowl 1.5: Unified structure learning for ocr-free document understanding
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 61516f06-91f8-4f28-844d-dc6468b132b1 · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Layoutlmv3: Pre-training for document ai with unified text and image masking
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cc7178e2-3580-4242-84b1-5f5fa48438c4 · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Revisiting scene text recognition: A data perspective
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 66fabf0d-d267-4115-9e41-5301b98f32d8 · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Icdar 2015 competition on robust reading
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3277cd3a-9c04-4a71-8da4-7d9890a273ea · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Icdar 2013 robust reading competition
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7e7f5bc5-1c8c-4735-8a4c-e6eef1e25f3d · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Ocr-free document understanding transformer
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2a2c6ce8-c79e-4450-8adf-cec8c915efe4 · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Berg, Wan-Yen Lo, Piotr Dollár, and Ross Girshick
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7a008935-daa7-4bdf-b6f4-02a4e774caa9 · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Open images v5 text annotation and yet another mask text spotter
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 18557dce-9ea7-425c-9c58-b4f7c57eaf36 · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Cat-net: Compression artifact tracing network for detection and localization of image splicing
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 564d561b-e063-452a-822f-368e60237fd4 · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Towards better structured and less noisy web data: Oscar with register annotations
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2ded9b07-7271-4dd3-a061-49d38b8e186e · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Pix2struct: Screenshot parsing as pretraining for visual language understanding
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6b25127e-d632-4c5f-9459-162906b9b441 · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Building a test collection for complex document information processing
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation faa64d08-d7f7-47e4-97af-c5ac7fa7ff1d · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI HunyuanOCR-1.5: Making Lightweight OCR VLMs Faster and Better
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8b4844e4-6028-4cf6-a810-4a09f07cb554 · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Dit: Self-supervised pre-training for document image transformer
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a596092e-5e08-4b02-9ce0-75614f61e452 · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Trocr: Transformer-based optical character recognition with pre- trained models
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 17d8b603-2104-4512-a017-df91a66ce365 · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Openvision: A fully-open, cost- effective family of advanced vision encoders for multimodal learning
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 972ce4b4-7d9b-4c2f-8d84-41cb11d90c3f · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Exploring plain vision transformer backbones for object detection
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 75c1fa4d-8313-48ff-aecb-966bc386a37e · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI dots.ocr: Multilingual document layout parsing in a single vision-language model.arXiv preprint arXiv:2512.02498, 2025
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1a3d7953-d21a-49ba-b273-52d579d63f9b · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Mdpbench: A benchmark for multilingual document parsing in real-world scenarios.arXiv preprint arXiv:2603.28130, 2026
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6227d56c-f8fb-4f99-8f2b-58e87f72a8a1 · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Monkeyocr: Document parsing with a structure-recognition-relation triplet paradigm.arXiv preprint arXiv:2506.05218, 2025
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0eb68467-df77-4ba9-803f-9c9d34794062 · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Monkey: Image resolution and text label are important things for large multi-modal models
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 144f30e9-a5b7-4326-88a3-24dcb77487ae · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Real-time scene text detection with differentiable binarization
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e2980ed5-5696-471b-a537-96f3857086ed · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Improved baselines with visual instruction tuning
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6d9657ec-972e-42d9-b38a-3dc965cd5665 · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Unresolved cited work
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c7491a30-bd47-4025-b790-957ef6961e93 · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Multi-scenario overlapping text segmen- tation with depth awareness
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e298b4ae-4a98-4dba-93b5-159648ec8af2 · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Openvision 2: A family of generative pretrained visual encoders for multimodal learning
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9daeb8bc-1982-4271-b4a8-c63c5472b6c1 · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Multilingual denoising pre-training for neural machine translation.Transactions of the Association for Computational Linguistics, 8:726–742, 2020
Reference 64
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 42d42d74-75d9-4113-a0c0-3fdcdee9a47d · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Curved scene text detection via transverse and longitudinal sequence connection.Pattern Recognition, 90:337–345, 2019
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 57fdb7dc-d7e1-4feb-a155-8d432855c425 · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Ocrbench: on the hidden mystery of ocr in large multimodal models.Science China Information Sciences, 67(12):220102, 2024
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation caf6b75e-cab9-4e8e-9984-09ba86f7f41b · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Textmonkey: An ocr-free large multimodal model for understanding document.IEEE Transac- tions on Pattern Analysis and Machine Intelligence, 48(5):6008–6019, 2026
Reference 67
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fcaa1f93-3a10-4c4c-b150-b04ba9dd335c · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Swin transformer: Hierarchical vision transformer using shifted windows
Reference 68
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5f565518-7516-45c3-bbd0-00ff66612847 · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI A convnet for the 2020s
Reference 69
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b336712c-f93f-4e92-9be6-207c7451ff63 · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Towards end-to-end unified scene text detection and layout analysis
Reference 70
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a177eb79-54e5-4330-a81b-ef6bcfe04176 · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Toward real text manipulation detection: New dataset and new solution.Pattern Recognition, 157:110828, 2025
Reference 71
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 027643cd-dad3-4d0c-a59c-2d95b2fab0ba · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Chartqa: A benchmark for question answering about charts with visual and logical reasoning
Reference 72
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 726ce31e-9039-4cdc-8a26-46d5f706e084 · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Infographicvqa
Reference 73
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 52273c3b-b0fe-478e-8f44-b37eea290133 · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Docvqa: A dataset for vqa on document images
Reference 74
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 486f3503-e9fa-4868-963f-12630bd07d44 · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Scene text recognition using higher order language priors
Reference 75
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 56fee384-a9b9-465d-ae78-1fde29b31db2 · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Mineru2.5: A decoupled vision-language model for efficient high-resolution document parsing
Reference 76
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 70d33beb-373c-4ec1-866e-3aaef0c7c233 · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI DINOv2: Learning Robust Visual Features without Supervision
Reference 77
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 850e8f6c-6ad3-4b18-bb0f-7bc8f9d7b94d · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Omnidocbench: Benchmarking diverse pdf document parsing with comprehensive annotations
Reference 78
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 35a7642d-62b0-4fcc-a69d-e83b270aa870 · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Compositional semantic parsing on semi-structured tables
Reference 79
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 064ab8cd-112e-4251-b2a0-726ce6ab182a · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Doclaynet: A large human-annotated dataset for document-layout segmentation
Reference 80
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4b0091e8-aa73-44c7-a41d-f62da0244813 · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Recogniz- ing text with perspective distortion in natural scenes
Reference 81
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a4f5d735-11b3-4965-ac60-9b86b1252f34 · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI olmocr: Unlocking trillions of tokens in pdfs with vision language models.arXiv preprint arXiv:2502.18443, 2025
Reference 82
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b3a23569-f991-49a6-900c-d324df1eeefe · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI olmocr 2: Unit test rewards for document ocr
Reference 83
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2a064dbd-0d0f-4f43-9adf-068886e97a4d · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Towards robust tampered text detection in document image: New dataset and new solution
Reference 84
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 392ca933-f89d-43f3-a82b-618558fdfbad · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Learning transferable visual models from natural language supervision
Reference 85
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 91df2e0b-c7d5-4bfa-921d-57421cb11d10 · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Am-radio: Agglomerative vision foundation model reduce all domains into one
Reference 86
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2733783f-b31b-4c71-a4eb-2cb09c86fb7a · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Sam 2: Segment anything in images and videos
Reference 87
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c54722b0-ffed-417f-b052-83957b64419f · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI A robust arbitrary text detection system for natural scene images.Expert Systems with Applications, 41(18):8027–8048, 2014
Reference 88
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a1fe7fa5-869c-42a3-be32-390deda8a3ba · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI U-net: Convolutional networks for biomedical image segmentation
Reference 89
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fda23e05-0668-4144-8aa4-635402b19c27 · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Wikimatrix: Mining 135m parallel sentences in 1620 language pairs from wikipedia
Reference 90
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 95f0fb92-a295-428f-a230-5edf53b252dd · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Unresolved cited work
Reference 91
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 986a6732-752f-4795-84d8-61b6e2cc50cf · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI DINOv3
Reference 92
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a4ccb308-4618-4552-b641-99b777849004 · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Textocr: Towards large-scale end-to-end reasoning for arbitrary-shaped scene text
Reference 93
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 32c0aab1-af3d-41b6-b8c9-6631daf00e51 · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Kleister: key information extraction datasets involving long documents with complex layouts
Reference 94
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b2d65c39-6b62-4449-bdea-047b42cd2138 · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Icdar 2019 competition on large-scale street view text with partial labeling-rrc-lsvt
Reference 95
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1dcea6d2-649f-439f-bf23-8555ab3c4094 · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Deepform: Understand structured documents at scale.Weights & Biases report, 4, 2020
Reference 96
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2e4ede48-72c3-4f26-b347-235866852955 · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Hunyuanocr technical report.arXiv preprint arXiv:2511.19575, 2025
Reference 97
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1908098a-9a02-4bdf-b4b0-a2e905dd33c1 · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Kimi K2.5: Visual Agentic Intelligence
Reference 98
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bf371730-be48-44f0-93b4-d31f8282461e · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Kwai Keye-VL Technical Report
Reference 99
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4d2e50e9-25bc-4e2e-a0bf-22e9519f3f8c · outbound
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI SigLIP 2: Multilingual Vision-Language Encoders with Improved Semantic Understanding, Localization, and Dense Features
Reference 100
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.