Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T22:38:52.885866Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 69 of 69 outbound references and 1 inbound Pith citation observation for arXiv:2506.21055.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T22:38:52.885866Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-06T01:00:45.178631Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-06T01:00:46.378040Z
69 of 69 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 921fc7de-ed9d-4b22-b318-c52ffdcc0946 · outbound
Class-Agnostic Region-of-Interest Matching in Document Images Visual text processing: A comprehensive review and unified evaluation, 2025
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation df18eaf1-bf02-408a-ab43-ce6eb87f244d · outbound
Class-Agnostic Region-of-Interest Matching in Document Images TextCtrl: Diffusion-based scene text editing with prior guidance control
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1d63aa5d-024f-478c-85b4-aac2aa887073 · outbound
Class-Agnostic Region-of-Interest Matching in Document Images The devil is in fine-tuning and long-tailed problems: A new benchmark for scene text detection
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c0631c98-09bc-439b-96dc-6aab131559b3 · outbound
Class-Agnostic Region-of-Interest Matching in Document Images Perceiving ambiguity and semantics without recognition: An efficientandeffectiveambiguousscenetextdetector
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation b7258021-25c7-4364-aaf7-fb2b4d237f6a · outbound
Class-Agnostic Region-of-Interest Matching in Document Images Self-training for domain adaptive scene text detection
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a18ddd1f-7c39-40ab-9387-09568c8e5927 · outbound
Class-Agnostic Region-of-Interest Matching in Document Images SEED: Semantics enhanced encoder-decoder framework for scene text recognition
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation befb8960-ebf7-4d31-93d5-b3eedda41466 · outbound
Class-Agnostic Region-of-Interest Matching in Document Images PIMNet: A parallel, iterative and mimicking network for scene text recognition
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e2015c11-edbe-4f8a-bbc2-2ca4c0337895 · outbound
Class-Agnostic Region-of-Interest Matching in Document Images IPAD: Iterative, parallel, and diffusion- based network for scene text recognition.IJCV, 2025
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 8f606f82-ca5a-4ca6-a420-9a38f1192faf · outbound
Class-Agnostic Region-of-Interest Matching in Document Images Linguistics-aware masked image modeling for self-supervised scene text recognition
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 8c5985b4-d06c-44e9-89e7-35297a2116db · outbound
Class-Agnostic Region-of-Interest Matching in Document Images Divide rows and conquer cells: Towards structure recognition for large tables
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 2612607d-2fcd-45c7-b656-4f5e1e9d79aa · outbound
Class-Agnostic Region-of-Interest Matching in Document Images Arbitrary reading order scene text spotter with local semantics guidance
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 072b1ef4-639b-4fd7-92cc-c02dec23a787 · outbound
Class-Agnostic Region-of-Interest Matching in Document Images TPSNet: Reverse thinking of thin plate splines for arbitrary shape scene text representation
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation bf975ee4-71aa-4dee-b357-45d8d148301d · outbound
Class-Agnostic Region-of-Interest Matching in Document Images TextBlockV2: Towards precise-detection-free scene text spotting with pre-trained language model.TOMM, 2025
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 21ebe37c-f3b9-43a7-b60f-1a18f559afda · outbound
Class-Agnostic Region-of-Interest Matching in Document Images Beyond cropped regions: New benchmark and corresponding baseline for chinese scene text retrieval in diverse layouts
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 454355a7-ea8f-4663-9f4e-cead2a9f40ee · outbound
Class-Agnostic Region-of-Interest Matching in Document Images Focus, distinguish, and prompt: Unleashing clip for efficient and flexible scene text retrieval
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 4099a824-0b80-49ba-b8e6-9108ffe30297 · outbound
Class-Agnostic Region-of-Interest Matching in Document Images LDP: Generalizing to multilingual visual information extraction by language decoupled pretraining
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7733c693-127c-4ebf-a474-78af895e95e5 · outbound
Class-Agnostic Region-of-Interest Matching in Document Images Beyond OCR+ VQA: Involving ocr into the flow for robust and accurate textvqa
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 60a873fb-d4fa-43f5-89c3-1a5efd2ab323 · outbound
Class-Agnostic Region-of-Interest Matching in Document Images Publaynet: largest dataset ever for document layout analysis
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation cd3e0c06-de16-45ab-a706-873f2996d133 · outbound
Class-Agnostic Region-of-Interest Matching in Document Images Learning to extract semantic structure from documents using multimodal fully convolutional neural networks
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 8767b378-67d6-4a8c-9eee-a328f48daba3 · outbound
Class-Agnostic Region-of-Interest Matching in Document Images VSR: a unified framework for document layout analysis combining vision, semantics and relations
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 0cb19cfa-b771-4b1d-8803-ce424b18c711 · outbound
Class-Agnostic Region-of-Interest Matching in Document Images Attention where it matters: Rethinking visualdocumentunderstandingwithselectiveregionconcentration
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e4ec6bb0-f924-43c6-b61f-e0cd703c6372 · outbound
Class-Agnostic Region-of-Interest Matching in Document Images Bros: A pre-trained language model focusing on text and layout for better key information extraction from documents
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a4c66329-f255-4372-805d-1e8ff5af512b · outbound
Class-Agnostic Region-of-Interest Matching in Document Images StrucText: Structured text under- standing with multi-modal transformers
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 2bdf5d4c-4b64-4e8f-b2e2-536026302043 · outbound
Class-Agnostic Region-of-Interest Matching in Document Images Cross-domain few-shot semantic segmentation
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1770ce1e-e27e-436b-ac3b-e21f24fc2741 · outbound
Class-Agnostic Region-of-Interest Matching in Document Images Pixel-by- pixel cross-domain alignment for few-shot semantic segmentation
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation fd412c3c-44a6-4cf5-a217-accfda4fc02a · outbound
Class-Agnostic Region-of-Interest Matching in Document Images Restnet: Boosting cross-domain few-shot segmentation with residual transformation network
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 8fb3ee38-27cd-4b28-aa17-9fb62b5c1263 · outbound
Class-Agnostic Region-of-Interest Matching in Document Images Adapt before comparison: A new perspective on cross-domain few- shot segmentation
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1d813d51-284f-4518-8788-82ea6d872de7 · outbound
Class-Agnostic Region-of-Interest Matching in Document Images Qwen2-vl: Enhancing vision-language model’s perception of the world at any resolution.arXiv, 2024
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 6549c2e4-07d8-425e-ace6-cd27321e2a16 · outbound
Class-Agnostic Region-of-Interest Matching in Document Images Internvl: Scaling up vision foundation models and aligning for generic visual-linguistic tasks
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1a78d34e-21c0-48f7-89c7-24794415f3a5 · outbound
Class-Agnostic Region-of-Interest Matching in Document Images Faster R-CNN: To- wards real-time object detection with region proposal networks.IEEE TPAMI, 39(6):1137–1149, 2016
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 38ef0995-0746-40c9-a746-8330974fa688 · outbound
Class-Agnostic Region-of-Interest Matching in Document Images Mask R-CNN
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ccd31602-49ff-4d2a-bad1-7d7c4c2d1948 · outbound
Class-Agnostic Region-of-Interest Matching in Document Images M6doc: A large-scale multi-format, multi- type, multi-layout, multi-language, multi-annotation category dataset for modern document layout analysis
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 45a2b99d-e596-4b45-9262-528555db8a48 · outbound
Class-Agnostic Region-of-Interest Matching in Document Images Swindocseg- menter: An end-to-end unified domain adaptive transformer for document instance segmentation
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation bdb1cb97-3bf9-4cb2-8b7d-2609161a6dc4 · outbound
Class-Agnostic Region-of-Interest Matching in Document Images Dino: Detr with improved denoising anchor boxes for end-to- end object detection
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d3f580ce-5f7a-44ab-837c-11e42be7dd61 · outbound
Class-Agnostic Region-of-Interest Matching in Document Images Selfdocseg: A self-supervised vision- based approach towards document segmentation
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 6c2e2959-6cb4-42e9-ab25-eede11b423c8 · outbound
Class-Agnostic Region-of-Interest Matching in Document Images Bootstrap your own latent-a new approach to self-supervised learning.NeurIPS, 33:21271–21284, 2020
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 4fd51609-9479-4c31-8847-8f59f16e751c · outbound
Class-Agnostic Region-of-Interest Matching in Document Images LayoutLM: Pre-training of text and layout for document image understanding
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation bd1cdc0b-a5f2-44ec-9107-a78563609979 · outbound
Class-Agnostic Region-of-Interest Matching in Document Images LayoutLMv2: Multi-modal pre-training for visually-rich document understanding
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c7a268cc-318d-4031-a919-113652112c41 · outbound
Class-Agnostic Region-of-Interest Matching in Document Images LayoutLMv3: Pre-training for document ai with unified text and image masking
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ba44459e-6ed5-46a9-be01-a926abcc3ff1 · outbound
Class-Agnostic Region-of-Interest Matching in Document Images BEiT: Bert pre-training of image transformers
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ed2a9473-153f-43df-9191-7e23a1845954 · outbound
Class-Agnostic Region-of-Interest Matching in Document Images Dit: Self-supervised pre-training for document image transformer
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation bec3a72a-68e0-4b71-8ee4-6e1ef1ee57f5 · outbound
Class-Agnostic Region-of-Interest Matching in Document Images Unidoc: Unified pretraining framework for document understanding.NeurIPS, 34:39–50, 2021
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 9e3fc858-264f-42e4-85a7-4d89f5acca7e · outbound
Class-Agnostic Region-of-Interest Matching in Document Images StrucTexTv2: Masked visual-textual prediction for document image pre-training
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c2564ae8-2255-4ef7-828e-ca25a88b8418 · outbound
Class-Agnostic Region-of-Interest Matching in Document Images Apseg: auto-prompt network for cross-domain few-shot semantic seg- mentation
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 80159f3f-5dad-4ee6-9529-1f87ec88f970 · outbound
Class-Agnostic Region-of-Interest Matching in Document Images LayoutReader: Pre-training of Text and Layout for Reading Order Detection
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0390b3f2-4f8d-4df2-aa48-69cc5f54a847 · outbound
Class-Agnostic Region-of-Interest Matching in Document Images Reading order matters: Information extraction from visually-rich documents by token path prediction
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 6b26cf17-dd66-49f4-9a30-dd7cdb7bedb3 · outbound
Class-Agnostic Region-of-Interest Matching in Document Images LayoutXLM: Multimodal Pre-training for Multilingual Visually-rich Document Understanding
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7e986b6d-11bb-439b-bcee-4ef7f39d01c5 · outbound
Class-Agnostic Region-of-Interest Matching in Document Images Query-driven generative network for document infor- mation extraction in the wild
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 31d1c692-08ff-4ada-8599-57c0c2b10db9 · outbound
Class-Agnostic Region-of-Interest Matching in Document Images Unifying vision, text, and layout for universal document processing
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 68e6649e-86f3-4a96-b3eb-48491fb2c388 · outbound
Class-Agnostic Region-of-Interest Matching in Document Images Pix2struct: Screenshot parsing as pretraining for visual lan- guage understanding
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 64f29873-8af6-4ea0-a559-09d700bfab73 · outbound
Class-Agnostic Region-of-Interest Matching in Document Images Prestu: Pre-training for scene-text understanding
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e4bb6a95-268f-453a-8315-0ee722efac8b · outbound
Class-Agnostic Region-of-Interest Matching in Document Images Ocr-free document understanding transformer
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 3a004b56-f5cd-4783-b87c-90adcaba4625 · outbound
Class-Agnostic Region-of-Interest Matching in Document Images Omniparser: A unified framework for text spotting key information extraction and table recognition
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation da40e932-5620-4d21-a338-cd5687f3217c · outbound
Class-Agnostic Region-of-Interest Matching in Document Images Icdar 2023 competition on structured text extraction from visually-rich document images
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 94aaa468-dfda-487a-ae04-e1e26405d1f2 · outbound
Class-Agnostic Region-of-Interest Matching in Document Images Visual information extraction in the wild: practical dataset and end-to-end solution
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 026f45d3-d935-4b2d-95da-1b95a43a1431 · outbound
Class-Agnostic Region-of-Interest Matching in Document Images Towards robust visual information extraction in real world: new dataset and novel solution
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation bdb503a7-3326-4655-8877-de82da18290f · outbound
Class-Agnostic Region-of-Interest Matching in Document Images Vision grid transformer for document layout analysis
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation dc544755-391a-4288-92c4-ff83b9bb14cc · outbound
Class-Agnostic Region-of-Interest Matching in Document Images Fully convolutional networks for semantic segmentation
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 01ab366c-0e36-4a60-bf49-d9fff2658d86 · outbound
Class-Agnostic Region-of-Interest Matching in Document Images Real-time scene text detection with differentiable binarization
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a0e08d6b-a1dc-47a5-97b3-d87dc9b97941 · outbound
Class-Agnostic Region-of-Interest Matching in Document Images Attention is all you need
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fe67d8f9-c8f3-483c-9b5e-a9ed0e9f5986 · outbound
Class-Agnostic Region-of-Interest Matching in Document Images Efficientandaccuratearbitrary-shapedtextdetectionwith pixel aggregation network
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 6fc7b2c6-a439-431f-ae7e-d7dd0309327d · outbound
Class-Agnostic Region-of-Interest Matching in Document Images Shape robust text detection with progressive scale expansion network
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation edf911ce-c13f-4f4a-a370-2b9959512f2a · outbound
Class-Agnostic Region-of-Interest Matching in Document Images A generic solution to polygon clipping.Communications of the ACM, 35(7):56–63, 1992
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 395d7cfd-e4d7-463e-9f3e-02eaa991ba05 · outbound
Class-Agnostic Region-of-Interest Matching in Document Images V-net: Fully convolu- tional neural networks for volumetric medical image segmentation
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 09de82e0-2bbc-4445-9909-693faa3671fb · outbound
Class-Agnostic Region-of-Interest Matching in Document Images Training region-based object detectors with online hard example mining
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 94ac7c68-494d-4262-8737-05185d07ddc9 · outbound
Class-Agnostic Region-of-Interest Matching in Document Images Dmt-net: Deep multiple networks for low-light image enhancement based on retinex model.IEEE Access, 11:132147–132161, 2023
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 999875d5-e24c-4821-9bb2-7c0f694d397b · outbound
Class-Agnostic Region-of-Interest Matching in Document Images Unresolved cited work
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 465e695b-7e91-463e-8c67-c07c5f834a6f · outbound
Class-Agnostic Region-of-Interest Matching in Document Images Deep residual learning for image recognition
Reference 68
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 052211b5-7a88-4440-8b13-39a46988d0f8 · outbound
Class-Agnostic Region-of-Interest Matching in Document Images A survey on curriculum learning.IEEE TPAMI, 44(9):4555–4576, 2021
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 661b1476-0953-407c-9c64-6fdd8c1b667d · inbound
Uni-DocDiff: A Unified Document Restoration Model Based on Diffusion Class-Agnostic Region-of-Interest Matching in Document Images
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.