Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T13:40:39.718350Z
Paper Citation Record · LEDGER
As of 18 August 2026, this Paper Citation Record lists 54 of 54 outbound references and 0 inbound Pith citation observations for arXiv:2412.12902.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T13:40:39.718350Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
54 of 54 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 66ec3495-180b-4d7b-a51c-83e6a5ce15bc · outbound
DoPTA: Improving Document Layout Analysis using Patch-Text Alignment Docformer: End-to-end transformer for document understanding
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8fd7490b-cab5-4011-a31a-31d79b7f9cbf · outbound
DoPTA: Improving Document Layout Analysis using Patch-Text Alignment Visual and textual deep feature fusion for document image classification
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 9bea97bc-1f37-4130-a7d9-2aaf2f2a2339 · outbound
DoPTA: Improving Document Layout Analysis using Patch-Text Alignment Eaml: Ensemble self-attention-based mu- tual learning network for document image classification,
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 0333c493-8e97-4c62-9425-584319afcd0e · outbound
DoPTA: Improving Document Layout Analysis using Patch-Text Alignment BEiT: BERT Pre-Training of Image Transformers
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e9c3414a-0dbb-4003-9c61-65d09ea10892 · outbound
DoPTA: Improving Document Layout Analysis using Patch-Text Alignment Gritsenko, Matthias Minderer, Charles Blundell, Razvan Pascanu, and Jovana Mitrovi´c
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation b65dd56a-4296-4227-854b-d3697cd3fa76 · outbound
DoPTA: Improving Document Layout Analysis using Patch-Text Alignment Cascade r-cnn: High quality object detection and instance segmentation
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 682cc618-ae98-41f9-bca2-cc528aa83daf · outbound
DoPTA: Improving Document Layout Analysis using Patch-Text Alignment Emerg- ing properties in self-supervised vision transformers
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8400b773-7954-4b58-8269-14422081365a · outbound
DoPTA: Improving Document Layout Analysis using Patch-Text Alignment Conceptual 12m: Pushing web-scale image-text pre- training to recognize long-tail visual concepts
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 49f50b91-07a4-4063-9c8f-30f4706b3949 · outbound
DoPTA: Improving Document Layout Analysis using Patch-Text Alignment A simple framework for contrastive learning of visual representations
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 89747ae7-cac3-433a-8592-12cfd1b1d7e9 · outbound
DoPTA: Improving Document Layout Analysis using Patch-Text Alignment Big self-supervised mod- els are strong semi-supervised learners
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation f32e35b1-f074-45e7-a5b6-6eb998f3ce3a · outbound
DoPTA: Improving Document Layout Analysis using Patch-Text Alignment Uniter: Universal image-text representation learning
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 67578e09-632f-4fae-a273-f183d55d4f1f · outbound
DoPTA: Improving Document Layout Analysis using Patch-Text Alignment Vision grid transformer for document layout analysis
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 70680e1c-a978-45ea-84b2-2887f04cb98f · outbound
DoPTA: Improving Document Layout Analysis using Patch-Text Alignment An image is worth 16x16 words: Transformers for image recognition at scale
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation b16bcab3-423f-4d11-bfad-252ed5f3bbe1 · outbound
DoPTA: Improving Document Layout Analysis using Patch-Text Alignment Bootstrap your own latent-a new approach to self-supervised learning
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e633ae88-9220-4280-8de8-f86f6947bfe1 · outbound
DoPTA: Improving Document Layout Analysis using Patch-Text Alignment Evaluation of deep convolutional nets for document image classification and retrieval
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 8d25c0fe-45d6-41a7-afdd-f20c1c32357d · outbound
DoPTA: Improving Document Layout Analysis using Patch-Text Alignment Momentum contrast for unsupervised visual rep- resentation learning
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a02407b9-2880-48ab-b9de-61c0d382e819 · outbound
DoPTA: Improving Document Layout Analysis using Patch-Text Alignment Masked autoencoders are scalable vision learners
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 3383098b-8d62-40c2-869e-743b5fe97593 · outbound
DoPTA: Improving Document Layout Analysis using Patch-Text Alignment Bros: A pre-trained lan- guage model focusing on text and layout for better key infor- mation extraction from documents
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 60f325e4-3b2f-490f-a045-690fcb2b1e8b · outbound
DoPTA: Improving Document Layout Analysis using Patch-Text Alignment Layoutlmv3: Pre-training for document ai with unified text and image masking
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 323b5738-e0c2-494a-bf73-211acd795769 · outbound
DoPTA: Improving Document Layout Analysis using Patch-Text Alignment Icdar2019 compe- tition on scanned receipt ocr and information extraction
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 4b133606-b0da-42ab-a007-0f1a1399d918 · outbound
DoPTA: Improving Document Layout Analysis using Patch-Text Alignment Unresolved cited work
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 444b683c-04af-42ed-a478-d1265eb77b62 · outbound
DoPTA: Improving Document Layout Analysis using Patch-Text Alignment Funsd: A dataset for form understanding in noisy scanned documents, 2019
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 4d4173ea-df47-4d37-9d8f-d67018a45527 · outbound
DoPTA: Improving Document Layout Analysis using Patch-Text Alignment Scaling up visual and vision-language representa- tion learning with noisy text supervision
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ea22a01c-c224-4ff8-8aa3-aaa0034c5661 · outbound
DoPTA: Improving Document Layout Analysis using Patch-Text Alignment Ocr-free document understanding transformer
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation ea3e7d8c-81de-4470-aa51-c34996a3bc19 · outbound
DoPTA: Improving Document Layout Analysis using Patch-Text Alignment Dit: Self-supervised pre-training for docu- ment image transformer
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation c8d01da3-b4e8-4be9-ab56-c66f1ed1f1f8 · outbound
DoPTA: Improving Document Layout Analysis using Patch-Text Alignment Grounded language-image pre-training, 2022
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bfc282d1-3098-4457-bef1-d54bcbc8b982 · outbound
DoPTA: Improving Document Layout Analysis using Patch-Text Alignment Grounded language-image pre-training
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation c4734ac9-db7e-4647-8924-73316f72819b · outbound
DoPTA: Improving Document Layout Analysis using Patch-Text Alignment DocBank: A Benchmark Dataset for Document Layout Analysis
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 43b1ff01-eb7e-4580-8647-fe82aa0ad910 · outbound
DoPTA: Improving Document Layout Analysis using Patch-Text Alignment Bi-VLDoc: Bidirectional Vision-Language Modeling for Visually-Rich Document Understanding
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 74d9812c-b55e-4de7-ab11-fc769ecdbf2e · outbound
DoPTA: Improving Document Layout Analysis using Patch-Text Alignment Docvqa: A dataset for vqa on document images
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 560bf376-4f24-4416-a1b2-1b777202dd5b · outbound
DoPTA: Improving Document Layout Analysis using Patch-Text Alignment Infographicvqa
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 953a4027-a9fa-4bcc-8324-f6f675b6ec19 · outbound
DoPTA: Improving Document Layout Analysis using Patch-Text Alignment DINOv2: Learning Robust Visual Features without Supervision
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d869e613-2f33-43ad-9ea8-26a9094cacf4 · outbound
DoPTA: Improving Document Layout Analysis using Patch-Text Alignment {CORD}: A consolidated receipt dataset for post-{ocr} parsing
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 93011028-67ea-4c9f-a3cc-353573530b1f · outbound
DoPTA: Improving Document Layout Analysis using Patch-Text Alignment Doclaynet: A large human- annotated dataset for document-layout segmentation
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 47f192b0-31d2-4a0b-9bac-37cddd362576 · outbound
DoPTA: Improving Document Layout Analysis using Patch-Text Alignment Going full-tilt boogie on document understanding with text-image-layout transformer
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation fa9145ef-c848-49f8-9306-68dbc95b562f · outbound
DoPTA: Improving Document Layout Analysis using Patch-Text Alignment Learning transferable visual models from natural language supervi- sion
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 62979f90-35aa-4ffc-9a99-a2bb6acb48b9 · outbound
DoPTA: Improving Document Layout Analysis using Patch-Text Alignment Imagenet large scale visual recognition challenge
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1e2f8194-9003-49d7-8273-f33e93b20574 · outbound
DoPTA: Improving Document Layout Analysis using Patch-Text Alignment Laion-5b: An open large-scale dataset for training next generation image-text models
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 65695760-db70-443f-aee2-426a6b7ce92f · outbound
DoPTA: Improving Document Layout Analysis using Patch-Text Alignment Complex document information processing (cdip) dataset, 2022
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 345b4aa3-8ecc-4782-8dd1-45a40c9bd442 · outbound
DoPTA: Improving Document Layout Analysis using Patch-Text Alignment Kleister: key in- formation extraction datasets involving long documents with complex layouts
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 5cabec59-b2ba-4b11-8a65-3bbbc638f022 · outbound
DoPTA: Improving Document Layout Analysis using Patch-Text Alignment Vl-bert: Pre-training of generic visual- linguistic representations
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 2028df8f-6ae4-4a69-aa05-8d38aa65529e · outbound
DoPTA: Improving Document Layout Analysis using Patch-Text Alignment Revisiting unreasonable effectiveness of data in deep learning era
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 8e70fc14-c584-4731-89ec-6bc277c3b6d3 · outbound
DoPTA: Improving Document Layout Analysis using Patch-Text Alignment Unifying vision, text, and layout for universal document processing
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 7bb084c8-b4ae-4cd3-9a6d-86c1bb614771 · outbound
DoPTA: Improving Document Layout Analysis using Patch-Text Alignment Yfcc100m: The new data in multimedia research
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c4e96163-fed6-436d-98d9-d8d38625d646 · outbound
DoPTA: Improving Document Layout Analysis using Patch-Text Alignment Training data-efficient image transformers & distillation through at- tention, 2021
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 91bc9e1f-cdd3-455c-81da-290aef275597 · outbound
DoPTA: Improving Document Layout Analysis using Patch-Text Alignment Detectron2
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b5a48ba7-8102-4142-8dab-dd75a01539d3 · outbound
DoPTA: Improving Document Layout Analysis using Patch-Text Alignment Aggregated residual transformations for deep neural networks, 2017
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 90ce9152-23a7-469c-bfe6-609d473f6bb1 · outbound
DoPTA: Improving Document Layout Analysis using Patch-Text Alignment Layoutlm: Pre-training of text and layout for document image understanding
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d8d7e7f7-c9b3-4c1b-a072-821c6f2242fb · outbound
DoPTA: Improving Document Layout Analysis using Patch-Text Alignment LayoutLMv2: Multi-modal Pre-training for Visually-Rich Document Understanding
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b6ede840-0a43-439c-b526-01ee219b47c5 · outbound
DoPTA: Improving Document Layout Analysis using Patch-Text Alignment FILIP: Fine-grained interactive language- image pre-training
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 3c7e747a-4390-4ab8-b9ed-b9d9b70a1ea0 · outbound
DoPTA: Improving Document Layout Analysis using Patch-Text Alignment StrucTexTv2: Masked Visual-Textual Prediction for Document Image Pre-training
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 97cf0576-2402-4cb0-9cd7-0e7c19cab44c · outbound
DoPTA: Improving Document Layout Analysis using Patch-Text Alignment Sigmoid loss for language image pre-training,
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2c154ed6-f2a5-4984-8f95-5f5575dfe960 · outbound
DoPTA: Improving Document Layout Analysis using Patch-Text Alignment Zhang, H
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 58d87be3-8112-40c9-9d50-d2d0c2364bd7 · outbound
DoPTA: Improving Document Layout Analysis using Patch-Text Alignment Pub- laynet: largest dataset ever for document layout analysis
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
No inbound Pith citation observations are available.