Pith. sign in

Paper Citation Record · LEDGER

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition

As of 14 August 2026, this Paper Citation Record lists 60 of 60 outbound references and 0 inbound Pith citation observations for arXiv:2411.11219.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.11219 v2

Coverage vector

measured 60 of 60 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T18:51:51.260140Z

measured 60 of 60 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

60 of 60 outbound references displayed

  • verified exact1
  • verified fuzzy31
  • unresolved28
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation cfcb1dca-9492-43bb-9f53-93c1fc5400a8 · outbound

This paper cites Unsupervised feature learning via non-parametric instance discrimination,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Unsupervised feature learning via non-parametric instance discrimination,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:53.038960Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T18:51:50.760300Z digest=sha256:f58e3b1c26832206a726614f2801964611fc13968268051cc5b4573289ccd73b

Observation ed596de8-4d2b-42d5-bca6-f7660c7ef85f · outbound

This paper cites Representation Learning with Contrastive Predictive Coding.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Representation Learning with Contrastive Predictive Coding

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:50.769703Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:50.769703Z digest=sha256:db33dde27fa3093e0bce14826e03310e4a9181948df5bbce72085cb78ae60fc0

Observation fcf23d59-0ff2-495a-ac06-a129da57c68a · outbound

This paper cites Momentum contrast for unsupervised visual representation learning,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Momentum contrast for unsupervised visual representation learning,

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:50.776195Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:50.776195Z digest=sha256:302512873c7ba8c6534f57f2bda9817c8f02da5af5aa7e2c9eb89a4c3204cb91

Observation f58e6fd3-0700-4869-901f-bfd79e4cdb88 · outbound

This paper cites A simple framework for contrastive learning of visual representations,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition A simple framework for contrastive learning of visual representations,

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:50.783247Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:50.783247Z digest=sha256:7968553698e75dee277107648b185317ed6c3bdbbb5f4b3cf6462f576361e71a

Observation 00ff5207-7c73-482f-b8e4-d43d0f45d396 · outbound

This paper cites BEiT: BERT Pre-Training of Image Transformers.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition BEiT: BERT Pre-Training of Image Transformers

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:50.789743Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:50.789743Z digest=sha256:195960ecbbd6c9e1e7161a76584ff2549ef75a91c8413be2a81b81d3b1614719

Observation d8f214a5-cd92-45f6-85c0-69b4b94695b8 · outbound

This paper cites Masked au- toencoders are scalable vision learners,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Masked au- toencoders are scalable vision learners,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:50.796041Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:50.796041Z digest=sha256:0890ab2a01973271b12ca92a3280f1d8bb7446d7d29c19e8df9c1e121523f963

Observation e4dc2277-8f68-4774-87c0-54b6a897e8bc · outbound

This paper cites Simmim: A simple framework for masked image modeling,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Simmim: A simple framework for masked image modeling,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.963075Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T18:51:50.805358Z digest=sha256:fa5ac2ee9da19519768e5ca8d8cff9d6696563261616be0b198100a7608e6feb

Observation b08f27d2-29f0-450f-ace9-59ed3504491d · outbound

This paper cites Read like humans: Autonomous, bidirectional and iterative language modeling for scene text recognition,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Read like humans: Autonomous, bidirectional and iterative language modeling for scene text recognition,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.946358Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T18:51:50.816191Z digest=sha256:770e341ddac47377526cea34d3a360159a71ed09c23337deb4c0ca0026fe681a

Observation c8ab4525-95a7-48b8-864a-f1e500388b9f · outbound

This paper cites Sequence-to-sequence contrastive learn- ing for text recognition,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Sequence-to-sequence contrastive learn- ing for text recognition,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:50.837653Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:50.837653Z digest=sha256:0e42c5751c78d7a2bcec72cf95385f0fb2bc43397bb9238babbf95469efec7ed

Observation 88f6549d-a09e-4af1-a8f8-5efb4b2cafdb · outbound

This paper cites Perceiving stroke-semantic context: Hierarchical contrastive learning for robust scene text recognition,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Perceiving stroke-semantic context: Hierarchical contrastive learning for robust scene text recognition,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.841789Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T18:51:50.844333Z digest=sha256:17da892ef4d61cc11c26e205143db405ff519f51bdbb90910b5f27c604f3ee81

Observation 33351973-32fb-4f94-9b5f-471929891d37 · outbound

This paper cites Reading and writing: Discriminative and generative modeling for self-supervised text recognition,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Reading and writing: Discriminative and generative modeling for self-supervised text recognition,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:50.852070Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:50.852070Z digest=sha256:8692f3b8eb5f489fa2d6d497ae90928fca5a4cb34c6abe9a351f0e10536334da

Observation 7229e8de-c376-4324-8a17-e29d8fa895b1 · outbound

This paper cites An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:50.857745Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:50.857745Z digest=sha256:f2b106a147695170e2dd767cb34120f1fc862b52cf4aaa28744263b1323e24fd

Observation 3e913a1d-5220-4e8e-86fd-8c3c80b5d31f · outbound

This paper cites MaskOCR: Text Recognition with Masked Encoder-Decoder Pretraining.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition MaskOCR: Text Recognition with Masked Encoder-Decoder Pretraining

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:50.864066Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:50.864066Z digest=sha256:7c330c74e0ee937601b953ca235a2c84cff0e5acfed86aa237fa57d6f943ab08

Observation fc24f0a7-bc9b-4f8e-a57c-6ed45b9eefa6 · outbound

This paper cites Towards accurate scene text recognition with semantic reasoning networks,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Towards accurate scene text recognition with semantic reasoning networks,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.783768Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T18:51:50.870943Z digest=sha256:2e993df711f93575bd23a40b44bdae20ee1c7de127299db82e596a9821139e4b

Observation 964ec38a-f7e9-4439-994f-5dde280b75d9 · outbound

This paper cites Seed: Semantics enhanced encoder-decoder framework for scene text recognition,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Seed: Semantics enhanced encoder-decoder framework for scene text recognition,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.754387Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T18:51:50.877425Z digest=sha256:4dca86791af520941d78a9ab59108d7ad7be29cf8591cd7c1b03483616ff6d90

Observation a3838276-dd2e-420c-9379-821d21227ddf · outbound

This paper cites On vocabulary reliance in scene text recognition,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition On vocabulary reliance in scene text recognition,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.717253Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T18:51:50.885900Z digest=sha256:b08c038311ccb71db02280508f770eb04da8e73824767d7e54225b6e90c4cbf1

Observation 70b6e185-263e-4787-9fa5-a6a20773d936 · outbound

This paper cites Ressl: Relational self-supervised learning with weak augmentation,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Ressl: Relational self-supervised learning with weak augmentation,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.676841Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T18:51:50.892516Z digest=sha256:4b265ce4fb0d7b356389e7f8e01267ea961d8d16d180095d474bb529ffe9d562

Observation a0266b4a-835a-475c-959f-757b35f52242 · outbound

This paper cites Synthetic data for text localisation in natural images,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Synthetic data for text localisation in natural images,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:50.899694Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:50.899694Z digest=sha256:7fbda1371370584d6bb9fc26075427198b04a2c6afb1c3797dbec054245b0de0

Observation 2f0c00b6-91b6-4bfd-98bc-4ef005f9cb42 · outbound

This paper cites Rethink- ing text segmentation: A novel dataset and a text-specific refinement approach,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Rethink- ing text segmentation: A novel dataset and a text-specific refinement approach,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.623318Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T18:51:50.908188Z digest=sha256:9b148d435472e3540120623591a8de296c56dd30a4467ed36a01591bafbacc1a

Observation 1c9b8423-121c-463b-8ab5-362d7b5c435f · outbound

This paper cites Relational contrastive learning for scene text recognition,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Relational contrastive learning for scene text recognition,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.598854Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T18:51:50.914335Z digest=sha256:de6586d42072860a754142adbc256d8dc334027a2516c42cd1538302c7402711

Observation 60134332-9388-44ca-a69a-d6ea42533e25 · outbound

This paper cites Unsupervised learning of visual features by contrasting cluster assign- ments,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Unsupervised learning of visual features by contrasting cluster assign- ments,

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:50.923630Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:50.923630Z digest=sha256:eede2b0e1f68c7eb530588b9bac44fce91533b5b828ce21723749dc6d9903cb9

Observation 183fe724-b5f8-4c46-a267-ca6f8f487dd6 · outbound

This paper cites Bootstrap your own latent-a new approach to self-supervised learning,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Bootstrap your own latent-a new approach to self-supervised learning,

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:50.930894Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:50.930894Z digest=sha256:14c6b1dde1fa628d44bf586852438e1b7f197d54144b0c9dfe05822ccb669968

Observation 4a006d29-137f-4795-99a6-e0c851631114 · outbound

This paper cites Exploring simple siamese representation learning,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Exploring simple siamese representation learning,

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:50.940508Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:50.940508Z digest=sha256:2009757af36ee3338ecce9d76cc508ca76fcba276448ebc00b49a818e85e4682

Observation 27ea1a1d-6ecd-45c2-b8f2-0073a227c032 · outbound

This paper cites Emerging properties in self-supervised vision transformers,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Emerging properties in self-supervised vision transformers,

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:50.945778Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:50.945778Z digest=sha256:dd90353b14215a2948c58dc26b33220e44b30a8a40fac2a6578b0cc38495993e

Observation b91e51fb-4ca1-4577-9ff7-223bb5a18fae · outbound

This paper cites iBOT: Image BERT Pre-Training with Online Tokenizer.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition iBOT: Image BERT Pre-Training with Online Tokenizer

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:50.951689Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:50.951689Z digest=sha256:4ef4aed94678aed703421861e8b64cb8ff1991a1bac38ac5744bdae4238ef519

Observation 01b76be6-5c32-471f-a601-c6c1fe7fa17e · outbound

This paper cites DINOv2: Learning Robust Visual Features without Supervision.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition DINOv2: Learning Robust Visual Features without Supervision

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:50.962274Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:50.962274Z digest=sha256:47f1a1b8e3ef9d1687e30d4b4f22da16a97a19ab3aa81d1dbddcac9c1a824f93

Observation d26d25f2-e35c-457a-99d7-0e3c0d0c14f2 · outbound

This paper cites Contrastive learning with stronger augmenta- tions,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Contrastive learning with stronger augmenta- tions,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.477016Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T18:51:50.969284Z digest=sha256:542da8031ba5209261b30b8a4ca74fc7ccb682f2a436a292811febef62bf9dc4

Observation a3086599-781f-4b39-b4fd-f2ccd2f2e188 · outbound

This paper cites What if we only use real datasets for scene text recognition? toward scene text recognition with fewer labels,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition What if we only use real datasets for scene text recognition? toward scene text recognition with fewer labels,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.440978Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T18:51:50.978025Z digest=sha256:2c0250ac37aa5bbe72177244dbe35c790cad03d8f250624ab1319273d056f6be

Observation ddbd37fd-da62-400b-9b57-70ae7bf42747 · outbound

This paper cites Siman: Exploring self-supervised rep- resentation learning of scene text via similarity-aware normalization,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Siman: Exploring self-supervised rep- resentation learning of scene text via similarity-aware normalization,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.404622Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T18:51:50.987664Z digest=sha256:6ee164d4e6aa2c9f95dd8edd48956cfd9a0ceb56b1f6d97389ce9e1ca47cbb0b

Observation 93e80a5d-fbbe-4a0f-8e1d-f1c8f777b9b4 · outbound

This paper cites Self- supervised character-to-character distillation for text recognition,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Self- supervised character-to-character distillation for text recognition,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.377343Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T18:51:50.996746Z digest=sha256:aa51ae33cc9b2d5aeb5ee85c65d0e307afc023afa391237b2441563d0e7e4ba4

Observation b2444aa3-aacf-4583-ab26-7478617f1f99 · outbound

This paper cites Reading scene text in deep convolutional sequences,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Reading scene text in deep convolutional sequences,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.350919Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T18:51:51.007434Z digest=sha256:4bb3bdeaa49e16a1fc75bc9f70138c38b936fbc733fddb06a2ec16bc2e6574db

Observation 70bd8651-a4a6-4636-800a-56f7253b28c1 · outbound

This paper cites Accurate recognition of words in scenes without char- acter segmentation using recurrent neural network,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Accurate recognition of words in scenes without char- acter segmentation using recurrent neural network,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.322809Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T18:51:51.039572Z digest=sha256:00a90b211b3ff7d865156b81fb8f92331d81f5a03ac359fc08758f205925a338

Observation b8a625bb-108a-48a2-bdbb-333c455433da · outbound

This paper cites An end-to-end trainable neural network for image-based sequence recognition and its application to scene text recognition,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition An end-to-end trainable neural network for image-based sequence recognition and its application to scene text recognition,

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:51.049965Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:51.049965Z digest=sha256:e5498f8e817bfcd7eaf8bbdf3f02d629de10c65f7b05c6a6e826a98bdda7659f

Observation e4f90b37-65c5-4c45-bb1b-5e6f01e50b3a · outbound

This paper cites Aster: An attentional scene text recognizer with flexible rectification,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Aster: An attentional scene text recognizer with flexible rectification,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.252200Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T18:51:51.057681Z digest=sha256:38db886be7dcd55535e1d3473207ff892697f35b445a33804472dc84ae574059

Observation 5f53a680-4315-4f3f-b2fd-9430fae75fa3 · outbound

This paper cites Learning to read irregular text with attention mechanisms.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Learning to read irregular text with attention mechanisms

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.222611Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T18:51:51.068453Z digest=sha256:de3632b54d82a37f18613af3dbc6b8ad992a9a61a9ed259262aed80abee1711c

Observation d06b7ef9-7e6b-49b8-855b-4bbb1c8a0710 · outbound

This paper cites Attention-based extraction of structured information from street view imagery,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Attention-based extraction of structured information from street view imagery,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.190639Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T18:51:51.078112Z digest=sha256:79791c82c4d5663bf82c5bfe77ebfc6c587419a2e84442dca9e6ba036adb0392

Observation 68f78fcc-ae89-4967-80d1-844cb149945e · outbound

This paper cites On recognizing texts of arbitrary shapes with 2d self-attention,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition On recognizing texts of arbitrary shapes with 2d self-attention,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.902039Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T18:51:51.091409Z digest=sha256:4c45db324fe7ff2b6535ee5f7c9d2979efdcd9d4294bf310ad4f0615b8863a68

Observation 55dbeeb6-2eee-4546-a882-dfbeea03c7b6 · outbound

This paper cites Master: Multi-aspect non-local network for scene text recognition,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Master: Multi-aspect non-local network for scene text recognition,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.167233Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T18:51:51.099312Z digest=sha256:128ca0bfcad50c795583ce3bd2311d1f1d4ffe94cbfd050c4935f3cdb30ffc81

Observation 36aab0a8-24ed-44ac-b68e-46bf06759c6a · outbound

This paper cites Context-based contrastive learning for scene text recognition,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Context-based contrastive learning for scene text recognition,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.130710Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T18:51:51.110585Z digest=sha256:9236138ec9e9fd43c96936b4a25eab0b16cb5f4ae064e969500968a50e6f74c1

Observation 30d36c02-f4b6-46db-b23f-b3e5ebda92f5 · outbound

This paper cites What Do Self-Supervised Vision Transformers Learn?.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition What Do Self-Supervised Vision Transformers Learn?

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:51.115642Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:51.115642Z digest=sha256:0c1e133116300a03ef12d2214deb35e7330cf5ea3c5ac94b096bdbaa8d044f20

Observation 2c07420e-0c87-43e3-8e7a-ba146686b80e · outbound

This paper cites Masked Siamese ConvNets.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Masked Siamese ConvNets

Reference 43

Resolution
verified exact
local_arxiv, observed 2026-08-12T18:51:51.459743Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T18:51:51.131141Z digest=sha256:2f41ae7da1e3600b8d5b3db937fa373929246c00cee2f5fbcc8ce1eb9e02e389

Observation d946ce2f-def7-4683-b9ff-12f493d29f49 · outbound

This paper cites Convnext v2: Co-designing and scaling convnets with masked autoen- coders,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Convnext v2: Co-designing and scaling convnets with masked autoen- coders,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.099905Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T18:51:51.140411Z digest=sha256:c040ba8131aa36c84c651b653e59a10cdb69c32f67dee42e13b02c5ae3dc8588

Observation 3d00ae23-eaea-49d2-ba9d-22c07fd90ffa · outbound

This paper cites Scene text recognition using higher order language priors,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Scene text recognition using higher order language priors,

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:51.149908Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:51.149908Z digest=sha256:eda02af5fae16e714698e7bdbe04db0d95f97c81a829ec94a297fb226ca11bc1

Observation daa89420-6b8a-4705-9b91-a40edd7ea10a · outbound

This paper cites Icdar 2003 robust reading competitions: entries, results, and future directions,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Icdar 2003 robust reading competitions: entries, results, and future directions,

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.050339Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T18:51:51.158424Z digest=sha256:10bfd09dafbdc88ef73d33463187922687d50bd7c61e83807ea0c5fb605cd0df

Observation 3e6d85b9-1692-43e9-a2a4-341c139fd1cc · outbound

This paper cites Icdar 2013 robust reading competition,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Icdar 2013 robust reading competition,

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:51.165266Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:51.165266Z digest=sha256:1cad8d33b135554712772c1c1f248f0c2d1eecd5da8c531387f64872681aeb8b

Observation f7cfbb70-e9fa-40b3-b8fc-1a160fc60d81 · outbound

This paper cites End-to-end scene text recog- nition,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition End-to-end scene text recog- nition,

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:51.171251Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:51.171251Z digest=sha256:d01279d69d2879a37dee70e7ac34d4c4b8252fce74be2f93cea23809825f7d40

Observation d0ec2b24-e01f-451e-a75e-c357c32a4627 · outbound

This paper cites Icdar 2015 competition on robust reading,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Icdar 2015 competition on robust reading,

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:51.176641Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:51.176641Z digest=sha256:413a0ecd85d383407f9eb8298e364ccfaa7d34ffbefb553f1bb59810046b5696

Observation 709297c3-c80f-4f4d-85dd-92e3c2e0324b · outbound

This paper cites Recognizing text with perspective distortion in natural scenes,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Recognizing text with perspective distortion in natural scenes,

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:51.181284Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:51.181284Z digest=sha256:4de49ff5a3dc6b65909deb66d70fc3c988868e648488f00382b9cae2228819ea

Observation a5e9a834-5d93-4e84-a1b0-8f6ea61ef8eb · outbound

This paper cites A robust arbitrary text detection system for natural scene images,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition A robust arbitrary text detection system for natural scene images,

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:51.189723Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:51.189723Z digest=sha256:797d7dcfc9eaed2c19d626abf7d51ae20ff1d6640ddc5e0c3f31cd894c46e273

Observation 24933bd1-0dce-42e8-994c-af1d4374b14e · outbound

This paper cites COCO-Text: Dataset and Benchmark for Text Detection and Recognition in Natural Images.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition COCO-Text: Dataset and Benchmark for Text Detection and Recognition in Natural Images

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:51.194473Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:51.194473Z digest=sha256:461a1a9ffc73775635727ea5f59f0c4758c903e80ecf36e2d1f061284ab230ce

Observation 45cc2fd2-fd34-4f99-945c-0eb27d4d8b0f · outbound

This paper cites Curved scene text detection via transverse and longitudinal sequence connection,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Curved scene text detection via transverse and longitudinal sequence connection,

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:51.919960Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T18:51:51.201544Z digest=sha256:14bd6a015dfeed9893ac57a8742846d9693f47b490167549e248e02b296bb0ea

Observation 5bd2736b-1dfe-4b13-9ac6-07ee9568cf60 · outbound

This paper cites Total-text: A comprehensive dataset for scene text detection and recognition,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Total-text: A comprehensive dataset for scene text detection and recognition,

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:51.892724Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T18:51:51.208011Z digest=sha256:a2627fd4fe95d399bac773da88ac1135a3e95b89a0a68b2551ba13e251e312d7

Observation 711f8c5e-dace-4049-bcfb-c1af68d97e23 · outbound

This paper cites From two to one: A new scene text recognizer with visual language modeling network,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition From two to one: A new scene text recognizer with visual language modeling network,

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.924987Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T18:51:51.216055Z digest=sha256:00859b92b8b0ffa39bbb4d70d8dcc56de9b18149b92fa7b9116f415c73a430d3

Observation c48a0081-eca3-4332-ad4f-8b03272225dd · outbound

This paper cites Robust scene text recognition with automatic rectification,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Robust scene text recognition with automatic rectification,

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:51.862579Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T18:51:51.221565Z digest=sha256:bea40d9d22c597efe628784bf077e5765b3f5ac590eabbe41c214cc1a078b6ba

Observation 567b5ee9-14fe-446f-9024-14baf5d428e4 · outbound

This paper cites What is wrong with scene text recognition model comparisons? dataset and model analysis,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition What is wrong with scene text recognition model comparisons? dataset and model analysis,

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:51.838323Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T18:51:51.226112Z digest=sha256:79c250d06ec45613989b374be8adb256a8a67334aee49bed34de8d52280b5c77

Observation bb93bc87-90ea-4726-864b-ee8d92d23afd · outbound

This paper cites Synthetic Data and Artificial Neural Networks for Natural Scene Text Recognition.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Synthetic Data and Artificial Neural Networks for Natural Scene Text Recognition

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:51.233345Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:51.233345Z digest=sha256:ca27f9c7cf5f3de2c48d8e0563a1a4567fa7f744aeaa34bcbcf6de55c5754093

Observation 269c4479-5ac7-498e-a195-c818e760d4f0 · outbound

This paper cites Benchmarking Chinese Text Recognition: Datasets, Baselines, and an Empirical Study.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Benchmarking Chinese Text Recognition: Datasets, Baselines, and an Empirical Study

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:51.239927Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:51.239927Z digest=sha256:be894949348f7c6666f8202574f1c5aeb128bd3589ed4ee951396726433a814b

Observation 92f2fc5d-958a-44b0-93bd-e502b61e316d · outbound

This paper cites The iam-database: an english sentence database for offline handwriting recognition,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition The iam-database: an english sentence database for offline handwriting recognition,

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:51.817737Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T18:51:51.248254Z digest=sha256:29e3e78bd3243093b7201f935cb101086838a3df9a0504a2bfb85cc7be8cce0e

Observation c6b7fa5a-4e6d-477d-8e3f-187a58c3ad9f · outbound

This paper cites Cvl-database: An off-line database for writer retrieval, writer identification and word spotting,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Cvl-database: An off-line database for writer retrieval, writer identification and word spotting,

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:51.783072Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T18:51:51.254718Z digest=sha256:b5079f54dab574cb656577e69b6f1cd5d2b77ede34153d2d6986193ee59fd821

Observation e5059467-309a-4b32-9523-a682fb31f211 · outbound

This paper cites Stochastic neighbor embedding,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Stochastic neighbor embedding,

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:51.260140Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:51.260140Z digest=sha256:f95f481742b0931615f34bc5f48a7ad844a42fbf1a8e9361752b6951eb36ce12

Pith citing papers

No inbound Pith citation observations are available.