Pith. sign in

Paper Citation Record · LEDGER

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition

As of 23 August 2026, this Paper Citation Record lists 60 of 60 outbound references and 0 inbound Pith citation observations for arXiv:2411.11219.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.11219 v2

Coverage vector

measured 60 of 60 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T18:51:51.260140Z

measured 60 of 60 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

60 of 60 outbound references displayed

  • verified exact1
  • verified fuzzy31
  • unresolved28
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation cfcb1dca-9492-43bb-9f53-93c1fc5400a8 · outbound

This paper cites Unsupervised feature learning via non-parametric instance discrimination,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Unsupervised feature learning via non-parametric instance discrimination,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:53.038960Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-12T18:51:50.760300Z digest=sha256:56745c731d46f7d30efa3476875cf52fc065f0e3a85e84c5f8cd4635e459b8c8

Observation ed596de8-4d2b-42d5-bca6-f7660c7ef85f · outbound

This paper cites Representation Learning with Contrastive Predictive Coding.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Representation Learning with Contrastive Predictive Coding

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:50.769703Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:50.769703Z digest=sha256:803804b4784581b9fc4e51273d3ff9c088a56718ab0f24b4652df8fd74238800

Observation fcf23d59-0ff2-495a-ac06-a129da57c68a · outbound

This paper cites Momentum contrast for unsupervised visual representation learning,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Momentum contrast for unsupervised visual representation learning,

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:50.776195Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:50.776195Z digest=sha256:bfc965b83ba3738634c6d61648a9597235f7e7a2df94554e5b3b06e02f8d5942

Observation f58e6fd3-0700-4869-901f-bfd79e4cdb88 · outbound

This paper cites A simple framework for contrastive learning of visual representations,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition A simple framework for contrastive learning of visual representations,

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:50.783247Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:50.783247Z digest=sha256:a43dd639c3ece05010b2353f75502f1d5efcf36c7d9ad7208349a41afadcea6f

Observation 00ff5207-7c73-482f-b8e4-d43d0f45d396 · outbound

This paper cites BEiT: BERT Pre-Training of Image Transformers.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition BEiT: BERT Pre-Training of Image Transformers

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:50.789743Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:50.789743Z digest=sha256:bab64a9a82d416b8c767b65b069f0e5cb55d164b76b37fbcef8e0c800c602fb5

Observation d8f214a5-cd92-45f6-85c0-69b4b94695b8 · outbound

This paper cites Masked au- toencoders are scalable vision learners,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Masked au- toencoders are scalable vision learners,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:50.796041Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:50.796041Z digest=sha256:61df7123a5e215e1e24ee681e941299717dc6200ae5e1974a95bea869a8f5c0a

Observation e4dc2277-8f68-4774-87c0-54b6a897e8bc · outbound

This paper cites Simmim: A simple framework for masked image modeling,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Simmim: A simple framework for masked image modeling,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.963075Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-12T18:51:50.805358Z digest=sha256:3671b8f6966ff488ff202b5c70a559cc56e594b32ca498316e0678a393083f12

Observation b08f27d2-29f0-450f-ace9-59ed3504491d · outbound

This paper cites Read like humans: Autonomous, bidirectional and iterative language modeling for scene text recognition,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Read like humans: Autonomous, bidirectional and iterative language modeling for scene text recognition,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.946358Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-12T18:51:50.816191Z digest=sha256:ec1d884ae687a80d4cf9eb7ba46d88ae77cc4f6926f4aa52e6bd67a23d8aaa75

Observation c8ab4525-95a7-48b8-864a-f1e500388b9f · outbound

This paper cites Sequence-to-sequence contrastive learn- ing for text recognition,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Sequence-to-sequence contrastive learn- ing for text recognition,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:50.837653Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:50.837653Z digest=sha256:ef9430d5d93a2bfc10a992baddc0d12cd8d3306764be8386e34063285a5302da

Observation 88f6549d-a09e-4af1-a8f8-5efb4b2cafdb · outbound

This paper cites Perceiving stroke-semantic context: Hierarchical contrastive learning for robust scene text recognition,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Perceiving stroke-semantic context: Hierarchical contrastive learning for robust scene text recognition,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.841789Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-12T18:51:50.844333Z digest=sha256:d31a1ab4dbc228cbc97ce8205a2a239cc615d651987a4d2ab425fb03cb7a14e5

Observation 33351973-32fb-4f94-9b5f-471929891d37 · outbound

This paper cites Reading and writing: Discriminative and generative modeling for self-supervised text recognition,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Reading and writing: Discriminative and generative modeling for self-supervised text recognition,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:50.852070Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:50.852070Z digest=sha256:6c4884a3a43ea11fc9887ea4a58b9d2f99b1f0a363d65020606bf27a0d613026

Observation 7229e8de-c376-4324-8a17-e29d8fa895b1 · outbound

This paper cites An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:50.857745Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:50.857745Z digest=sha256:6a1a3cd0641a51348189c70220e22c6770a8ad2243c987bcf54c776cb2913f2f

Observation 3e913a1d-5220-4e8e-86fd-8c3c80b5d31f · outbound

This paper cites MaskOCR: Text Recognition with Masked Encoder-Decoder Pretraining.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition MaskOCR: Text Recognition with Masked Encoder-Decoder Pretraining

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:50.864066Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:50.864066Z digest=sha256:acdd6e6d4d148ed6866470ecec426940cdd3c43b3ec703fd85a5426617e952a8

Observation fc24f0a7-bc9b-4f8e-a57c-6ed45b9eefa6 · outbound

This paper cites Towards accurate scene text recognition with semantic reasoning networks,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Towards accurate scene text recognition with semantic reasoning networks,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.783768Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-12T18:51:50.870943Z digest=sha256:07255e4b0af116ea164331aebd500ec9a06b05c2149d5b026b2ce4faa4b50e50

Observation 964ec38a-f7e9-4439-994f-5dde280b75d9 · outbound

This paper cites Seed: Semantics enhanced encoder-decoder framework for scene text recognition,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Seed: Semantics enhanced encoder-decoder framework for scene text recognition,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.754387Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-12T18:51:50.877425Z digest=sha256:846dcfb110c17056e0b3e661d5f43b8892f7a17349ed00865fb72f2e79595b7f

Observation a3838276-dd2e-420c-9379-821d21227ddf · outbound

This paper cites On vocabulary reliance in scene text recognition,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition On vocabulary reliance in scene text recognition,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.717253Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-12T18:51:50.885900Z digest=sha256:c256eb99a4351d1f0dd3e05f20f61e26bcd6405e59d4d87e2849aca7ac86ee69

Observation 70b6e185-263e-4787-9fa5-a6a20773d936 · outbound

This paper cites Ressl: Relational self-supervised learning with weak augmentation,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Ressl: Relational self-supervised learning with weak augmentation,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.676841Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-12T18:51:50.892516Z digest=sha256:10514867bd468cbb6969cc72967c0c971f0656df360226c2a7faefa9815446ca

Observation a0266b4a-835a-475c-959f-757b35f52242 · outbound

This paper cites Synthetic data for text localisation in natural images,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Synthetic data for text localisation in natural images,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:50.899694Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:50.899694Z digest=sha256:be90a2492e422ff06bf090f26ccf570834e3d03e2674e655a0913a59aaac1205

Observation 2f0c00b6-91b6-4bfd-98bc-4ef005f9cb42 · outbound

This paper cites Rethink- ing text segmentation: A novel dataset and a text-specific refinement approach,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Rethink- ing text segmentation: A novel dataset and a text-specific refinement approach,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.623318Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-12T18:51:50.908188Z digest=sha256:ef901b94d3ced978fc851b05f677bc71f84511c7f34d747f12f69e3f5d88582b

Observation 1c9b8423-121c-463b-8ab5-362d7b5c435f · outbound

This paper cites Relational contrastive learning for scene text recognition,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Relational contrastive learning for scene text recognition,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.598854Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-12T18:51:50.914335Z digest=sha256:03489b1973d3aebc4cde861cf8dea250b8b34b2b12df299508297a31ed67d918

Observation 60134332-9388-44ca-a69a-d6ea42533e25 · outbound

This paper cites Unsupervised learning of visual features by contrasting cluster assign- ments,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Unsupervised learning of visual features by contrasting cluster assign- ments,

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:50.923630Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:50.923630Z digest=sha256:edd51673487c5677cf2de857aa61d42692f52e25c96b1d2a97c4ef521a80eef5

Observation 183fe724-b5f8-4c46-a267-ca6f8f487dd6 · outbound

This paper cites Bootstrap your own latent-a new approach to self-supervised learning,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Bootstrap your own latent-a new approach to self-supervised learning,

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:50.930894Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:50.930894Z digest=sha256:452d1a625d38a32546f2350f7130989fae1481fe9e79aa5501ffb9b03562f9c1

Observation 4a006d29-137f-4795-99a6-e0c851631114 · outbound

This paper cites Exploring simple siamese representation learning,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Exploring simple siamese representation learning,

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:50.940508Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:50.940508Z digest=sha256:d6e59e444356479f7f8fa3660d70f7f12d065d600d8273b3e5de6d061142b06b

Observation 27ea1a1d-6ecd-45c2-b8f2-0073a227c032 · outbound

This paper cites Emerging properties in self-supervised vision transformers,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Emerging properties in self-supervised vision transformers,

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:50.945778Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:50.945778Z digest=sha256:c087d6ee7ef4eaedb53a356068be327b9a6c5201d83c06423e1ca5dd6b8d6431

Observation b91e51fb-4ca1-4577-9ff7-223bb5a18fae · outbound

This paper cites iBOT: Image BERT Pre-Training with Online Tokenizer.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition iBOT: Image BERT Pre-Training with Online Tokenizer

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:50.951689Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:50.951689Z digest=sha256:63caeca0f644f96e0a2549f65e2c5274a61b552479f295c22da911669979d039

Observation 01b76be6-5c32-471f-a601-c6c1fe7fa17e · outbound

This paper cites DINOv2: Learning Robust Visual Features without Supervision.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition DINOv2: Learning Robust Visual Features without Supervision

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:50.962274Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:50.962274Z digest=sha256:7338da24314c0dcf941080d400a59ac3c087eb8ff9936bf5081821dccbba7dba

Observation d26d25f2-e35c-457a-99d7-0e3c0d0c14f2 · outbound

This paper cites Contrastive learning with stronger augmenta- tions,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Contrastive learning with stronger augmenta- tions,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.477016Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-12T18:51:50.969284Z digest=sha256:695a529b895c4d9eb6f11206b6d322a221ae0a9b035c45b771bbdb58387ce99b

Observation a3086599-781f-4b39-b4fd-f2ccd2f2e188 · outbound

This paper cites What if we only use real datasets for scene text recognition? toward scene text recognition with fewer labels,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition What if we only use real datasets for scene text recognition? toward scene text recognition with fewer labels,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.440978Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-12T18:51:50.978025Z digest=sha256:32f687eac111f5baa1cdda6f049eebf23da0ed8b66a6fa7bce4f3af24625923a

Observation ddbd37fd-da62-400b-9b57-70ae7bf42747 · outbound

This paper cites Siman: Exploring self-supervised rep- resentation learning of scene text via similarity-aware normalization,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Siman: Exploring self-supervised rep- resentation learning of scene text via similarity-aware normalization,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.404622Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-12T18:51:50.987664Z digest=sha256:2e86a986ad1028e20e719fc8885d11445fbdb765da9447993a8863c68fe0f4f5

Observation 93e80a5d-fbbe-4a0f-8e1d-f1c8f777b9b4 · outbound

This paper cites Self- supervised character-to-character distillation for text recognition,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Self- supervised character-to-character distillation for text recognition,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.377343Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-12T18:51:50.996746Z digest=sha256:2f31eec013f0289c902b972aac8ca2b07404189737c3d8b5a80fc5a4db8cb795

Observation b2444aa3-aacf-4583-ab26-7478617f1f99 · outbound

This paper cites Reading scene text in deep convolutional sequences,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Reading scene text in deep convolutional sequences,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.350919Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-12T18:51:51.007434Z digest=sha256:3a901225ba1a8e3a668b6ceee2e63886bfe02579d713e7f7a777255d35c4e3bf

Observation 70bd8651-a4a6-4636-800a-56f7253b28c1 · outbound

This paper cites Accurate recognition of words in scenes without char- acter segmentation using recurrent neural network,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Accurate recognition of words in scenes without char- acter segmentation using recurrent neural network,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.322809Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-12T18:51:51.039572Z digest=sha256:875afec1b0f914755ff45b6cfb710178d4295d6bdb8a84f453a4c8961d62ecf3

Observation b8a625bb-108a-48a2-bdbb-333c455433da · outbound

This paper cites An end-to-end trainable neural network for image-based sequence recognition and its application to scene text recognition,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition An end-to-end trainable neural network for image-based sequence recognition and its application to scene text recognition,

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:51.049965Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:51.049965Z digest=sha256:ed8b6ea38b43777012f5132561e321a4fc0b20dc305f0dd1aca50fad6f433b0b

Observation e4f90b37-65c5-4c45-bb1b-5e6f01e50b3a · outbound

This paper cites Aster: An attentional scene text recognizer with flexible rectification,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Aster: An attentional scene text recognizer with flexible rectification,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.252200Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-12T18:51:51.057681Z digest=sha256:b877cc6fe394f9a0b45fdf146560c986229581319a626385700897e107043c03

Observation 5f53a680-4315-4f3f-b2fd-9430fae75fa3 · outbound

This paper cites Learning to read irregular text with attention mechanisms.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Learning to read irregular text with attention mechanisms

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.222611Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-12T18:51:51.068453Z digest=sha256:f5ade228eee12d01f1206825db8a37390df695d02a66146c27c995f0fe447bb9

Observation d06b7ef9-7e6b-49b8-855b-4bbb1c8a0710 · outbound

This paper cites Attention-based extraction of structured information from street view imagery,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Attention-based extraction of structured information from street view imagery,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.190639Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-12T18:51:51.078112Z digest=sha256:628d696c2786ac4e7ca776be31650a9428e7aeee1840feed75f82e03ade384f4

Observation 68f78fcc-ae89-4967-80d1-844cb149945e · outbound

This paper cites On recognizing texts of arbitrary shapes with 2d self-attention,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition On recognizing texts of arbitrary shapes with 2d self-attention,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.902039Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-12T18:51:51.091409Z digest=sha256:2381e2ab5c446b11331008d8c1c72a9ca529b44e19c58f16f13aab5b23613d05

Observation 55dbeeb6-2eee-4546-a882-dfbeea03c7b6 · outbound

This paper cites Master: Multi-aspect non-local network for scene text recognition,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Master: Multi-aspect non-local network for scene text recognition,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.167233Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-12T18:51:51.099312Z digest=sha256:f3b908f2cff430e938f0159655aa46ce71778c9075fca69f83615d702f9378fb

Observation 36aab0a8-24ed-44ac-b68e-46bf06759c6a · outbound

This paper cites Context-based contrastive learning for scene text recognition,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Context-based contrastive learning for scene text recognition,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.130710Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-12T18:51:51.110585Z digest=sha256:2a8747f29e5eaaca1177b451f9169d451e027cee6f0b9e27dab81d1aa6d1c624

Observation 30d36c02-f4b6-46db-b23f-b3e5ebda92f5 · outbound

This paper cites What Do Self-Supervised Vision Transformers Learn?.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition What Do Self-Supervised Vision Transformers Learn?

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:51.115642Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:51.115642Z digest=sha256:fdd37d14e76210422394678ceca8d09fa394a8e1e3247139d368d772702e9766

Observation 2c07420e-0c87-43e3-8e7a-ba146686b80e · outbound

This paper cites Masked Siamese ConvNets.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Masked Siamese ConvNets

Reference 43

Resolution
verified exact
local_arxiv, observed 2026-08-12T18:51:51.459743Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-12T18:51:51.131141Z digest=sha256:b1af690f7d455e45826fc475077405aa11fc79f60e597024c3a966969e28e2c1

Observation d946ce2f-def7-4683-b9ff-12f493d29f49 · outbound

This paper cites Convnext v2: Co-designing and scaling convnets with masked autoen- coders,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Convnext v2: Co-designing and scaling convnets with masked autoen- coders,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.099905Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-12T18:51:51.140411Z digest=sha256:f5ca13d479f2a5cafac076a4a389755d58904f9f16eac6ee920c056d22b343fa

Observation 3d00ae23-eaea-49d2-ba9d-22c07fd90ffa · outbound

This paper cites Scene text recognition using higher order language priors,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Scene text recognition using higher order language priors,

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:51.149908Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:51.149908Z digest=sha256:943191ccb44918408ec29894ce5ee84a7f3677ba21568d3656dfb49b17af7347

Observation daa89420-6b8a-4705-9b91-a40edd7ea10a · outbound

This paper cites Icdar 2003 robust reading competitions: entries, results, and future directions,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Icdar 2003 robust reading competitions: entries, results, and future directions,

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.050339Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-12T18:51:51.158424Z digest=sha256:5cdc956fdb9f2a829189fb1cf692f8a004f25c96a7b9417cf091a7aadc0374b3

Observation 3e6d85b9-1692-43e9-a2a4-341c139fd1cc · outbound

This paper cites Icdar 2013 robust reading competition,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Icdar 2013 robust reading competition,

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:51.165266Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:51.165266Z digest=sha256:89a9475e5bd0ed7c20525f6fab7b9431615ff3c6be0bcec0c1a8de6a58896a90

Observation f7cfbb70-e9fa-40b3-b8fc-1a160fc60d81 · outbound

This paper cites End-to-end scene text recog- nition,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition End-to-end scene text recog- nition,

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:51.171251Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:51.171251Z digest=sha256:c82369f84a24031dfa45323b067d68329648912656b013cda89937fd5987a997

Observation d0ec2b24-e01f-451e-a75e-c357c32a4627 · outbound

This paper cites Icdar 2015 competition on robust reading,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Icdar 2015 competition on robust reading,

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:51.176641Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:51.176641Z digest=sha256:28b1903b096e5e2552c9b2efe5482c5489bf97c161b1a829e4e8e01cf0e5b38a

Observation 709297c3-c80f-4f4d-85dd-92e3c2e0324b · outbound

This paper cites Recognizing text with perspective distortion in natural scenes,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Recognizing text with perspective distortion in natural scenes,

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:51.181284Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:51.181284Z digest=sha256:25755e39e88e7663db61223db62fc3fab43b5ebb56e321cc21ef7f9236abfb27

Observation a5e9a834-5d93-4e84-a1b0-8f6ea61ef8eb · outbound

This paper cites A robust arbitrary text detection system for natural scene images,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition A robust arbitrary text detection system for natural scene images,

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:51.189723Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:51.189723Z digest=sha256:454f2cba23228f1949604a561b0e3dc3e5eea0bc35aed275c449d3088352fbe3

Observation 24933bd1-0dce-42e8-994c-af1d4374b14e · outbound

This paper cites COCO-Text: Dataset and Benchmark for Text Detection and Recognition in Natural Images.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition COCO-Text: Dataset and Benchmark for Text Detection and Recognition in Natural Images

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:51.194473Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:51.194473Z digest=sha256:f23e3601d0b91fa039fc491bcb2acf104a4679d2be371d68295f15f7d839f57c

Observation 45cc2fd2-fd34-4f99-945c-0eb27d4d8b0f · outbound

This paper cites Curved scene text detection via transverse and longitudinal sequence connection,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Curved scene text detection via transverse and longitudinal sequence connection,

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:51.919960Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-12T18:51:51.201544Z digest=sha256:565c62be10b817df59ff12c653d123b60f1aed8bbec989734579452b2031fa88

Observation 5bd2736b-1dfe-4b13-9ac6-07ee9568cf60 · outbound

This paper cites Total-text: A comprehensive dataset for scene text detection and recognition,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Total-text: A comprehensive dataset for scene text detection and recognition,

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:51.892724Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-12T18:51:51.208011Z digest=sha256:d466281d0e05e20790a815f1946c5548270458fa3c3cdae438578bdecd44a57e

Observation 711f8c5e-dace-4049-bcfb-c1af68d97e23 · outbound

This paper cites From two to one: A new scene text recognizer with visual language modeling network,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition From two to one: A new scene text recognizer with visual language modeling network,

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.924987Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-12T18:51:51.216055Z digest=sha256:cd53d2ce29061114942629019e2f599a05eebcc1445718103f18864b09509529

Observation c48a0081-eca3-4332-ad4f-8b03272225dd · outbound

This paper cites Robust scene text recognition with automatic rectification,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Robust scene text recognition with automatic rectification,

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:51.862579Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-12T18:51:51.221565Z digest=sha256:3b437bb5172a638a0edb3f76eb9e0e3af3bb91d996b99860234e10acdee3d939

Observation 567b5ee9-14fe-446f-9024-14baf5d428e4 · outbound

This paper cites What is wrong with scene text recognition model comparisons? dataset and model analysis,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition What is wrong with scene text recognition model comparisons? dataset and model analysis,

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:51.838323Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-12T18:51:51.226112Z digest=sha256:8f08a5adaeb7ba458fcb6b56e35f84ef3445d5ff951572e956c26807dbccfb6f

Observation bb93bc87-90ea-4726-864b-ee8d92d23afd · outbound

This paper cites Synthetic Data and Artificial Neural Networks for Natural Scene Text Recognition.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Synthetic Data and Artificial Neural Networks for Natural Scene Text Recognition

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:51.233345Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:51.233345Z digest=sha256:6f9a9699b9f92e41867a5198008e796c42577be5fe6df73636cdeddb03f2ff2b

Observation 269c4479-5ac7-498e-a195-c818e760d4f0 · outbound

This paper cites Benchmarking Chinese Text Recognition: Datasets, Baselines, and an Empirical Study.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Benchmarking Chinese Text Recognition: Datasets, Baselines, and an Empirical Study

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:51.239927Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:51.239927Z digest=sha256:b5f971831d47699a568237327771f416a85438d0a0f3e1c9e5f3e3cc14589799

Observation 92f2fc5d-958a-44b0-93bd-e502b61e316d · outbound

This paper cites The iam-database: an english sentence database for offline handwriting recognition,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition The iam-database: an english sentence database for offline handwriting recognition,

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:51.817737Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-12T18:51:51.248254Z digest=sha256:1c2127fbd865c146af6588f1876c59a1662661310a55028f89d73a0eaf1ae211

Observation c6b7fa5a-4e6d-477d-8e3f-187a58c3ad9f · outbound

This paper cites Cvl-database: An off-line database for writer retrieval, writer identification and word spotting,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Cvl-database: An off-line database for writer retrieval, writer identification and word spotting,

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:51.783072Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-12T18:51:51.254718Z digest=sha256:077934cf757ed3963074ee0cd6db1305851643ba1f8f24ef8de674b4451a690d

Observation e5059467-309a-4b32-9523-a682fb31f211 · outbound

This paper cites Stochastic neighbor embedding,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Stochastic neighbor embedding,

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:51.260140Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:51.260140Z digest=sha256:db0c9df894ed591c0138bc96aed8e8ffaa942afe44fc502be5bdaedbe4b8fbe3

Pith citing papers

No inbound Pith citation observations are available.