Pith. sign in

Paper Citation Record · LEDGER

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition

As of 17 August 2026, this Paper Citation Record lists 60 of 60 outbound references and 0 inbound Pith citation observations for arXiv:2411.11219.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.11219 v2

Coverage vector

measured 60 of 60 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T18:51:51.260140Z

measured 60 of 60 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

60 of 60 outbound references displayed

  • verified exact1
  • verified fuzzy31
  • unresolved28
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation cfcb1dca-9492-43bb-9f53-93c1fc5400a8 · outbound

This paper cites Unsupervised feature learning via non-parametric instance discrimination,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Unsupervised feature learning via non-parametric instance discrimination,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:53.038960Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T18:51:50.760300Z digest=sha256:9b2affd0a05e8ebfba69df016bf8373279505e7d2c8de29f03621766b44bc603

Observation ed596de8-4d2b-42d5-bca6-f7660c7ef85f · outbound

This paper cites Representation Learning with Contrastive Predictive Coding.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Representation Learning with Contrastive Predictive Coding

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:50.769703Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:50.769703Z digest=sha256:869eacb13d9f9ec8bc6d0b0891a5384643fff9833906127de20e65cafde382bd

Observation fcf23d59-0ff2-495a-ac06-a129da57c68a · outbound

This paper cites Momentum contrast for unsupervised visual representation learning,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Momentum contrast for unsupervised visual representation learning,

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:50.776195Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:50.776195Z digest=sha256:021f4468b889b0559b9a4e34d88e5881c323ee7dbe22ab41d30874ea1ef61ab5

Observation f58e6fd3-0700-4869-901f-bfd79e4cdb88 · outbound

This paper cites A simple framework for contrastive learning of visual representations,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition A simple framework for contrastive learning of visual representations,

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:50.783247Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:50.783247Z digest=sha256:cb7e3febb3a0ec1ca89615cd682fb031cf4cf2c562d309cd823d171f33d852b2

Observation 00ff5207-7c73-482f-b8e4-d43d0f45d396 · outbound

This paper cites BEiT: BERT Pre-Training of Image Transformers.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition BEiT: BERT Pre-Training of Image Transformers

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:50.789743Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:50.789743Z digest=sha256:697ccf92fa11c3c9114b17174dda109ff2314cda712397485214c399076d008f

Observation d8f214a5-cd92-45f6-85c0-69b4b94695b8 · outbound

This paper cites Masked au- toencoders are scalable vision learners,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Masked au- toencoders are scalable vision learners,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:50.796041Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:50.796041Z digest=sha256:fe6c6de81097b4c7e4d714fd9e60f9832f565ca78395e5bb4a06430e71613863

Observation e4dc2277-8f68-4774-87c0-54b6a897e8bc · outbound

This paper cites Simmim: A simple framework for masked image modeling,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Simmim: A simple framework for masked image modeling,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.963075Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T18:51:50.805358Z digest=sha256:1e7fef79c8a3c6ef43ee7cee210cc0831ee96b3e09bbef84f0549f15c2ec9810

Observation b08f27d2-29f0-450f-ace9-59ed3504491d · outbound

This paper cites Read like humans: Autonomous, bidirectional and iterative language modeling for scene text recognition,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Read like humans: Autonomous, bidirectional and iterative language modeling for scene text recognition,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.946358Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T18:51:50.816191Z digest=sha256:52ecd8bc9fe7a1ea9ab1c354fb3c85a0440a37ea89b1e57d2eeb93c55210e2da

Observation c8ab4525-95a7-48b8-864a-f1e500388b9f · outbound

This paper cites Sequence-to-sequence contrastive learn- ing for text recognition,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Sequence-to-sequence contrastive learn- ing for text recognition,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:50.837653Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:50.837653Z digest=sha256:b81b6e8a791e3a31946bd61dd33f31ff6d5bfe109a97c6b6d3d49788998405e7

Observation 88f6549d-a09e-4af1-a8f8-5efb4b2cafdb · outbound

This paper cites Perceiving stroke-semantic context: Hierarchical contrastive learning for robust scene text recognition,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Perceiving stroke-semantic context: Hierarchical contrastive learning for robust scene text recognition,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.841789Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T18:51:50.844333Z digest=sha256:6d6780660e4daaf0b1212681da00cae3e64df5be1b1162676dd2e17c4b01a9af

Observation 33351973-32fb-4f94-9b5f-471929891d37 · outbound

This paper cites Reading and writing: Discriminative and generative modeling for self-supervised text recognition,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Reading and writing: Discriminative and generative modeling for self-supervised text recognition,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:50.852070Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:50.852070Z digest=sha256:ac305cb71b747641738b1f26487fb0f825a737d483ce0d019707dab3488bbd5e

Observation 7229e8de-c376-4324-8a17-e29d8fa895b1 · outbound

This paper cites An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:50.857745Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:50.857745Z digest=sha256:591611eab62e45897329d8ef76efd8056a0c760bd5b98f458648d0fbc7f2ee2d

Observation 3e913a1d-5220-4e8e-86fd-8c3c80b5d31f · outbound

This paper cites MaskOCR: Text Recognition with Masked Encoder-Decoder Pretraining.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition MaskOCR: Text Recognition with Masked Encoder-Decoder Pretraining

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:50.864066Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:50.864066Z digest=sha256:3ed6e6387a3b385a55d6d6c61f978548684de6d83a9aa41e74677bbbc0704c2b

Observation fc24f0a7-bc9b-4f8e-a57c-6ed45b9eefa6 · outbound

This paper cites Towards accurate scene text recognition with semantic reasoning networks,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Towards accurate scene text recognition with semantic reasoning networks,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.783768Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T18:51:50.870943Z digest=sha256:b25207d337c859240dd73ebc163955c24a8862c216e54b78dd871b096a557b87

Observation 964ec38a-f7e9-4439-994f-5dde280b75d9 · outbound

This paper cites Seed: Semantics enhanced encoder-decoder framework for scene text recognition,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Seed: Semantics enhanced encoder-decoder framework for scene text recognition,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.754387Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T18:51:50.877425Z digest=sha256:8e34df894bb4e0028c3c25aea9f8be57f2c638cd6ec2e9926d3282e1872477e7

Observation a3838276-dd2e-420c-9379-821d21227ddf · outbound

This paper cites On vocabulary reliance in scene text recognition,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition On vocabulary reliance in scene text recognition,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.717253Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T18:51:50.885900Z digest=sha256:4c8628129bb1df978c067726db33686e56ebc04bc92b830350b217fc2d64eeec

Observation 70b6e185-263e-4787-9fa5-a6a20773d936 · outbound

This paper cites Ressl: Relational self-supervised learning with weak augmentation,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Ressl: Relational self-supervised learning with weak augmentation,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.676841Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T18:51:50.892516Z digest=sha256:30f99af8b2d24e58d662c4f69b6d5997bba0299c52240369c3aff2f416614440

Observation a0266b4a-835a-475c-959f-757b35f52242 · outbound

This paper cites Synthetic data for text localisation in natural images,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Synthetic data for text localisation in natural images,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:50.899694Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:50.899694Z digest=sha256:3d83c30f6ac51509b2c4d08446ddb82e9804879d87c1d7ee151a332e3ddcd700

Observation 2f0c00b6-91b6-4bfd-98bc-4ef005f9cb42 · outbound

This paper cites Rethink- ing text segmentation: A novel dataset and a text-specific refinement approach,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Rethink- ing text segmentation: A novel dataset and a text-specific refinement approach,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.623318Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T18:51:50.908188Z digest=sha256:8dcd7170ad58b0249955777cc83bcf98d1fce5e242f0e146d38e4e0f42d772e9

Observation 1c9b8423-121c-463b-8ab5-362d7b5c435f · outbound

This paper cites Relational contrastive learning for scene text recognition,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Relational contrastive learning for scene text recognition,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.598854Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T18:51:50.914335Z digest=sha256:3de1f23621e7d4db59b33b01b6c5e126447f3b28d3cf84ddd512b4d4d72d3564

Observation 60134332-9388-44ca-a69a-d6ea42533e25 · outbound

This paper cites Unsupervised learning of visual features by contrasting cluster assign- ments,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Unsupervised learning of visual features by contrasting cluster assign- ments,

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:50.923630Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:50.923630Z digest=sha256:0805f8884446105ab45f1cba9b86ce2b9c997f85afdc6bf411c391e86d712a13

Observation 183fe724-b5f8-4c46-a267-ca6f8f487dd6 · outbound

This paper cites Bootstrap your own latent-a new approach to self-supervised learning,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Bootstrap your own latent-a new approach to self-supervised learning,

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:50.930894Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:50.930894Z digest=sha256:39966aa4f435bd18aa06a9c0b4123fb252f966b7da1e92794ff40fe8e5587245

Observation 4a006d29-137f-4795-99a6-e0c851631114 · outbound

This paper cites Exploring simple siamese representation learning,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Exploring simple siamese representation learning,

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:50.940508Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:50.940508Z digest=sha256:69b1610f5ef2d44440daab3b55363e16f3201e17b337738c762e925e666a8b3d

Observation 27ea1a1d-6ecd-45c2-b8f2-0073a227c032 · outbound

This paper cites Emerging properties in self-supervised vision transformers,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Emerging properties in self-supervised vision transformers,

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:50.945778Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:50.945778Z digest=sha256:0d8d43760c94d4af3254bb3aeb73f3449eb2a2d211e7fb12788e113c1c5ec9a3

Observation b91e51fb-4ca1-4577-9ff7-223bb5a18fae · outbound

This paper cites iBOT: Image BERT Pre-Training with Online Tokenizer.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition iBOT: Image BERT Pre-Training with Online Tokenizer

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:50.951689Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:50.951689Z digest=sha256:7da6db4fb18369fc7e146d2fe3129a0c8e65cb935aa6cb81b551a4954eee4a47

Observation 01b76be6-5c32-471f-a601-c6c1fe7fa17e · outbound

This paper cites DINOv2: Learning Robust Visual Features without Supervision.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition DINOv2: Learning Robust Visual Features without Supervision

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:50.962274Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:50.962274Z digest=sha256:f47e685afda5e4f5204a6fb6efa1f3b5d6ef21374b4be33973a9bd393d524c59

Observation d26d25f2-e35c-457a-99d7-0e3c0d0c14f2 · outbound

This paper cites Contrastive learning with stronger augmenta- tions,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Contrastive learning with stronger augmenta- tions,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.477016Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T18:51:50.969284Z digest=sha256:3c35d45b33206d1ff75de1be69fabfffbaa61c6263ca081c2efedd8f8498ff47

Observation a3086599-781f-4b39-b4fd-f2ccd2f2e188 · outbound

This paper cites What if we only use real datasets for scene text recognition? toward scene text recognition with fewer labels,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition What if we only use real datasets for scene text recognition? toward scene text recognition with fewer labels,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.440978Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T18:51:50.978025Z digest=sha256:999c92df8e072026a7052f5ed30c0aecae8d6695ee5ca172a2ed726dfdd35d49

Observation ddbd37fd-da62-400b-9b57-70ae7bf42747 · outbound

This paper cites Siman: Exploring self-supervised rep- resentation learning of scene text via similarity-aware normalization,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Siman: Exploring self-supervised rep- resentation learning of scene text via similarity-aware normalization,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.404622Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T18:51:50.987664Z digest=sha256:76c01cc4b44829e157453a30f70693943cfc362596b798428ed62cf3a3f34e15

Observation 93e80a5d-fbbe-4a0f-8e1d-f1c8f777b9b4 · outbound

This paper cites Self- supervised character-to-character distillation for text recognition,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Self- supervised character-to-character distillation for text recognition,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.377343Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T18:51:50.996746Z digest=sha256:4b7066b2c275c0ede4df1f10e5152690c2e21b88eb7c85fcca0c35d70e7ce70d

Observation b2444aa3-aacf-4583-ab26-7478617f1f99 · outbound

This paper cites Reading scene text in deep convolutional sequences,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Reading scene text in deep convolutional sequences,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.350919Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T18:51:51.007434Z digest=sha256:312ac12aeb661d223f77544aad72861b9bd4fefcf7b463fbe5cdcb95bc264c36

Observation 70bd8651-a4a6-4636-800a-56f7253b28c1 · outbound

This paper cites Accurate recognition of words in scenes without char- acter segmentation using recurrent neural network,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Accurate recognition of words in scenes without char- acter segmentation using recurrent neural network,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.322809Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T18:51:51.039572Z digest=sha256:8483d7dfbbf62ce2c8ff075c77cf3d3a67b9d28d81ed6fd92286feb046ce9e59

Observation b8a625bb-108a-48a2-bdbb-333c455433da · outbound

This paper cites An end-to-end trainable neural network for image-based sequence recognition and its application to scene text recognition,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition An end-to-end trainable neural network for image-based sequence recognition and its application to scene text recognition,

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:51.049965Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:51.049965Z digest=sha256:35812bd4ea9534eecced26d97a643740ef477db364f9da395e74a6964c75104b

Observation e4f90b37-65c5-4c45-bb1b-5e6f01e50b3a · outbound

This paper cites Aster: An attentional scene text recognizer with flexible rectification,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Aster: An attentional scene text recognizer with flexible rectification,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.252200Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T18:51:51.057681Z digest=sha256:4aec81f8778b3599b478e20e821bec0050420ace743f3b3a4e240b85d5330f62

Observation 5f53a680-4315-4f3f-b2fd-9430fae75fa3 · outbound

This paper cites Learning to read irregular text with attention mechanisms.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Learning to read irregular text with attention mechanisms

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.222611Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T18:51:51.068453Z digest=sha256:a7b0d90cd3059a07374c2064c6fd64503850957bced21ba1dac65c9083d3d852

Observation d06b7ef9-7e6b-49b8-855b-4bbb1c8a0710 · outbound

This paper cites Attention-based extraction of structured information from street view imagery,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Attention-based extraction of structured information from street view imagery,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.190639Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T18:51:51.078112Z digest=sha256:49f953b201d12a13a675643a1a2df7f18a3b002f376523158b6861e335ea4ade

Observation 68f78fcc-ae89-4967-80d1-844cb149945e · outbound

This paper cites On recognizing texts of arbitrary shapes with 2d self-attention,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition On recognizing texts of arbitrary shapes with 2d self-attention,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.902039Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T18:51:51.091409Z digest=sha256:964d4e0ba379da8cba96ba53fb17b1c95b12cdf2c3df608227ac71d39acddd05

Observation 55dbeeb6-2eee-4546-a882-dfbeea03c7b6 · outbound

This paper cites Master: Multi-aspect non-local network for scene text recognition,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Master: Multi-aspect non-local network for scene text recognition,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.167233Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T18:51:51.099312Z digest=sha256:3d278d05d8826f07360679b4c5d12ebb76f08add3cbc62e83981c0591d826886

Observation 36aab0a8-24ed-44ac-b68e-46bf06759c6a · outbound

This paper cites Context-based contrastive learning for scene text recognition,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Context-based contrastive learning for scene text recognition,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.130710Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T18:51:51.110585Z digest=sha256:5065ed481a8f244dfc27913672a1d3249867f0e24f63113de15f7981945ab287

Observation 30d36c02-f4b6-46db-b23f-b3e5ebda92f5 · outbound

This paper cites What Do Self-Supervised Vision Transformers Learn?.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition What Do Self-Supervised Vision Transformers Learn?

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:51.115642Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:51.115642Z digest=sha256:08d352de1ef6269a9c82b2572431c6051281decd7490a11e13d2d07a8d86bc1e

Observation 2c07420e-0c87-43e3-8e7a-ba146686b80e · outbound

This paper cites Masked Siamese ConvNets.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Masked Siamese ConvNets

Reference 43

Resolution
verified exact
local_arxiv, observed 2026-08-12T18:51:51.459743Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T18:51:51.131141Z digest=sha256:317aace3e8bad36e9bcba23a8c0bb6420c5294c7910a0a5fed5be1549ebf7937

Observation d946ce2f-def7-4683-b9ff-12f493d29f49 · outbound

This paper cites Convnext v2: Co-designing and scaling convnets with masked autoen- coders,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Convnext v2: Co-designing and scaling convnets with masked autoen- coders,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.099905Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T18:51:51.140411Z digest=sha256:06c7b8e97762f56d7e0090ebdd1c90f65567757c6e79c8da0234214404123c86

Observation 3d00ae23-eaea-49d2-ba9d-22c07fd90ffa · outbound

This paper cites Scene text recognition using higher order language priors,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Scene text recognition using higher order language priors,

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:51.149908Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:51.149908Z digest=sha256:6b5216746bcfaf22369626ade48136e00ab620ae4aca4570160ac65bf19d7036

Observation daa89420-6b8a-4705-9b91-a40edd7ea10a · outbound

This paper cites Icdar 2003 robust reading competitions: entries, results, and future directions,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Icdar 2003 robust reading competitions: entries, results, and future directions,

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.050339Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T18:51:51.158424Z digest=sha256:da94e475dbcadce602284ccb47435719def2b14c9213cdf2973f35ae5f91bfdb

Observation 3e6d85b9-1692-43e9-a2a4-341c139fd1cc · outbound

This paper cites Icdar 2013 robust reading competition,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Icdar 2013 robust reading competition,

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:51.165266Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:51.165266Z digest=sha256:4c41e510ce8d13bd050c2a44c92d774ff1db7d44f40130e34b5fad481a13e775

Observation f7cfbb70-e9fa-40b3-b8fc-1a160fc60d81 · outbound

This paper cites End-to-end scene text recog- nition,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition End-to-end scene text recog- nition,

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:51.171251Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:51.171251Z digest=sha256:fe5e41b03da6fe2cc65d5d2407c4de619a42cabd0b45f1f0693a62885eb7eccb

Observation d0ec2b24-e01f-451e-a75e-c357c32a4627 · outbound

This paper cites Icdar 2015 competition on robust reading,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Icdar 2015 competition on robust reading,

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:51.176641Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:51.176641Z digest=sha256:f0d3f92df5c0034df06c3b3a4b8a9e9819ee9592e963a1afcf635f4b4187e400

Observation 709297c3-c80f-4f4d-85dd-92e3c2e0324b · outbound

This paper cites Recognizing text with perspective distortion in natural scenes,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Recognizing text with perspective distortion in natural scenes,

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:51.181284Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:51.181284Z digest=sha256:a2dc479fb3dd2bf307757c8e02246fc1822adb548c58d999e657633b12439858

Observation a5e9a834-5d93-4e84-a1b0-8f6ea61ef8eb · outbound

This paper cites A robust arbitrary text detection system for natural scene images,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition A robust arbitrary text detection system for natural scene images,

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:51.189723Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:51.189723Z digest=sha256:9dae6b954450aeae3068263ba52b2bcf4f939f1c09daac98f2bf446c41851447

Observation 24933bd1-0dce-42e8-994c-af1d4374b14e · outbound

This paper cites COCO-Text: Dataset and Benchmark for Text Detection and Recognition in Natural Images.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition COCO-Text: Dataset and Benchmark for Text Detection and Recognition in Natural Images

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:51.194473Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:51.194473Z digest=sha256:2edd3eae7f41bff26d071fd8cefabdf8b38698d57050d973e2f65580fd8d0c43

Observation 45cc2fd2-fd34-4f99-945c-0eb27d4d8b0f · outbound

This paper cites Curved scene text detection via transverse and longitudinal sequence connection,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Curved scene text detection via transverse and longitudinal sequence connection,

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:51.919960Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T18:51:51.201544Z digest=sha256:7b095484836f99c13e5c111a9838e94f6210fa1427c03f09093001c0945cd45d

Observation 5bd2736b-1dfe-4b13-9ac6-07ee9568cf60 · outbound

This paper cites Total-text: A comprehensive dataset for scene text detection and recognition,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Total-text: A comprehensive dataset for scene text detection and recognition,

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:51.892724Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T18:51:51.208011Z digest=sha256:a75cc479ce8b82a5649169f20b1851b797c10c6e33ec8292f4ebeb3102143b98

Observation 711f8c5e-dace-4049-bcfb-c1af68d97e23 · outbound

This paper cites From two to one: A new scene text recognizer with visual language modeling network,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition From two to one: A new scene text recognizer with visual language modeling network,

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.924987Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T18:51:51.216055Z digest=sha256:a024b45338dbb264bf0ddb58cbb370b03551489c7a14712560e5aaa75e46a156

Observation c48a0081-eca3-4332-ad4f-8b03272225dd · outbound

This paper cites Robust scene text recognition with automatic rectification,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Robust scene text recognition with automatic rectification,

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:51.862579Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T18:51:51.221565Z digest=sha256:46401a8787b76cdf8f6f884a008d1655ca517a0942750252433c0afe37dfa127

Observation 567b5ee9-14fe-446f-9024-14baf5d428e4 · outbound

This paper cites What is wrong with scene text recognition model comparisons? dataset and model analysis,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition What is wrong with scene text recognition model comparisons? dataset and model analysis,

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:51.838323Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T18:51:51.226112Z digest=sha256:5d001fc8ade522d4667ef465fea1f25803dd364ca69b27ba86db4f0f5850f9ba

Observation bb93bc87-90ea-4726-864b-ee8d92d23afd · outbound

This paper cites Synthetic Data and Artificial Neural Networks for Natural Scene Text Recognition.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Synthetic Data and Artificial Neural Networks for Natural Scene Text Recognition

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:51.233345Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:51.233345Z digest=sha256:cad0c88b7ee03d433a970229f4292dcd533b073ef90693543224fa13461c070e

Observation 269c4479-5ac7-498e-a195-c818e760d4f0 · outbound

This paper cites Benchmarking Chinese Text Recognition: Datasets, Baselines, and an Empirical Study.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Benchmarking Chinese Text Recognition: Datasets, Baselines, and an Empirical Study

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:51.239927Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:51.239927Z digest=sha256:4d9c2ab9dfdd301ef4748b66ae0d977213f9f1f545583d06c7f0d4500e8986d5

Observation 92f2fc5d-958a-44b0-93bd-e502b61e316d · outbound

This paper cites The iam-database: an english sentence database for offline handwriting recognition,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition The iam-database: an english sentence database for offline handwriting recognition,

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:51.817737Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T18:51:51.248254Z digest=sha256:e9779658a396167d5be7d9c9c84b7476c55774add6e5cc3d92c637f37311f047

Observation c6b7fa5a-4e6d-477d-8e3f-187a58c3ad9f · outbound

This paper cites Cvl-database: An off-line database for writer retrieval, writer identification and word spotting,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Cvl-database: An off-line database for writer retrieval, writer identification and word spotting,

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:51.783072Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-12T18:51:51.254718Z digest=sha256:3111d5393a2f6dd034d59e716f5fd06fb22bab50a8402f59a39a92e57bb5c913

Observation e5059467-309a-4b32-9523-a682fb31f211 · outbound

This paper cites Stochastic neighbor embedding,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Stochastic neighbor embedding,

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:51.260140Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:51.260140Z digest=sha256:c4cbfcfe5c90ad01a85de17274a1cbff7ca28e88a8cfae3fa13269353ae8294f

Pith citing papers

No inbound Pith citation observations are available.