Pith. sign in

Paper Citation Record · LEDGER

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition

As of 17 August 2026, this Paper Citation Record lists 60 of 60 outbound references and 0 inbound Pith citation observations for arXiv:2411.11219.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.11219 v2

Coverage vector

measured 60 of 60 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T18:51:51.260140Z

measured 60 of 60 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

60 of 60 outbound references displayed

  • verified exact1
  • verified fuzzy31
  • unresolved28
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation cfcb1dca-9492-43bb-9f53-93c1fc5400a8 · outbound

This paper cites Unsupervised feature learning via non-parametric instance discrimination,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Unsupervised feature learning via non-parametric instance discrimination,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:53.038960Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T18:51:50.760300Z digest=sha256:36cded2c7a3ce3563300c64714c6517ee8167f95f6e3479d9aa3654eea796123

Observation ed596de8-4d2b-42d5-bca6-f7660c7ef85f · outbound

This paper cites Representation Learning with Contrastive Predictive Coding.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Representation Learning with Contrastive Predictive Coding

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:50.769703Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:50.769703Z digest=sha256:869eacb13d9f9ec8bc6d0b0891a5384643fff9833906127de20e65cafde382bd

Observation fcf23d59-0ff2-495a-ac06-a129da57c68a · outbound

This paper cites Momentum contrast for unsupervised visual representation learning,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Momentum contrast for unsupervised visual representation learning,

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:50.776195Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:50.776195Z digest=sha256:021f4468b889b0559b9a4e34d88e5881c323ee7dbe22ab41d30874ea1ef61ab5

Observation f58e6fd3-0700-4869-901f-bfd79e4cdb88 · outbound

This paper cites A simple framework for contrastive learning of visual representations,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition A simple framework for contrastive learning of visual representations,

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:50.783247Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:50.783247Z digest=sha256:cb7e3febb3a0ec1ca89615cd682fb031cf4cf2c562d309cd823d171f33d852b2

Observation 00ff5207-7c73-482f-b8e4-d43d0f45d396 · outbound

This paper cites BEiT: BERT Pre-Training of Image Transformers.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition BEiT: BERT Pre-Training of Image Transformers

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:50.789743Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:50.789743Z digest=sha256:2ea054d3c92c153d7e5b8077b046541a5e5ca2e86ee18e7af837961d17304450

Observation d8f214a5-cd92-45f6-85c0-69b4b94695b8 · outbound

This paper cites Masked au- toencoders are scalable vision learners,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Masked au- toencoders are scalable vision learners,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:50.796041Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:50.796041Z digest=sha256:fe6c6de81097b4c7e4d714fd9e60f9832f565ca78395e5bb4a06430e71613863

Observation e4dc2277-8f68-4774-87c0-54b6a897e8bc · outbound

This paper cites Simmim: A simple framework for masked image modeling,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Simmim: A simple framework for masked image modeling,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.963075Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T18:51:50.805358Z digest=sha256:8674e6f587bfa2c4291e9809ce05a3c4056608d11c8195785b4e6ad9beb1fe18

Observation b08f27d2-29f0-450f-ace9-59ed3504491d · outbound

This paper cites Read like humans: Autonomous, bidirectional and iterative language modeling for scene text recognition,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Read like humans: Autonomous, bidirectional and iterative language modeling for scene text recognition,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.946358Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T18:51:50.816191Z digest=sha256:f6107e6d618f901dd6e564636875a326d48d493f3f10b8000dc026702c78cbd6

Observation c8ab4525-95a7-48b8-864a-f1e500388b9f · outbound

This paper cites Sequence-to-sequence contrastive learn- ing for text recognition,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Sequence-to-sequence contrastive learn- ing for text recognition,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:50.837653Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:50.837653Z digest=sha256:b81b6e8a791e3a31946bd61dd33f31ff6d5bfe109a97c6b6d3d49788998405e7

Observation 88f6549d-a09e-4af1-a8f8-5efb4b2cafdb · outbound

This paper cites Perceiving stroke-semantic context: Hierarchical contrastive learning for robust scene text recognition,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Perceiving stroke-semantic context: Hierarchical contrastive learning for robust scene text recognition,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.841789Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T18:51:50.844333Z digest=sha256:e7cc5d4b415fdb6b02cf0231005b5f03fbd2cdfeb4c2fa63ad796090763f0649

Observation 33351973-32fb-4f94-9b5f-471929891d37 · outbound

This paper cites Reading and writing: Discriminative and generative modeling for self-supervised text recognition,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Reading and writing: Discriminative and generative modeling for self-supervised text recognition,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:50.852070Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:50.852070Z digest=sha256:ac305cb71b747641738b1f26487fb0f825a737d483ce0d019707dab3488bbd5e

Observation 7229e8de-c376-4324-8a17-e29d8fa895b1 · outbound

This paper cites An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:50.857745Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:50.857745Z digest=sha256:591611eab62e45897329d8ef76efd8056a0c760bd5b98f458648d0fbc7f2ee2d

Observation 3e913a1d-5220-4e8e-86fd-8c3c80b5d31f · outbound

This paper cites MaskOCR: Text Recognition with Masked Encoder-Decoder Pretraining.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition MaskOCR: Text Recognition with Masked Encoder-Decoder Pretraining

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:50.864066Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:50.864066Z digest=sha256:3ed6e6387a3b385a55d6d6c61f978548684de6d83a9aa41e74677bbbc0704c2b

Observation fc24f0a7-bc9b-4f8e-a57c-6ed45b9eefa6 · outbound

This paper cites Towards accurate scene text recognition with semantic reasoning networks,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Towards accurate scene text recognition with semantic reasoning networks,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.783768Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T18:51:50.870943Z digest=sha256:686d3644568ab8154fec21fd7951e81d4d503f6637100bcf001f18f37670fbe0

Observation 964ec38a-f7e9-4439-994f-5dde280b75d9 · outbound

This paper cites Seed: Semantics enhanced encoder-decoder framework for scene text recognition,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Seed: Semantics enhanced encoder-decoder framework for scene text recognition,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.754387Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T18:51:50.877425Z digest=sha256:531b8586360cabeb2cd0815f6e5acab6d0e2be53b841650c1dd6302b907a2449

Observation a3838276-dd2e-420c-9379-821d21227ddf · outbound

This paper cites On vocabulary reliance in scene text recognition,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition On vocabulary reliance in scene text recognition,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.717253Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T18:51:50.885900Z digest=sha256:093def4adf03697b4b3ccebc1632e050b63bafb7165f3761dc493cf02dbc7536

Observation 70b6e185-263e-4787-9fa5-a6a20773d936 · outbound

This paper cites Ressl: Relational self-supervised learning with weak augmentation,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Ressl: Relational self-supervised learning with weak augmentation,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.676841Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T18:51:50.892516Z digest=sha256:786f6f04aee7affe72912c68b3910e3c2fb8a40321fb3a192418abac4cb4fbb0

Observation a0266b4a-835a-475c-959f-757b35f52242 · outbound

This paper cites Synthetic data for text localisation in natural images,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Synthetic data for text localisation in natural images,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:50.899694Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:50.899694Z digest=sha256:3d83c30f6ac51509b2c4d08446ddb82e9804879d87c1d7ee151a332e3ddcd700

Observation 2f0c00b6-91b6-4bfd-98bc-4ef005f9cb42 · outbound

This paper cites Rethink- ing text segmentation: A novel dataset and a text-specific refinement approach,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Rethink- ing text segmentation: A novel dataset and a text-specific refinement approach,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.623318Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T18:51:50.908188Z digest=sha256:fcc02800c7a6d97e90393929b583e5640a0b35117b2847f3e53267a831783a1e

Observation 1c9b8423-121c-463b-8ab5-362d7b5c435f · outbound

This paper cites Relational contrastive learning for scene text recognition,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Relational contrastive learning for scene text recognition,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.598854Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T18:51:50.914335Z digest=sha256:9fcc75dd436d857639ed63905f7b5bf46ba892a38e30b9ce2d3b1efbae91b0c9

Observation 60134332-9388-44ca-a69a-d6ea42533e25 · outbound

This paper cites Unsupervised learning of visual features by contrasting cluster assign- ments,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Unsupervised learning of visual features by contrasting cluster assign- ments,

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:50.923630Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:50.923630Z digest=sha256:0805f8884446105ab45f1cba9b86ce2b9c997f85afdc6bf411c391e86d712a13

Observation 183fe724-b5f8-4c46-a267-ca6f8f487dd6 · outbound

This paper cites Bootstrap your own latent-a new approach to self-supervised learning,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Bootstrap your own latent-a new approach to self-supervised learning,

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:50.930894Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:50.930894Z digest=sha256:39966aa4f435bd18aa06a9c0b4123fb252f966b7da1e92794ff40fe8e5587245

Observation 4a006d29-137f-4795-99a6-e0c851631114 · outbound

This paper cites Exploring simple siamese representation learning,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Exploring simple siamese representation learning,

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:50.940508Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:50.940508Z digest=sha256:69b1610f5ef2d44440daab3b55363e16f3201e17b337738c762e925e666a8b3d

Observation 27ea1a1d-6ecd-45c2-b8f2-0073a227c032 · outbound

This paper cites Emerging properties in self-supervised vision transformers,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Emerging properties in self-supervised vision transformers,

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:50.945778Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:50.945778Z digest=sha256:0d8d43760c94d4af3254bb3aeb73f3449eb2a2d211e7fb12788e113c1c5ec9a3

Observation b91e51fb-4ca1-4577-9ff7-223bb5a18fae · outbound

This paper cites iBOT: Image BERT Pre-Training with Online Tokenizer.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition iBOT: Image BERT Pre-Training with Online Tokenizer

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:50.951689Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:50.951689Z digest=sha256:7da6db4fb18369fc7e146d2fe3129a0c8e65cb935aa6cb81b551a4954eee4a47

Observation 01b76be6-5c32-471f-a601-c6c1fe7fa17e · outbound

This paper cites DINOv2: Learning Robust Visual Features without Supervision.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition DINOv2: Learning Robust Visual Features without Supervision

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:50.962274Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:50.962274Z digest=sha256:111ea85f7264c4e13a51ebf27628197fbbe8151b2fd178d19753eb7e46b7f830

Observation d26d25f2-e35c-457a-99d7-0e3c0d0c14f2 · outbound

This paper cites Contrastive learning with stronger augmenta- tions,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Contrastive learning with stronger augmenta- tions,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.477016Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T18:51:50.969284Z digest=sha256:23fc17948b20bfa5b36f7222a18a3970f44c3694d2492b3ecdb92412fd824383

Observation a3086599-781f-4b39-b4fd-f2ccd2f2e188 · outbound

This paper cites What if we only use real datasets for scene text recognition? toward scene text recognition with fewer labels,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition What if we only use real datasets for scene text recognition? toward scene text recognition with fewer labels,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.440978Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T18:51:50.978025Z digest=sha256:485cdff00d5f34ee4296e78bd0d6edfe5b21b6e33012c35cf9cba48d8f4b3666

Observation ddbd37fd-da62-400b-9b57-70ae7bf42747 · outbound

This paper cites Siman: Exploring self-supervised rep- resentation learning of scene text via similarity-aware normalization,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Siman: Exploring self-supervised rep- resentation learning of scene text via similarity-aware normalization,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.404622Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T18:51:50.987664Z digest=sha256:161b74eb563b0a110933a4ebd4adb5608a7ba0d57325663f7754a78b15559001

Observation 93e80a5d-fbbe-4a0f-8e1d-f1c8f777b9b4 · outbound

This paper cites Self- supervised character-to-character distillation for text recognition,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Self- supervised character-to-character distillation for text recognition,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.377343Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T18:51:50.996746Z digest=sha256:c4770d13ade32c9e168b1bce7aec152b0d78bed61e890602da307003d39b8099

Observation b2444aa3-aacf-4583-ab26-7478617f1f99 · outbound

This paper cites Reading scene text in deep convolutional sequences,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Reading scene text in deep convolutional sequences,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.350919Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T18:51:51.007434Z digest=sha256:926ea4e8f3b511c1e90ea87facd1229132a68af637e56d0f7f6f54fa1b78268f

Observation 70bd8651-a4a6-4636-800a-56f7253b28c1 · outbound

This paper cites Accurate recognition of words in scenes without char- acter segmentation using recurrent neural network,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Accurate recognition of words in scenes without char- acter segmentation using recurrent neural network,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.322809Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T18:51:51.039572Z digest=sha256:276ac5403ec2e17f772c24bdf9c21ee66c07e4990d5c3db22f1338d5c7ea7a8e

Observation b8a625bb-108a-48a2-bdbb-333c455433da · outbound

This paper cites An end-to-end trainable neural network for image-based sequence recognition and its application to scene text recognition,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition An end-to-end trainable neural network for image-based sequence recognition and its application to scene text recognition,

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:51.049965Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:51.049965Z digest=sha256:35812bd4ea9534eecced26d97a643740ef477db364f9da395e74a6964c75104b

Observation e4f90b37-65c5-4c45-bb1b-5e6f01e50b3a · outbound

This paper cites Aster: An attentional scene text recognizer with flexible rectification,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Aster: An attentional scene text recognizer with flexible rectification,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.252200Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T18:51:51.057681Z digest=sha256:9ca14ca38695d54620d1bd99107353c63110513e6b5e77733b77da0555e3bad8

Observation 5f53a680-4315-4f3f-b2fd-9430fae75fa3 · outbound

This paper cites Learning to read irregular text with attention mechanisms.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Learning to read irregular text with attention mechanisms

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.222611Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T18:51:51.068453Z digest=sha256:eb3aab0b08e016f0d73d84554938172a70ea20b294b570391e3c94342162e9aa

Observation d06b7ef9-7e6b-49b8-855b-4bbb1c8a0710 · outbound

This paper cites Attention-based extraction of structured information from street view imagery,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Attention-based extraction of structured information from street view imagery,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.190639Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T18:51:51.078112Z digest=sha256:0c74bc3deb1068fef02bb55564b32b3cb6b4bc16ef5c50b655baf742eee02306

Observation 68f78fcc-ae89-4967-80d1-844cb149945e · outbound

This paper cites On recognizing texts of arbitrary shapes with 2d self-attention,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition On recognizing texts of arbitrary shapes with 2d self-attention,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.902039Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T18:51:51.091409Z digest=sha256:f3d8ccd31daa1dda497415dbe4561f58adb3d550cfca968d09921f1606ab649a

Observation 55dbeeb6-2eee-4546-a882-dfbeea03c7b6 · outbound

This paper cites Master: Multi-aspect non-local network for scene text recognition,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Master: Multi-aspect non-local network for scene text recognition,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.167233Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T18:51:51.099312Z digest=sha256:e07663e84aacba10fe4964551f2c76415e2e70ad9269c11740b382f43d9d21c8

Observation 36aab0a8-24ed-44ac-b68e-46bf06759c6a · outbound

This paper cites Context-based contrastive learning for scene text recognition,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Context-based contrastive learning for scene text recognition,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.130710Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T18:51:51.110585Z digest=sha256:d8d67b8f34fe97c0db034b1c04fab0f8d5394619acd5b1856060063ecfa54171

Observation 30d36c02-f4b6-46db-b23f-b3e5ebda92f5 · outbound

This paper cites What Do Self-Supervised Vision Transformers Learn?.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition What Do Self-Supervised Vision Transformers Learn?

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:51.115642Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:51.115642Z digest=sha256:08d352de1ef6269a9c82b2572431c6051281decd7490a11e13d2d07a8d86bc1e

Observation 2c07420e-0c87-43e3-8e7a-ba146686b80e · outbound

This paper cites Masked Siamese ConvNets.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Masked Siamese ConvNets

Reference 43

Resolution
verified exact
local_arxiv, observed 2026-08-12T18:51:51.459743Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T18:51:51.131141Z digest=sha256:22b59c7dcd571b295a17493e7c26af58563a7ef0bec5c1ee0977143b0be256ba

Observation d946ce2f-def7-4683-b9ff-12f493d29f49 · outbound

This paper cites Convnext v2: Co-designing and scaling convnets with masked autoen- coders,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Convnext v2: Co-designing and scaling convnets with masked autoen- coders,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.099905Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T18:51:51.140411Z digest=sha256:66b43bb395712443a8ccd03f70d4b62f2bb7e5debe24e002fe077fa6ed7fe0ba

Observation 3d00ae23-eaea-49d2-ba9d-22c07fd90ffa · outbound

This paper cites Scene text recognition using higher order language priors,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Scene text recognition using higher order language priors,

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:51.149908Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:51.149908Z digest=sha256:6b5216746bcfaf22369626ade48136e00ab620ae4aca4570160ac65bf19d7036

Observation daa89420-6b8a-4705-9b91-a40edd7ea10a · outbound

This paper cites Icdar 2003 robust reading competitions: entries, results, and future directions,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Icdar 2003 robust reading competitions: entries, results, and future directions,

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.050339Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T18:51:51.158424Z digest=sha256:e64d2806facffe0ece2fda65acc685a64e6d92daad7960d0ecaecc0fd3bb5abf

Observation 3e6d85b9-1692-43e9-a2a4-341c139fd1cc · outbound

This paper cites Icdar 2013 robust reading competition,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Icdar 2013 robust reading competition,

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:51.165266Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:51.165266Z digest=sha256:4c41e510ce8d13bd050c2a44c92d774ff1db7d44f40130e34b5fad481a13e775

Observation f7cfbb70-e9fa-40b3-b8fc-1a160fc60d81 · outbound

This paper cites End-to-end scene text recog- nition,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition End-to-end scene text recog- nition,

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:51.171251Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:51.171251Z digest=sha256:fe5e41b03da6fe2cc65d5d2407c4de619a42cabd0b45f1f0693a62885eb7eccb

Observation d0ec2b24-e01f-451e-a75e-c357c32a4627 · outbound

This paper cites Icdar 2015 competition on robust reading,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Icdar 2015 competition on robust reading,

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:51.176641Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:51.176641Z digest=sha256:f0d3f92df5c0034df06c3b3a4b8a9e9819ee9592e963a1afcf635f4b4187e400

Observation 709297c3-c80f-4f4d-85dd-92e3c2e0324b · outbound

This paper cites Recognizing text with perspective distortion in natural scenes,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Recognizing text with perspective distortion in natural scenes,

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:51.181284Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:51.181284Z digest=sha256:a2dc479fb3dd2bf307757c8e02246fc1822adb548c58d999e657633b12439858

Observation a5e9a834-5d93-4e84-a1b0-8f6ea61ef8eb · outbound

This paper cites A robust arbitrary text detection system for natural scene images,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition A robust arbitrary text detection system for natural scene images,

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:51.189723Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:51.189723Z digest=sha256:9dae6b954450aeae3068263ba52b2bcf4f939f1c09daac98f2bf446c41851447

Observation 24933bd1-0dce-42e8-994c-af1d4374b14e · outbound

This paper cites COCO-Text: Dataset and Benchmark for Text Detection and Recognition in Natural Images.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition COCO-Text: Dataset and Benchmark for Text Detection and Recognition in Natural Images

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:51.194473Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:51.194473Z digest=sha256:2edd3eae7f41bff26d071fd8cefabdf8b38698d57050d973e2f65580fd8d0c43

Observation 45cc2fd2-fd34-4f99-945c-0eb27d4d8b0f · outbound

This paper cites Curved scene text detection via transverse and longitudinal sequence connection,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Curved scene text detection via transverse and longitudinal sequence connection,

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:51.919960Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T18:51:51.201544Z digest=sha256:58d36de510966dfc6517da8910c806a2fc5941e447a30cbb23bc308529825a63

Observation 5bd2736b-1dfe-4b13-9ac6-07ee9568cf60 · outbound

This paper cites Total-text: A comprehensive dataset for scene text detection and recognition,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Total-text: A comprehensive dataset for scene text detection and recognition,

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:51.892724Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T18:51:51.208011Z digest=sha256:6d2675ebb6fc78b29259ba8f5551060322847d4446ef5ec4007e25ee610dafa5

Observation 711f8c5e-dace-4049-bcfb-c1af68d97e23 · outbound

This paper cites From two to one: A new scene text recognizer with visual language modeling network,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition From two to one: A new scene text recognizer with visual language modeling network,

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:52.924987Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T18:51:51.216055Z digest=sha256:c254fd7bdeb948b8f6350ec47c381c2dcbef6eba397b6561e549e4ab72826def

Observation c48a0081-eca3-4332-ad4f-8b03272225dd · outbound

This paper cites Robust scene text recognition with automatic rectification,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Robust scene text recognition with automatic rectification,

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:51.862579Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T18:51:51.221565Z digest=sha256:d9c21d999036b792a8d1d8005c11bc31324454a01005f42ea612d6e935ae27ef

Observation 567b5ee9-14fe-446f-9024-14baf5d428e4 · outbound

This paper cites What is wrong with scene text recognition model comparisons? dataset and model analysis,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition What is wrong with scene text recognition model comparisons? dataset and model analysis,

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:51.838323Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T18:51:51.226112Z digest=sha256:a528b303cee5e2bc47aa2181511978e9808453b815fae4c8e826e841f953c027

Observation bb93bc87-90ea-4726-864b-ee8d92d23afd · outbound

This paper cites Synthetic Data and Artificial Neural Networks for Natural Scene Text Recognition.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Synthetic Data and Artificial Neural Networks for Natural Scene Text Recognition

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:51.233345Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:51.233345Z digest=sha256:cad0c88b7ee03d433a970229f4292dcd533b073ef90693543224fa13461c070e

Observation 269c4479-5ac7-498e-a195-c818e760d4f0 · outbound

This paper cites Benchmarking Chinese Text Recognition: Datasets, Baselines, and an Empirical Study.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Benchmarking Chinese Text Recognition: Datasets, Baselines, and an Empirical Study

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:51.239927Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:51.239927Z digest=sha256:4d9c2ab9dfdd301ef4748b66ae0d977213f9f1f545583d06c7f0d4500e8986d5

Observation 92f2fc5d-958a-44b0-93bd-e502b61e316d · outbound

This paper cites The iam-database: an english sentence database for offline handwriting recognition,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition The iam-database: an english sentence database for offline handwriting recognition,

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:51.817737Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T18:51:51.248254Z digest=sha256:4c9cc3611279d622211b6451661696ce154d0b6f112a6693be9a983b7f854721

Observation c6b7fa5a-4e6d-477d-8e3f-187a58c3ad9f · outbound

This paper cites Cvl-database: An off-line database for writer retrieval, writer identification and word spotting,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Cvl-database: An off-line database for writer retrieval, writer identification and word spotting,

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:51:51.783072Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T18:51:51.254718Z digest=sha256:775188b24913d81b9f5d286298316238b7bfe9a019e689b874abd1ae8f567118

Observation e5059467-309a-4b32-9523-a682fb31f211 · outbound

This paper cites Stochastic neighbor embedding,.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition Stochastic neighbor embedding,

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:51.260140Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:51.260140Z digest=sha256:c4cbfcfe5c90ad01a85de17274a1cbff7ca28e88a8cfae3fa13269353ae8294f

Pith citing papers

No inbound Pith citation observations are available.