Pith. sign in

Paper Citation Record · LEDGER

CLIP4STR: A Simple Baseline for Scene Text Recognition with Pre-trained Vision-Language Model

As of 12 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 2 inbound Pith citation observations for arXiv:2305.14014.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2305.14014 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 2 of 2 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T05:56:41.159898Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-06T05:56:41.326008Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 3cfee4c0-c6bb-4305-b7ef-1bc7fcb16e67 · inbound

TEACH: Text Encoding as Curriculum Hints for Scene Text Recognition cites this paper.

TEACH: Text Encoding as Curriculum Hints for Scene Text Recognition CLIP4STR: A Simple Baseline for Scene Text Recognition with Pre-trained Vision-Language Model

Reference 14

Resolution
verified exact
local_arxiv, observed 2026-08-06T05:56:41.329345Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T05:56:41.159898Z digest=sha256:b492f95faffa002cb14b6a317337e287d737a7f26c08167f74f8805c23c2715a

Observation 512aa118-4c3b-4bb9-926b-fdb0ed05ef79 · inbound

HAT Super-Resolution and a PARSeq+CLIP4STR Voting Ensemble for Extreme In-the-Wild License Plate Recognition cites this paper.

HAT Super-Resolution and a PARSeq+CLIP4STR Voting Ensemble for Extreme In-the-Wild License Plate Recognition CLIP4STR: A Simple Baseline for Scene Text Recognition with Pre-trained Vision-Language Model

Reference 3

Resolution
unresolved
no resolver link, observed 2026-07-13T05:58:43.042797Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T05:58:43.042797Z digest=sha256:391c83cf1251b9fa55b15bc315b18487655ac832fa47678b0710120c4e3d3827