Pith. sign in

Paper Citation Record · LEDGER

TrOCR: Transformer-based Optical Character Recognition with Pre-trained Models

As of 21 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 16 inbound Pith citation observations for arXiv:2109.10282.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2109.10282 v5

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 16 of 16 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 16 of 16 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T20:58:40.756810Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T09:59:45.215932Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation f885576d-2d45-4e59-adb5-288dda9c0720 · inbound

PaLM-E: An Embodied Multimodal Language Model cites this paper.

PaLM-E: An Embodied Multimodal Language Model TrOCR: Transformer-based Optical Character Recognition with Pre-trained Models

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-10T22:29:29.944797Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-10T22:29:29.631351Z digest=sha256:b5726d1e78a652db00567664e84c0fc00744d605f1b11303f8cca2e5fe94c381

Observation 6b220a2a-95d1-4458-864b-2d72902168ec · inbound

Cleansing Jewel: A Neural Spelling Correction Model Built On Google OCR-ed Tibetan Manuscripts cites this paper.

Cleansing Jewel: A Neural Spelling Correction Model Built On Google OCR-ed Tibetan Manuscripts TrOCR: Transformer-based Optical Character Recognition with Pre-trained Models

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-24T09:29:17.196979Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-24T09:28:35.914041Z digest=sha256:1db8b60f44d2416102b339c7a83a5dde16397bd6614bc775b84d92c3b15fa60f

Observation 4c45c11c-0420-4ef5-96b3-909ae99d3b00 · inbound

Nougat: Neural Optical Understanding for Academic Documents cites this paper.

Nougat: Neural Optical Understanding for Academic Documents TrOCR: Transformer-based Optical Character Recognition with Pre-trained Models

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-16T09:42:12.557852Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-16T09:42:12.463309Z digest=sha256:1be2094c36f162e73311f2cb9ab1902a5fc42db171827121569800d0497fb062

Observation 08ea73c4-027d-46c0-91cf-6eae061111b0 · inbound

Arabic-Nougat: Fine-Tuning Vision Transformers for Arabic OCR and Markdown Extraction cites this paper.

Arabic-Nougat: Fine-Tuning Vision Transformers for Arabic OCR and Markdown Extraction TrOCR: Transformer-based Optical Character Recognition with Pre-trained Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-12T17:33:44.993280Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T17:33:44.993280Z digest=sha256:e21d7774add0857c6ef68235393c1c574cfc5072630d13c3816b6a8431b92074

Observation 083a522b-f88d-4847-abc8-0065526ea0bb · inbound

HAND: Hierarchical Attention Network for Multi-Scale Handwritten Document Recognition and Layout Analysis cites this paper.

HAND: Hierarchical Attention Network for Multi-Scale Handwritten Document Recognition and Layout Analysis TrOCR: Transformer-based Optical Character Recognition with Pre-trained Models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-11T01:04:46.694685Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T01:04:46.694685Z digest=sha256:52d91cb6575e279402fb4454654f920b90165ef1150c8b5ea98688356202fe5d

Observation fdf4eebc-d0f3-45c8-9fff-3ee3fef44f98 · inbound

Low-Resource Language Processing: An OCR-Driven Summarization and Translation Pipeline cites this paper.

Low-Resource Language Processing: An OCR-Driven Summarization and Translation Pipeline TrOCR: Transformer-based Optical Character Recognition with Pre-trained Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T20:58:40.756810Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:58:40.756810Z digest=sha256:07651a44b9b768b693a9595864e2dfc2cac9a10c094065829e9a3ea8c199cca8

Observation 7b22ec8f-aea2-4168-9007-bd1a8a07bfb7 · inbound

WriteViT: Handwritten Text Generation with Vision Transformer cites this paper.

WriteViT: Handwritten Text Generation with Vision Transformer TrOCR: Transformer-based Optical Character Recognition with Pre-trained Models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-15T20:21:05.298478Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:21:05.298478Z digest=sha256:9047559321453c7596ea3bbceb301222c39569cde19bd1c7dacc4f02539e957c

Observation a726c355-ef02-40c6-a81b-97d9637db9a0 · inbound

Predicting the Past: Estimating Historical Appraisals with OCR and Machine Learning cites this paper.

Predicting the Past: Estimating Historical Appraisals with OCR and Machine Learning TrOCR: Transformer-based Optical Character Recognition with Pre-trained Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T12:21:13.469848Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:21:13.469848Z digest=sha256:f6738285b8e95eed33e68f7e371ed548b950b6799860fc74a13d1762e6c4c35d

Observation ec05221d-8ed6-4c63-9330-eef5e9d199f4 · inbound

Towards Selection of Large Multimodal Models as Engines for Burned-in Protected Health Information Detection in Medical Images cites this paper.

Towards Selection of Large Multimodal Models as Engines for Burned-in Protected Health Information Detection in Medical Images TrOCR: Transformer-based Optical Character Recognition with Pre-trained Models

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-22T13:26:35.616274Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-22T13:25:37.515812Z digest=sha256:0a473008bab3c2b79c34e8830774efab1c57ff6421872ec43770e4df4acdcdbd

Observation 0144df45-62ec-49cb-a19d-64a362904cb7 · inbound

LightOnOCR: A 1B End-to-End Multilingual Vision-Language Model for State-of-the-Art OCR cites this paper.

LightOnOCR: A 1B End-to-End Multilingual Vision-Language Model for State-of-the-Art OCR TrOCR: Transformer-based Optical Character Recognition with Pre-trained Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-03T09:20:16.742613Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T09:20:16.742613Z digest=sha256:23bce809b0236378131aa734346372dee1b1dd3a0e83a6839c1daf151d9ddb34

Observation 522a78a7-505f-41a3-af70-20212adcd08b · inbound

Seeing is Coding: On the Effectiveness of Vision Language Models in Code Understanding cites this paper.

Seeing is Coding: On the Effectiveness of Vision Language Models in Code Understanding TrOCR: Transformer-based Optical Character Recognition with Pre-trained Models

Reference 57

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T08:32:36.450368Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-16T08:30:50.984873Z digest=sha256:7371368c15355df751f1be3081ca741ea40bb4131090e7d47aa003e8a0f3b2ba

Observation c92cc229-d948-4f52-8b8d-d827f36004da · inbound

The Character Error Vector: Decomposable errors for page-level OCR evaluation cites this paper.

The Character Error Vector: Decomposable errors for page-level OCR evaluation TrOCR: Transformer-based Optical Character Recognition with Pre-trained Models

Reference 23

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T23:25:52.899638Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-10T19:09:26.075739Z digest=sha256:9cff780ff0cd981669575d92d6ab92cd79169118652f0a336bb5f85d76b7f0bb

Observation 018fe94b-b0d6-41d4-9e20-f5b5f6a7b352 · inbound

From Handwriting to Structured Data: Benchmarking AI Digitisation of Handwritten Forms cites this paper.

From Handwriting to Structured Data: Benchmarking AI Digitisation of Handwritten Forms TrOCR: Transformer-based Optical Character Recognition with Pre-trained Models

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-11T09:16:02.594049Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-10T16:08:30.056580Z digest=sha256:218f14eb61b1d566265bab427e05cd512a612ab194d76c7b17cfa8ceb2721d14

Observation 04d7cfb0-2882-48b5-8b33-d7d460d14f22 · inbound

RaV-IDP: A Reconstruction-as-Validation Framework for Faithful Intelligent Document Processing cites this paper.

RaV-IDP: A Reconstruction-as-Validation Framework for Faithful Intelligent Document Processing TrOCR: Transformer-based Optical Character Recognition with Pre-trained Models

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-11T21:11:10.569769Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-08T06:37:59.084380Z digest=sha256:313b4e23b0a059a0a4465934b85b72d5078329d7d427b9f3a613e3d42e6bd2df

Observation fca84310-5f07-45d0-8e00-0f8e1fcd1d03 · inbound

Comparative Analysis of Liquid Neural Networks and LSTM for Sequential Pattern Recognition: Robustness, Efficiency, and Clinical Utility cites this paper.

Comparative Analysis of Liquid Neural Networks and LSTM for Sequential Pattern Recognition: Robustness, Efficiency, and Clinical Utility TrOCR: Transformer-based Optical Character Recognition with Pre-trained Models

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-06-29T19:33:54.179957Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-29T19:27:13.782327Z digest=sha256:09606e3975c3fd52339b012c0abc488a8495d5d2a224c9ca2dfc8534d72a1a89

Observation 6f2f35d8-b82b-49d6-91f9-a953f9e9d047 · inbound

Koshur Pixel: a large-scale synthetic ocr dataset for kashmiri cites this paper.

Koshur Pixel: a large-scale synthetic ocr dataset for kashmiri TrOCR: Transformer-based Optical Character Recognition with Pre-trained Models

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-07-04T09:59:45.217732Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-26T09:19:11.131168Z digest=sha256:aeb468882780e7b5e49b8c5794b0ea5554ad8639c6b62ba781b6a819bb01ac7b