Pith. sign in

Paper Citation Record · LEDGER

A Token-level Text Image Foundation Model for Document Understanding

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2503.02304.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2503.02304 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:57:17.146130Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-20T14:43:22.323230Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation bbac3d35-4df6-44d7-9649-39c93dbbfd3c · inbound

OCR-Reasoning Benchmark: Unveiling the True Capabilities of MLLMs in Complex Text-Rich Image Reasoning cites this paper.

OCR-Reasoning Benchmark: Unveiling the True Capabilities of MLLMs in Complex Text-Rich Image Reasoning A Token-level Text Image Foundation Model for Document Understanding

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:17.146130Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:57:17.146130Z digest=sha256:cfd3942fbe31fecd5d12b31a4ad630553d5bc9553cd442e1666a35987f393681

Observation 97f94665-d118-465f-8682-43be64fe1d70 · inbound

CodePercept: Code-Grounded Visual STEM Perception for MLLMs cites this paper.

CodePercept: Code-Grounded Visual STEM Perception for MLLMs A Token-level Text Image Foundation Model for Document Understanding

Reference 17

Resolution
unresolved
no resolver link, observed 2026-07-14T23:22:13.847876Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T23:22:13.847876Z digest=sha256:58bd19f5d779c6f5bbcd9982465a3e3362e4521e048cc0aa0ef694844e488ee9

Observation 3a06ee62-7ad7-44a4-b8b4-d0b1bf4a9e40 · inbound

StyleTextGen: Style-Conditioned Multilingual Scene Text Generation cites this paper.

StyleTextGen: Style-Conditioned Multilingual Scene Text Generation A Token-level Text Image Foundation Model for Document Understanding

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-15T04:55:03.129477Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T04:54:52.729257Z digest=sha256:fa193446b0533c42af12f0f1bf99ae359ae3185332d1bffb3ea2cc21656db7f7

Observation aa67edf7-4fb4-44c2-b300-89d8c139ed19 · inbound

Beyond Detection: A Structure-Aware Framework for Scene Text Tracking cites this paper.

Beyond Detection: A Structure-Aware Framework for Scene Text Tracking A Token-level Text Image Foundation Model for Document Understanding

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-20T14:43:22.333828Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-20T14:42:38.961728Z digest=sha256:a8730e4ec5d9cd49f1a7ada1bb21242d0f48820d91f2ed6a357ccdc4029de556

Observation ed34658a-2942-4f20-8c00-76aefe254146 · inbound

DocOCR-Eval: A Correction-Based Framework for OCR Tool Selection Without Ground Truth cites this paper.

DocOCR-Eval: A Correction-Based Framework for OCR Tool Selection Without Ground Truth A Token-level Text Image Foundation Model for Document Understanding

Reference 227

Resolution
unresolved
no resolver link, observed 2026-08-02T14:55:43.470957Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T14:55:43.470957Z digest=sha256:08de276a966fb8026bc0f888ec56f5998f67d90c3b8fe449a1d1e1c7ff2e095a

Observation 21de2478-3a78-4df6-af50-ec66cee72b28 · inbound

SCOPE and SCION: A Benchmark and an Auditable Reference Pipeline for Schema Induction and Fusion from Text cites this paper.

SCOPE and SCION: A Benchmark and an Auditable Reference Pipeline for Schema Induction and Fusion from Text A Token-level Text Image Foundation Model for Document Understanding

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-02T13:36:56.239734Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T13:36:56.239734Z digest=sha256:3872ac12c0de1234d450a95a0ca246ea5a1b47ace38ae76f83932b52442f8f8e