Pith. sign in

Paper Citation Record · LEDGER

LayoutLMv3: Pre-training for Document AI with Unified Text and Image Masking

As of 19 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 16 inbound Pith citation observations for arXiv:2204.08387.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2204.08387 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 16 of 16 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 16 of 16 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T23:59:04.212299Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

31
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 4d97eea6-7760-4b16-95c6-70844403d09f · inbound

Nougat: Neural Optical Understanding for Academic Documents cites this paper.

Nougat: Neural Optical Understanding for Academic Documents LayoutLMv3: Pre-training for Document AI with Unified Text and Image Masking

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-16T09:42:12.591250Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-16T09:42:12.463309Z digest=sha256:bcdacf4a5b23c846f42aa5cfb5c030d835a51acd74d414f6e7463a1d22338815

Observation f0e9cf88-61f0-4e5c-aec6-84e7e3da4ffd · inbound

DocFusion: A Unified Framework for Document Parsing Tasks cites this paper.

DocFusion: A Unified Framework for Document Parsing Tasks LayoutLMv3: Pre-training for Document AI with Unified Text and Image Masking

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-11T14:04:51.738129Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:04:51.738129Z digest=sha256:e6a3e4a2971fe0aaa39d052c506844438cdab644b073b102aa3e0fd21cfc076c

Observation 921d78d4-9d5c-4ed5-af6c-1612d4efa661 · inbound

Survey on Question Answering over Visually Rich Documents: Methods, Challenges, and Trends cites this paper.

Survey on Question Answering over Visually Rich Documents: Methods, Challenges, and Trends LayoutLMv3: Pre-training for Document AI with Unified Text and Image Masking

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-10T22:17:27.678099Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T22:17:27.678099Z digest=sha256:7b9807d702145c65e46324075fa03a30390520eba40054b04903dc145a525bd2

Observation c11ec639-9821-4c69-8ae9-0726ac8c1cb7 · inbound

Granite Vision: a lightweight, open-source multimodal model for enterprise Intelligence cites this paper.

Granite Vision: a lightweight, open-source multimodal model for enterprise Intelligence LayoutLMv3: Pre-training for Document AI with Unified Text and Image Masking

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T20:06:00.756271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T20:06:00.756271Z digest=sha256:9ab2f5a1dc22080078e3cfe611ccc8654fc7bd4f29f9adb09dc2c427ba36fd84

Observation caeec3cf-8475-477a-85e8-49535a9cea2d · inbound

DocSpiral: A Platform for Integrated Assistive Document Annotation through Human-in-the-Spiral cites this paper.

DocSpiral: A Platform for Integrated Assistive Document Annotation through Human-in-the-Spiral LayoutLMv3: Pre-training for Document AI with Unified Text and Image Masking

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-15T23:59:04.212299Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T23:59:04.212299Z digest=sha256:db9cdcbf3d57f08bfc810dda5c668ff69261354a03583cddc10e2bc26121a1cb

Observation d306081e-cbcd-4759-8c8d-eeb26ad9708d · inbound

Lost in OCR Translation? Vision-Based Approaches to Robust Document Retrieval cites this paper.

Lost in OCR Translation? Vision-Based Approaches to Robust Document Retrieval LayoutLMv3: Pre-training for Document AI with Unified Text and Image Masking

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T23:03:10.410340Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:03:10.410340Z digest=sha256:16d75df5950e1f59eb38433a6815fa3bef596b64fc3e0980dc74afb5bd3cb8b1

Observation 6561f0ba-55d3-4927-8ac7-886e6123f8bc · inbound

Template-Based Schema Matching of Multi-Layout Tenancy Schedules:A Comparative Study of a Template-Based Hybrid Matcher and the ALITE Full Disjunction Model cites this paper.

Template-Based Schema Matching of Multi-Layout Tenancy Schedules:A Comparative Study of a Template-Based Hybrid Matcher and the ALITE Full Disjunction Model LayoutLMv3: Pre-training for Document AI with Unified Text and Image Masking

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T20:47:40.802834Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:47:40.802834Z digest=sha256:e357b90bd5a0ea00acad1e18f5fab3cbeebfc6020bea0cfe050f34eca1634bd7

Observation a9d26657-f784-4e93-8dd6-1848d3e1e81b · inbound

FRED: Financial Retrieval-Enhanced Detection and Editing of Hallucinations in Language Models cites this paper.

FRED: Financial Retrieval-Enhanced Detection and Editing of Hallucinations in Language Models LayoutLMv3: Pre-training for Document AI with Unified Text and Image Masking

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T13:10:36.384402Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T13:10:36.384402Z digest=sha256:7870c09ccf1263fc09ab2d9ba8f8823a863142294f61bb55068a07c2e9c3f7fa

Observation 1b37765a-afbd-4f95-814a-a82a19028506 · inbound

Vector embedding of multi-modal texts: a tool for discovery? cites this paper.

Vector embedding of multi-modal texts: a tool for discovery? LayoutLMv3: Pre-training for Document AI with Unified Text and Image Masking

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-04T21:05:25.170516Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T21:05:25.170516Z digest=sha256:ef1a0a2928d5bdeb316f8c73e30b0ac3bd6eed4b5b894a8898b9845f057f19cd

Observation 13a49ca6-5c57-4ccb-994e-7a477bcf66cd · inbound

Interfaze: The Future of AI is built on Task-Specific Small Models cites this paper.

Interfaze: The Future of AI is built on Task-Specific Small Models LayoutLMv3: Pre-training for Document AI with Unified Text and Image Masking

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-03T04:48:06.443689Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T04:48:06.443689Z digest=sha256:db77a13b910771f28846aed638df1d52552769c64d06ace415e8861bc2248384

Observation 4103739f-1479-4e63-adf2-4944ea4c36f4 · inbound

Improving Layout Representation Learning Across Inconsistently Annotated Datasets via Agentic Harmonization cites this paper.

Improving Layout Representation Learning Across Inconsistently Annotated Datasets via Agentic Harmonization LayoutLMv3: Pre-training for Document AI with Unified Text and Image Masking

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-11T08:50:58.269368Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-10T16:28:16.315767Z digest=sha256:4c390ae71b3af0c6cc8ce2b9dac9357ae39fa09f297773bb152dc40f3bb5f2c0

Observation 0626d58a-5ad4-4d49-a335-5774dcad8752 · inbound

Multimodal Approaches for Visually-Rich Document Type Classification: A Comparative Analysis cites this paper.

Multimodal Approaches for Visually-Rich Document Type Classification: A Comparative Analysis LayoutLMv3: Pre-training for Document AI with Unified Text and Image Masking

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-07-01T22:46:19.633352Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T15:03:16.112327Z digest=sha256:da4de9158572a199a61474394d908e770ba358a4dd201ee4d05db47d8aa2309e

Observation c325b111-1cd4-4337-ac24-e9bc7c863320 · inbound

ReforMe: Re-Shaping Documents with Contextual Prompting and Layout-Aware Propagation cites this paper.

ReforMe: Re-Shaping Documents with Contextual Prompting and Layout-Aware Propagation LayoutLMv3: Pre-training for Document AI with Unified Text and Image Masking

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-07-02T04:56:38.959909Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T08:43:27.671508Z digest=sha256:a8636c6efa3cd8a19995dee58343cd17a49b918e21233ae9ebda0e9aa2b41cf2

Observation 97825941-3fd8-439f-b3c0-8f51a614a3ce · inbound

RT-DocLayout: Real-Time End-to-End Document Layout Analysis with Reading Order in the Wild cites this paper.

RT-DocLayout: Real-Time End-to-End Document Layout Analysis with Reading Order in the Wild LayoutLMv3: Pre-training for Document AI with Unified Text and Image Masking

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-07-04T09:59:45.195190Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-26T09:19:38.285839Z digest=sha256:d5a7f19ecd4fc1246ccd8abb58f0485b8524d088bb3489ff8d68bd4c783ac0e4

Observation c82d3cd1-571a-4e75-bd28-b51d2340ede5 · inbound

Structure-Preserving Document Translation via Multi-Stage LLM Pipeline: A Case Study in Marathi cites this paper.

Structure-Preserving Document Translation via Multi-Stage LLM Pipeline: A Case Study in Marathi LayoutLMv3: Pre-training for Document AI with Unified Text and Image Masking

Reference 29

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T10:04:35.285146Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-06-30T09:59:20.602618Z digest=sha256:63fa7b865a63da188b74e87913a817364fac1e56117732f4b647851e875cf3a3

Observation d6fa453b-1ee5-416b-bf06-23ccc9d25c9e · inbound

IntelliAudit: Using Large Language Models to Evaluate Audit Controls cites this paper.

IntelliAudit: Using Large Language Models to Evaluate Audit Controls LayoutLMv3: Pre-training for Document AI with Unified Text and Image Masking

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-11T00:27:29.113753Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:27:29.113753Z digest=sha256:63b59220495354b2efdec6c05298ec69cccd1e6d971f0663ce8ff31a70d3753e