Pith. sign in

Paper Citation Record · LEDGER

OmniDocBench: Benchmarking Diverse PDF Document Parsing with Comprehensive Annotations

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 18 inbound Pith citation observations for arXiv:2412.07626.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.07626 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 18 of 18 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 18 of 18 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T23:35:56.618177Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T09:59:45.169409Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 1d0b8330-df9f-4d00-9d00-417141a413fa · inbound

Document Parsing Unveiled: Techniques, Challenges, and Prospects for Structured Information Extraction cites this paper.

Document Parsing Unveiled: Techniques, Challenges, and Prospects for Structured Information Extraction OmniDocBench: Benchmarking Diverse PDF Document Parsing with Comprehensive Annotations

Reference 174

Resolution
verified exact
arxiv_id, observed 2026-05-23T19:15:47.146781Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-23T19:15:21.695801Z digest=sha256:7b52ab448494fa2fad0259e1b8e6a015d6f046e8dd2b6c190b6053ccf2900807

Observation 2655b100-79ee-4075-a31b-8e8eb81030d6 · inbound

OCRBench v2: An Improved Benchmark for Evaluating Large Multimodal Models on Visual Text Localization and Reasoning cites this paper.

OCRBench v2: An Improved Benchmark for Evaluating Large Multimodal Models on Visual Text Localization and Reasoning OmniDocBench: Benchmarking Diverse PDF Document Parsing with Comprehensive Annotations

Reference 25

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T20:33:26.760260Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-17T20:33:26.613927Z digest=sha256:d2c89468df67dbb787c9f2df747b7e644c206c6b7e0570a00ec33ad1dfa29a01

Observation 0a4e7827-97ca-416c-80b4-8d28d47004fc · inbound

Ask in Any Modality: A Comprehensive Survey on Multimodal Retrieval-Augmented Generation cites this paper.

Ask in Any Modality: A Comprehensive Survey on Multimodal Retrieval-Augmented Generation OmniDocBench: Benchmarking Diverse PDF Document Parsing with Comprehensive Annotations

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T23:35:56.618177Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:35:56.618177Z digest=sha256:bf45ff974990fa114e6359cd5bc429140e826164f1bb93dca3fc3614c0a74f42

Observation b239c643-500f-4054-a261-c2b1bc967d0b · inbound

Qwen2.5-VL Technical Report cites this paper.

Qwen2.5-VL Technical Report OmniDocBench: Benchmarking Diverse PDF Document Parsing with Comprehensive Annotations

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-23T02:25:19.092640Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-23T02:25:04.405036Z digest=sha256:1be793cf4aaacf761998e801a6d65f572268075ff7fe87d0dd0e4d0a9fee1cc7

Observation 0c1ab032-5de0-4550-a6f5-1b60dc2dd58b · inbound

From LLM Reasoning to Autonomous AI Agents: A Comprehensive Review cites this paper.

From LLM Reasoning to Autonomous AI Agents: A Comprehensive Review OmniDocBench: Benchmarking Diverse PDF Document Parsing with Comprehensive Annotations

Reference 71

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T02:57:38.330816Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T02:57:37.873567Z digest=sha256:137e863df5fec471890977a18c628ae29bb84c27d620a522b7959c2ecc85195a

Observation 1576ac8f-c64d-4cae-bb3b-ee042d68991a · inbound

OCR-Reasoning Benchmark: Unveiling the True Capabilities of MLLMs in Complex Text-Rich Image Reasoning cites this paper.

OCR-Reasoning Benchmark: Unveiling the True Capabilities of MLLMs in Complex Text-Rich Image Reasoning OmniDocBench: Benchmarking Diverse PDF Document Parsing with Comprehensive Annotations

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:19.872212Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:57:19.872212Z digest=sha256:b3d8f166079bdcb6d19a890c193714b488ef4034a0364f84959949be893af8dd

Observation 46ecc3aa-3986-4892-9349-3a8f625cc6e7 · inbound

Ming-Omni: A Unified Multimodal Model for Perception and Generation cites this paper.

Ming-Omni: A Unified Multimodal Model for Perception and Generation OmniDocBench: Benchmarking Diverse PDF Document Parsing with Comprehensive Annotations

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T04:58:09.094956Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:58:09.094956Z digest=sha256:db2145f55ddc33505eb3d69eb07a5d546b50e35e523fe420b3325cc001c922de

Observation e07cabd4-b5ef-463b-842c-e40f57f4dbcc · inbound

FinCriticalED: A Visual Benchmark for Financial Fact-Level OCR cites this paper.

FinCriticalED: A Visual Benchmark for Financial Fact-Level OCR OmniDocBench: Benchmarking Diverse PDF Document Parsing with Comprehensive Annotations

Reference 25

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T20:20:11.555432Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-17T20:19:52.701262Z digest=sha256:b0fdf4b40c3531b397495054cdaf5a6c84b55346e9bda50beba810b79bd622a2

Observation ccfe5ca1-82e5-4598-869f-c1be30e45310 · inbound

Kimi K2.5: Visual Agentic Intelligence cites this paper.

Kimi K2.5: Visual Agentic Intelligence OmniDocBench: Benchmarking Diverse PDF Document Parsing with Comprehensive Annotations

Reference 45

Resolution
verified exact
arxiv_id, observed 2026-05-10T16:09:05.347444Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T16:09:05.225767Z digest=sha256:c063ecb9a2cbf16aabb316d207f9b1fcd17aa2515789b51f737ce4c94d6cc352

Observation b9d7b485-da8f-41f9-b5bb-5673b97830e4 · inbound

Real5-OmniDocBench: A Full-Scale Physical Reconstruction Benchmark for Robust Document Parsing in the Wild cites this paper.

Real5-OmniDocBench: A Full-Scale Physical Reconstruction Benchmark for Robust Document Parsing in the Wild OmniDocBench: Benchmarking Diverse PDF Document Parsing with Comprehensive Annotations

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-02T18:54:54.719570Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:54:54.719570Z digest=sha256:0fcfe8fc320f194b32f0b31640cc57213414d2ce0ff201c16bc0c3b390f0c3ae

Observation 25a464e7-47f8-4242-ba21-fdfb3a196fd4 · inbound

MedStruct-S: A Benchmark for Key Discovery, Key-Conditioned QA and Semi-Structured Extraction from OCR Clinical Reports cites this paper.

MedStruct-S: A Benchmark for Key Discovery, Key-Conditioned QA and Semi-Structured Extraction from OCR Clinical Reports OmniDocBench: Benchmarking Diverse PDF Document Parsing with Comprehensive Annotations

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-09T06:45:40.868022Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-08T18:10:43.434573Z digest=sha256:503aa647e9d4ab607856ecf338e5cab71f717f4af198baabc42c457929f95922

Observation 4778ce9a-db83-40b5-bd77-5309d39f0ec8 · inbound

Structured Layout Priors for Robust Out-of-Distribution Visual Document Understanding cites this paper.

Structured Layout Priors for Robust Out-of-Distribution Visual Document Understanding OmniDocBench: Benchmarking Diverse PDF Document Parsing with Comprehensive Annotations

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-05-20T05:53:05.075400Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-20T05:48:34.771799Z digest=sha256:e7b9275378b17919a657feae534603dda2b34b5a82a42c2f47906e054142d70d

Observation 104f0dd6-2d59-4a73-8732-d19bf6b2e051 · inbound

ABot-OCR Technical Report cites this paper.

ABot-OCR Technical Report OmniDocBench: Benchmarking Diverse PDF Document Parsing with Comprehensive Annotations

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-06-29T13:33:28.198034Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-29T13:29:17.221676Z digest=sha256:960bcbb8f9046743424bbfa6398e26d07b6ae1f4f03911a764679e177acd3603

Observation a912d313-509b-4f85-9375-4b3397adad2a · inbound

RealDocBench: A Benchmark for Field-Level QA and Layout Understanding on Real-World Regulated Documents cites this paper.

RealDocBench: A Benchmark for Field-Level QA and Layout Understanding on Real-World Regulated Documents OmniDocBench: Benchmarking Diverse PDF Document Parsing with Comprehensive Annotations

Reference 18

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T17:17:14.892915Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-27T22:08:16.858006Z digest=sha256:8be658a89c4dd84e3214efa340eb456ff5653730147dec5d986b9334519002bb

Observation 24fa7979-b727-4abd-8b14-d36e5a3549b6 · inbound

The Stanford EDGAR Filings Dataset: Reconstructing U.S. Corporate and Financial Disclosures into Layout-Faithful and Token-Efficient Pretraining Data cites this paper.

The Stanford EDGAR Filings Dataset: Reconstructing U.S. Corporate and Financial Disclosures into Layout-Faithful and Token-Efficient Pretraining Data OmniDocBench: Benchmarking Diverse PDF Document Parsing with Comprehensive Annotations

Reference 18

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T21:18:59.442364Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-27T00:42:40.477615Z digest=sha256:238fa256b9ca3c39c4bea6d6b7cdc1846fa9d4d9979b056ba18a126fe99c9927

Observation 4f46e2e2-d0f1-4e2d-ad3b-0e2d942feaf4 · inbound

RT-DocLayout: Real-Time End-to-End Document Layout Analysis with Reading Order in the Wild cites this paper.

RT-DocLayout: Real-Time End-to-End Document Layout Analysis with Reading Order in the Wild OmniDocBench: Benchmarking Diverse PDF Document Parsing with Comprehensive Annotations

Reference 25

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T09:59:45.170867Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-26T09:19:38.285839Z digest=sha256:9edd6eb50af31b0916dff790637d7ea35ad43999c785604b173fcc12e9a24e3e

Observation 61d43482-cf0f-4dea-948f-4b7d73d127c5 · inbound

Heterogeneous Element-Aware Cross-Version Differencing of Scientific Documents via Layout-Aware Alignment and Structure-Aware Reasoning cites this paper.

Heterogeneous Element-Aware Cross-Version Differencing of Scientific Documents via Layout-Aware Alignment and Structure-Aware Reasoning OmniDocBench: Benchmarking Diverse PDF Document Parsing with Comprehensive Annotations

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-02T14:39:21.796004Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T14:39:21.796004Z digest=sha256:83429f3cc20d9f7414399a632a69492f0ddf3900a57181a583214b6573281c91

Observation a3f0029a-ef24-433e-99c7-81542671c287 · inbound

LayoutLite: Token-Level Implicit Layout Analysis for Efficient Document OCR cites this paper.

LayoutLite: Token-Level Implicit Layout Analysis for Efficient Document OCR OmniDocBench: Benchmarking Diverse PDF Document Parsing with Comprehensive Annotations

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-01T05:34:45.592429Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T05:34:45.592429Z digest=sha256:b1ca08a7a8ebc674db6aa2c345b268df2deffc0191f0473219559fb8a0f1ac42