Pith. sign in

Paper Citation Record · LEDGER

OmniDocBench: Benchmarking Diverse PDF Document Parsing with Comprehensive Annotations

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 18 inbound Pith citation observations for arXiv:2412.07626.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.07626 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 18 of 18 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 18 of 18 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T23:35:56.618177Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T09:59:45.169409Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 1d0b8330-df9f-4d00-9d00-417141a413fa · inbound

Document Parsing Unveiled: Techniques, Challenges, and Prospects for Structured Information Extraction cites this paper.

Document Parsing Unveiled: Techniques, Challenges, and Prospects for Structured Information Extraction OmniDocBench: Benchmarking Diverse PDF Document Parsing with Comprehensive Annotations

Reference 174

Resolution
verified exact
arxiv_id, observed 2026-05-23T19:15:47.146781Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-23T19:15:21.695801Z digest=sha256:65a271ae003402a38bfce0f541a881a05a60cf3383f5c19ecce27187c1322abf

Observation 2655b100-79ee-4075-a31b-8e8eb81030d6 · inbound

OCRBench v2: An Improved Benchmark for Evaluating Large Multimodal Models on Visual Text Localization and Reasoning cites this paper.

OCRBench v2: An Improved Benchmark for Evaluating Large Multimodal Models on Visual Text Localization and Reasoning OmniDocBench: Benchmarking Diverse PDF Document Parsing with Comprehensive Annotations

Reference 25

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T20:33:26.760260Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-17T20:33:26.613927Z digest=sha256:689fb5a8d9937511dbdd4baad681087bac9f83e06e00ca3e007b7c192f268226

Observation 0a4e7827-97ca-416c-80b4-8d28d47004fc · inbound

Ask in Any Modality: A Comprehensive Survey on Multimodal Retrieval-Augmented Generation cites this paper.

Ask in Any Modality: A Comprehensive Survey on Multimodal Retrieval-Augmented Generation OmniDocBench: Benchmarking Diverse PDF Document Parsing with Comprehensive Annotations

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T23:35:56.618177Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:35:56.618177Z digest=sha256:18ea2f63f0fc6934249d6428ef840242862921312dd192e1bc9e68ec08c41fab

Observation b239c643-500f-4054-a261-c2b1bc967d0b · inbound

Qwen2.5-VL Technical Report cites this paper.

Qwen2.5-VL Technical Report OmniDocBench: Benchmarking Diverse PDF Document Parsing with Comprehensive Annotations

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-23T02:25:19.092640Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-23T02:25:04.405036Z digest=sha256:3249719a2976d3d825f7498de383f49a66048ffa949e3d1921f7104c84c7ae74

Observation 0c1ab032-5de0-4550-a6f5-1b60dc2dd58b · inbound

From LLM Reasoning to Autonomous AI Agents: A Comprehensive Review cites this paper.

From LLM Reasoning to Autonomous AI Agents: A Comprehensive Review OmniDocBench: Benchmarking Diverse PDF Document Parsing with Comprehensive Annotations

Reference 71

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T02:57:38.330816Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-15T02:57:37.873567Z digest=sha256:1b997610513cc4fed837e7e572648b0a47e6b1a22885ecb8e65b243552a99b07

Observation 1576ac8f-c64d-4cae-bb3b-ee042d68991a · inbound

OCR-Reasoning Benchmark: Unveiling the True Capabilities of MLLMs in Complex Text-Rich Image Reasoning cites this paper.

OCR-Reasoning Benchmark: Unveiling the True Capabilities of MLLMs in Complex Text-Rich Image Reasoning OmniDocBench: Benchmarking Diverse PDF Document Parsing with Comprehensive Annotations

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:19.872212Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:57:19.872212Z digest=sha256:b3d8f166079bdcb6d19a890c193714b488ef4034a0364f84959949be893af8dd

Observation 46ecc3aa-3986-4892-9349-3a8f625cc6e7 · inbound

Ming-Omni: A Unified Multimodal Model for Perception and Generation cites this paper.

Ming-Omni: A Unified Multimodal Model for Perception and Generation OmniDocBench: Benchmarking Diverse PDF Document Parsing with Comprehensive Annotations

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T04:58:09.094956Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:58:09.094956Z digest=sha256:db2145f55ddc33505eb3d69eb07a5d546b50e35e523fe420b3325cc001c922de

Observation e07cabd4-b5ef-463b-842c-e40f57f4dbcc · inbound

FinCriticalED: A Visual Benchmark for Financial Fact-Level OCR cites this paper.

FinCriticalED: A Visual Benchmark for Financial Fact-Level OCR OmniDocBench: Benchmarking Diverse PDF Document Parsing with Comprehensive Annotations

Reference 25

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T20:20:11.555432Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-17T20:19:52.701262Z digest=sha256:c5ed0ef065f298f9a1822b6f72d3d54b1ca22975723d56627068fe4ab8ec1c4c

Observation ccfe5ca1-82e5-4598-869f-c1be30e45310 · inbound

Kimi K2.5: Visual Agentic Intelligence cites this paper.

Kimi K2.5: Visual Agentic Intelligence OmniDocBench: Benchmarking Diverse PDF Document Parsing with Comprehensive Annotations

Reference 45

Resolution
verified exact
arxiv_id, observed 2026-05-10T16:09:05.347444Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T16:09:05.225767Z digest=sha256:17e10296b87d225826b1739f5c44020d702cf1599b5f500964e0d90ed5017638

Observation b9d7b485-da8f-41f9-b5bb-5673b97830e4 · inbound

Real5-OmniDocBench: A Full-Scale Physical Reconstruction Benchmark for Robust Document Parsing in the Wild cites this paper.

Real5-OmniDocBench: A Full-Scale Physical Reconstruction Benchmark for Robust Document Parsing in the Wild OmniDocBench: Benchmarking Diverse PDF Document Parsing with Comprehensive Annotations

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-02T18:54:54.719570Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:54:54.719570Z digest=sha256:0fcfe8fc320f194b32f0b31640cc57213414d2ce0ff201c16bc0c3b390f0c3ae

Observation 25a464e7-47f8-4242-ba21-fdfb3a196fd4 · inbound

MedStruct-S: A Benchmark for Key Discovery, Key-Conditioned QA and Semi-Structured Extraction from OCR Clinical Reports cites this paper.

MedStruct-S: A Benchmark for Key Discovery, Key-Conditioned QA and Semi-Structured Extraction from OCR Clinical Reports OmniDocBench: Benchmarking Diverse PDF Document Parsing with Comprehensive Annotations

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-09T06:45:40.868022Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-08T18:10:43.434573Z digest=sha256:eced0380a731bfaa73fed469f04b2a8e1a36e8ce9932418ffc1ab79ba2ceec4f

Observation 4778ce9a-db83-40b5-bd77-5309d39f0ec8 · inbound

Structured Layout Priors for Robust Out-of-Distribution Visual Document Understanding cites this paper.

Structured Layout Priors for Robust Out-of-Distribution Visual Document Understanding OmniDocBench: Benchmarking Diverse PDF Document Parsing with Comprehensive Annotations

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-05-20T05:53:05.075400Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-20T05:48:34.771799Z digest=sha256:5331255e76a1797131e069d3f3a3d4849c5b6c901887c8db2ee5bff2c072ee82

Observation 104f0dd6-2d59-4a73-8732-d19bf6b2e051 · inbound

ABot-OCR Technical Report cites this paper.

ABot-OCR Technical Report OmniDocBench: Benchmarking Diverse PDF Document Parsing with Comprehensive Annotations

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-06-29T13:33:28.198034Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-29T13:29:17.221676Z digest=sha256:2ff450d8314896be3092bf42ff59767e7b3d8b4dfc60d5a5318c9b38fa80f9cc

Observation a912d313-509b-4f85-9375-4b3397adad2a · inbound

RealDocBench: A Benchmark for Field-Level QA and Layout Understanding on Real-World Regulated Documents cites this paper.

RealDocBench: A Benchmark for Field-Level QA and Layout Understanding on Real-World Regulated Documents OmniDocBench: Benchmarking Diverse PDF Document Parsing with Comprehensive Annotations

Reference 18

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T17:17:14.892915Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-27T22:08:16.858006Z digest=sha256:d700a9f716648709e9b70fe560d49c1945b09396f39c6e1dcd47656198c4e4da

Observation 24fa7979-b727-4abd-8b14-d36e5a3549b6 · inbound

The Stanford EDGAR Filings Dataset: Reconstructing U.S. Corporate and Financial Disclosures into Layout-Faithful and Token-Efficient Pretraining Data cites this paper.

The Stanford EDGAR Filings Dataset: Reconstructing U.S. Corporate and Financial Disclosures into Layout-Faithful and Token-Efficient Pretraining Data OmniDocBench: Benchmarking Diverse PDF Document Parsing with Comprehensive Annotations

Reference 18

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T21:18:59.442364Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-27T00:42:40.477615Z digest=sha256:d76fb3a9119c8f2e257231bf0456cded26dc401fadba6377face06e4a5a01c24

Observation 4f46e2e2-d0f1-4e2d-ad3b-0e2d942feaf4 · inbound

RT-DocLayout: Real-Time End-to-End Document Layout Analysis with Reading Order in the Wild cites this paper.

RT-DocLayout: Real-Time End-to-End Document Layout Analysis with Reading Order in the Wild OmniDocBench: Benchmarking Diverse PDF Document Parsing with Comprehensive Annotations

Reference 25

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T09:59:45.170867Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-26T09:19:38.285839Z digest=sha256:2465926f3cfbad17c3d4c2668b0c9e15accffe2e6f7d4ea9d6d555373c196ad5

Observation 61d43482-cf0f-4dea-948f-4b7d73d127c5 · inbound

Heterogeneous Element-Aware Cross-Version Differencing of Scientific Documents via Layout-Aware Alignment and Structure-Aware Reasoning cites this paper.

Heterogeneous Element-Aware Cross-Version Differencing of Scientific Documents via Layout-Aware Alignment and Structure-Aware Reasoning OmniDocBench: Benchmarking Diverse PDF Document Parsing with Comprehensive Annotations

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-02T14:39:21.796004Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T14:39:21.796004Z digest=sha256:83429f3cc20d9f7414399a632a69492f0ddf3900a57181a583214b6573281c91

Observation a3f0029a-ef24-433e-99c7-81542671c287 · inbound

LayoutLite: Token-Level Implicit Layout Analysis for Efficient Document OCR cites this paper.

LayoutLite: Token-Level Implicit Layout Analysis for Efficient Document OCR OmniDocBench: Benchmarking Diverse PDF Document Parsing with Comprehensive Annotations

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-01T05:34:45.592429Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T05:34:45.592429Z digest=sha256:b1ca08a7a8ebc674db6aa2c345b268df2deffc0191f0473219559fb8a0f1ac42