Pith. sign in

Paper Citation Record · LEDGER

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding

As of 17 August 2026, this Paper Citation Record lists 59 of 59 outbound references and 11 inbound Pith citation observations for arXiv:2601.20430.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2601.20430 v2

Coverage vector

measured 59 of 59 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T15:44:17.303248Z

measured 70 of 70 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 11 of 11 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T21:04:30.042864Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T16:49:57.063058Z

Reference resolution

59 of 59 outbound references displayed

  • verified exact0
  • verified fuzzy10
  • unresolved49
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c3c7fcfe-c610-41ab-8f3a-4820dd6cf184 · outbound

This paper cites Document Parsing Unveiled: Techniques, Challenges, and Prospects for Structured Information Extraction.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding Document Parsing Unveiled: Techniques, Challenges, and Prospects for Structured Information Extraction

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:16.966655Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:16.966655Z digest=sha256:d2ffe4e78a554e95d32b515749fe72d016e0f8b80a9f5c8d382034e8ae7641fb

Observation 28aecddb-10c8-4960-a858-eb33f6df49bb · outbound

This paper cites Dolphin: Document Image Parsing via Heterogeneous Anchor Prompting.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding Dolphin: Document Image Parsing via Heterogeneous Anchor Prompting

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:16.973324Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:16.973324Z digest=sha256:86c2298c60fd7aca1b23e51258ef4aeb262ef041b089027d3fd5287c6d1063c8

Observation 2b1f6253-d19b-407b-aaa4-221515c6d48a · outbound

This paper cites Monkeyocr: Document parsing with a structure-recognition-relation triplet paradigm.arXiv preprint arXiv:2506.05218, 2025.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding Monkeyocr: Document parsing with a structure-recognition-relation triplet paradigm.arXiv preprint arXiv:2506.05218, 2025

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:16.978948Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:16.978948Z digest=sha256:9f4d04c9b6385ebf2696bb81fef4d55b5e5ad234d0d0505e62903273527b2dbd

Observation 38bc556d-76f8-4df6-bfcf-fa4f878dfe24 · outbound

This paper cites MinerU: An Open-Source Solution for Precise Document Content Extraction.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding MinerU: An Open-Source Solution for Precise Document Content Extraction

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:16.984659Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:16.984659Z digest=sha256:98d50519dbcd5a19c1242cd96e579da623bd96f44b2dc768887f3b80cf0ec209

Observation fbd59e9a-c7e3-48f3-87a6-bd58f72aaca8 · outbound

This paper cites olmocr: Unlocking trillions of tokens in pdfs with vision language models.arXiv preprint arXiv:2502.18443, 2025.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding olmocr: Unlocking trillions of tokens in pdfs with vision language models.arXiv preprint arXiv:2502.18443, 2025

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:16.990144Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:16.990144Z digest=sha256:12671e11a7531a071547001b85b2fb7b4e2045c383fc99321d461de0361ec8d5

Observation 4fb521d5-2558-4cef-b695-bda64cd4e81d · outbound

This paper cites an unresolved cited work.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding Unresolved cited work

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:16.995462Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:16.995462Z digest=sha256:778cf5e69cbfd93e3be015cd3c799fa276efa7d64cc49d2dd0d3422e84ff8aec

Observation 65178962-49ea-4845-b8ee-059fc51e11d0 · outbound

This paper cites MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:17.001936Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:17.001936Z digest=sha256:795c4b8d3494327bd208ff64b35ec9f1e3cdaca410b6d842ccb8a04d09664e61

Observation 81b97e13-f7cb-4139-a7dd-3ea4ed262be7 · outbound

This paper cites Paddleocr-vl: Boosting multilingual document parsing via a 0.9 b ultra-compact vision-language model.arXiv preprint arXiv:2510.14528, 2025.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding Paddleocr-vl: Boosting multilingual document parsing via a 0.9 b ultra-compact vision-language model.arXiv preprint arXiv:2510.14528, 2025

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:17.008593Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:17.008593Z digest=sha256:d8fde90e4885ce02bb7528350a66e05028ea918f98e9f803ac7da2f213952bd8

Observation 51db00ed-ed1d-40a9-8942-9d9547d83dc7 · outbound

This paper cites Language models are few-shot learners.Advances in neural information processing systems, 33:1877–1901, 2020.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding Language models are few-shot learners.Advances in neural information processing systems, 33:1877–1901, 2020

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:17.013892Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:17.013892Z digest=sha256:0db0325f130cabbc8ffde21f826b9e8774120f1faed17304af36e8aa22e0c105

Observation 6407f825-1915-4392-a1de-1b8e9630142b · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding LLaMA: Open and Efficient Foundation Language Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:17.018994Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:17.018994Z digest=sha256:c7de7d207cb7ee536b04dff0f8b70d7d24401833e5b95dfb8140f4bb587d3a55

Observation f82d6265-c768-4184-a7f5-7700df801663 · outbound

This paper cites GPT-4 Technical Report.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding GPT-4 Technical Report

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:17.025501Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:17.025501Z digest=sha256:3f3275bf6c2ccdac5ceedd991b88a1770080126c6e5602b849bd6a4b049d5be6

Observation c5043f4b-ba09-4296-bdae-802a4c47b9cc · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding Gemini: A Family of Highly Capable Multimodal Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:17.031620Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:17.031620Z digest=sha256:b5d85bad6e513ea41a90c0c5a334ef134c5f5762fcb78e3e988881f6af93722d

Observation 800a5e75-8673-4c26-9d98-870939a307e0 · outbound

This paper cites Mistral 7B.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding Mistral 7B

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:17.038573Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:17.038573Z digest=sha256:d6d5f623633e1e538f3dd462c8fffa3d0a66af5bf2b1dade420e0fbb72182f80

Observation e0258661-845a-4506-8560-49e6f5284cef · outbound

This paper cites Learning transferable visual models from natural language supervision.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding Learning transferable visual models from natural language supervision

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:17.044069Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:17.044069Z digest=sha256:b4c623e7c3332762f1e6f6399551f21b21e3fc9bf42b529ea8ea4a0924a5114a

Observation ca900daa-7c5c-4724-b9db-1b983b302a36 · outbound

This paper cites Visual instruction tuning.Advances in neural information processing systems, 36:34892–34916, 2023.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding Visual instruction tuning.Advances in neural information processing systems, 36:34892–34916, 2023

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:17.049410Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:17.049410Z digest=sha256:3f6408516f282c469252aa12d00de912fb4f04692efe53a2d7fdb064eb9fb194

Observation 1c583a0e-1b4d-4f34-9e07-a2e25a7241fd · outbound

This paper cites MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:17.056152Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:17.056152Z digest=sha256:def5364fa4f2228e5a328c1fc509c1a5c73fff282e59478fbea34c134b40fe8c

Observation 14265ea9-2609-4393-8036-c816c782a320 · outbound

This paper cites Instructblip: Towards general-purpose vision-language models with instruction tuning.Advances in neural information processing systems, 36:49250–49267, 2023.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding Instructblip: Towards general-purpose vision-language models with instruction tuning.Advances in neural information processing systems, 36:49250–49267, 2023

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:17.062657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:17.062657Z digest=sha256:e9ab476d1f84ac183c9d41851c750ca1d4360748e77c88504c9ef7feffe9c43b

Observation 275bff2b-0528-4d99-a769-5016c2a45a4c · outbound

This paper cites SigLIP 2: Multilingual Vision-Language Encoders with Improved Semantic Understanding, Localization, and Dense Features.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding SigLIP 2: Multilingual Vision-Language Encoders with Improved Semantic Understanding, Localization, and Dense Features

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:17.068839Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:17.068839Z digest=sha256:f42b646e7bf9183d3060b94790bc647ed08973bb756c004f7d43a1a5812ecc3f

Observation 8fdd8404-4467-4ca7-9efb-94bc36cf75f3 · outbound

This paper cites PaddleOCR 3.0 Technical Report.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding PaddleOCR 3.0 Technical Report

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:17.075783Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:17.075783Z digest=sha256:59c2b740e23de94d8e93d2bb3eac8bbe69b0ac105d4f113edc55a3c008d5ea0e

Observation abf9538b-1101-4abd-96e1-a9471f9bf8e9 · outbound

This paper cites Docling: An Efficient Open-Source Toolkit for AI-driven Document Conversion.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding Docling: An Efficient Open-Source Toolkit for AI-driven Document Conversion

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:17.083690Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:17.083690Z digest=sha256:45f232eaf979e79920bd43f3f49d84015031dab9ee587f26c46c04b88f68aa5b

Observation 021ed39c-3c1f-4c42-8cc1-66a02c9bd9d1 · outbound

This paper cites DeepSeek-OCR: Contexts Optical Compression.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding DeepSeek-OCR: Contexts Optical Compression

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:17.088989Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:17.088989Z digest=sha256:a7f61135233fcdbd40c77dfba95ef5a91c6a4537a1c5f41b218f08214c0ce5fa

Observation 265ca862-3f02-4b68-841b-abaf3cf10b11 · outbound

This paper cites Learning to Recover from Multi-Modality Errors for Non-Autoregressive Neural Machine Translation.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding Learning to Recover from Multi-Modality Errors for Non-Autoregressive Neural Machine Translation

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:17.094994Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:17.094994Z digest=sha256:fdc4f17459a9492886edeff98d2c5301d462d741e2de056d39a031ee126c8373

Observation 00b3ea6c-2fb6-4c2a-b774-495dd3199179 · outbound

This paper cites Maskgit: Masked generative image transformer.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding Maskgit: Masked generative image transformer

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:44:18.902835Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T15:44:17.099950Z digest=sha256:c91dd62f20cd6be73dc03c3ba5dc5f2c71515854a753dc6188893c53e9bb555d

Observation 350c4962-09c2-4acf-8802-8d1c8442f4bd · outbound

This paper cites Medusa: Simple LLM Inference Acceleration Framework with Multiple Decoding Heads.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding Medusa: Simple LLM Inference Acceleration Framework with Multiple Decoding Heads

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:17.104779Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:17.104779Z digest=sha256:a465ffdff879ef16580a6667c0ed3f02a78171b30bf9f51a7b96f32548dabac1

Observation 9057952b-921e-4f8e-ac46-f468fe359c36 · outbound

This paper cites Youtu-vl: Unleashing visual potential via unified vision-language supervision.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding Youtu-vl: Unleashing visual potential via unified vision-language supervision

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:17.109878Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:17.109878Z digest=sha256:cfbb0b7c441632ddf1c913cb5840707befba00aa5d162b0d27284c9211c78ca8

Observation 5c4cca28-a230-4bb3-a035-9d3ad42778c6 · outbound

This paper cites Youtu-llm: Unlocking the native agentic potential for lightweight large language models.arXiv preprint arXiv:2512.24618, 2025.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding Youtu-llm: Unlocking the native agentic potential for lightweight large language models.arXiv preprint arXiv:2512.24618, 2025

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:17.114650Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:17.114650Z digest=sha256:dbf9e2a0f0d5e1edf374c9806803608b239e25c9460509da1ce9cd6e833a730d

Observation 0ea370fd-202a-4d56-b225-e05a196ee111 · outbound

This paper cites Optimized table tokenization for table structure recognition.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding Optimized table tokenization for table structure recognition

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:17.119434Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:17.119434Z digest=sha256:41de7c1cd38de96580a344f358c904d5f52192a645b8a6194b577d31dd06027a

Observation 7db0fe3b-5fcf-4d01-aa20-ee8c4d0f55e0 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:17.124389Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:17.124389Z digest=sha256:641c1aa4604544f668354a00d7ba9ef8c9c80e6962be592bbb14eddae43b7c31

Observation a62de1df-af5b-440c-b49a-a0d43f451c13 · outbound

This paper cites Omnidocbench: Benchmarking diverse pdf document parsing with comprehensive annotations.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding Omnidocbench: Benchmarking diverse pdf document parsing with comprehensive annotations

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:17.129841Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:17.129841Z digest=sha256:014a688d7332d0c19da79860d4ef0e7f55ca5370fdc783d456847841a9b809a8

Observation 398feac7-ff3d-45e3-8f67-2da3619ec8de · outbound

This paper cites Marker.https://github.com/datalab-to/marker, 2025.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding Marker.https://github.com/datalab-to/marker, 2025

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:44:18.846582Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T15:44:17.134934Z digest=sha256:8cdae8c9e07389d9e253b33389a6dcfc5dac08cbad0497be52ad071455c1727b

Observation cfb6dc25-7455-4cb2-b3a5-8a114ff93372 · outbound

This paper cites GPT-4o System Card.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding GPT-4o System Card

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:17.140395Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:17.140395Z digest=sha256:90e529ab77b01a6a20884c16faae5542fd9a867e614e8b804b099e7eab19c2ea

Observation 913246a2-824c-438a-867a-78f96aa42f9b · outbound

This paper cites InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:17.145667Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:17.145667Z digest=sha256:ace70d59e3d9c1803bf4c05697479ac0c9d0d2bcc750a363481f085597dace32

Observation 5c7a5446-4ce2-4536-99ff-060339556d2b · outbound

This paper cites Qwen2.5-VL Technical Report.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding Qwen2.5-VL Technical Report

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:17.156679Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:17.156679Z digest=sha256:f553b56c5a11a6dbceeeeef717268f16c817311f82dfdf7d52f6832f8df41505

Observation 68e6045e-c5eb-48e3-8263-934d9cfab531 · outbound

This paper cites Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:17.162018Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:17.162018Z digest=sha256:610125f600f2085aa46e24c7789bdf74a253880bb92d0e4e1ff15592a5ae5c2b

Observation f2c7bcfa-27b1-4a28-9681-d38efc63ea09 · outbound

This paper cites Ocrflux.https://github.com/chatdoc- com/OCRFlux, 2025.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding Ocrflux.https://github.com/chatdoc- com/OCRFlux, 2025

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:44:18.828056Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T15:44:17.167479Z digest=sha256:444c3d18aaafe1c10b80c72ca417abad0d179c8649cf718882cb9e34075372f5

Observation 95e24531-f5e0-49de-b1ab-ac01fc966a01 · outbound

This paper cites Mistral-ocr.https://mistral.ai/news/mistral-ocr?utm source=ai-bot.cn, 2025.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding Mistral-ocr.https://mistral.ai/news/mistral-ocr?utm source=ai-bot.cn, 2025

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:44:18.806663Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T15:44:17.175157Z digest=sha256:ce3aedc94ff7193730cbb04c4b09f95f96e7afd011f483dff8c653b0c240072b

Observation b413d709-e95d-4c19-8996-3c5e53901baa · outbound

This paper cites Points-reader: Distillation-free adaptation of vision-language models for document conversion.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding Points-reader: Distillation-free adaptation of vision-language models for document conversion

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:44:18.788674Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T15:44:17.182344Z digest=sha256:b33a713ca0f7d86a6eb1f884a4449dbe0ae2341890e02c31d4fe10ddecfaf91a

Observation 8f6593be-c8f6-4088-a3b5-7443e6954bbb · outbound

This paper cites Nanonets-ocr-s: A model for transforming documents into structured markdown with intelligent content recognition and semantic tagging, 2025.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding Nanonets-ocr-s: A model for transforming documents into structured markdown with intelligent content recognition and semantic tagging, 2025

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:17.187713Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:17.187713Z digest=sha256:e35e06d712fff700402592a03ce5a07dcb4d6811b3beace0d753e05b50367385

Observation 584ac40b-b000-43d9-9c34-e88a801f4b7b · outbound

This paper cites Image over text: Transforming formula recognition evaluation with character detection matching.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding Image over text: Transforming formula recognition evaluation with character detection matching

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:17.192599Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:17.192599Z digest=sha256:ee8ac109a994e6b767b58fd8605d57a95b1b8a9b31f66c6466bb5dad3edeaf6b

Observation 8374201a-4916-4d89-8531-ceb6b2fd6130 · outbound

This paper cites General OCR Theory: Towards OCR-2.0 via a Unified End-to-end Model.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding General OCR Theory: Towards OCR-2.0 via a Unified End-to-end Model

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:17.198748Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:17.198748Z digest=sha256:6fd1e87a03f325c8fd9e60ce84e40b9e8469c0950ba5c04c75cd2c3f2ee1e830

Observation bfa9869d-5c7f-479b-ab76-594a6deb129d · outbound

This paper cites Google deepmind.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding Google deepmind

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:44:18.743820Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T15:44:17.204307Z digest=sha256:bf73daeb2029035d01eb3ca141e808706a72acb6669a51911474a92bcfcbc567

Observation f6f2c98a-57ea-4267-bbb0-7efbdebe247d · outbound

This paper cites Cc-ocr: A comprehensive and challenging ocr benchmark for evaluating large multimodal models in literacy.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding Cc-ocr: A comprehensive and challenging ocr benchmark for evaluating large multimodal models in literacy

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:17.209052Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:17.209052Z digest=sha256:3109dfcba7a5e66212df2490a355de684d45d36a5046419e3a5ba44c2edce154

Observation 4099a557-3283-4f1b-a3ea-7496250d29b4 · outbound

This paper cites OCRBench v2: An Improved Benchmark for Evaluating Large Multimodal Models on Visual Text Localization and Reasoning.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding OCRBench v2: An Improved Benchmark for Evaluating Large Multimodal Models on Visual Text Localization and Reasoning

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:17.214122Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:17.214122Z digest=sha256:70ddd81ca8337152224560d3ef8d2ed35f64f8efaf02186d01f2f954215f9132

Observation a96d7866-464b-4bde-99bc-dc0f3ea305eb · outbound

This paper cites Qwen3 Technical Report.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding Qwen3 Technical Report

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:17.219209Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:17.219209Z digest=sha256:053c2dbbf8f7980960448651757a347cd4d8ab9ff42c519ac046a2cc2fada036

Observation f1107f11-8a7f-4d10-bf80-4c163b796c7f · outbound

This paper cites Onechart: Purify the chart structural extraction via one auxiliary token.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding Onechart: Purify the chart structural extraction via one auxiliary token

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:44:18.712573Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T15:44:17.225007Z digest=sha256:233f96b5493e1b04db7460419e959caa55f4e688f217ef2563148577eb4f7e95

Observation 30f9e1a3-1ad4-458e-a7e2-ab5c0f8f3d14 · outbound

This paper cites Deplot: One-shot visual language reasoning by plot-to-table translation.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding Deplot: One-shot visual language reasoning by plot-to-table translation

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:17.230016Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:17.230016Z digest=sha256:8dd80cbf6f129a26ec777b90b6b62bdabd65b8095441b8e76e94073805889907

Observation 8eea3c0f-5258-4300-ad5b-07682ca6ffc8 · outbound

This paper cites DocLayout-YOLO: Enhancing Document Layout Analysis through Diverse Synthetic Data and Global-to-Local Adaptive Perception.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding DocLayout-YOLO: Enhancing Document Layout Analysis through Diverse Synthetic Data and Global-to-Local Adaptive Perception

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:17.234986Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:17.234986Z digest=sha256:bba195fc1d82b365df9a4610f988011307454395bff648c27be36b6c5d224df6

Observation a14a9b06-33ef-441e-83ad-40e11c6b45b9 · outbound

This paper cites PP-DocLayout: A Unified Document Layout Detection Model to Accelerate Large-Scale Data Construction.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding PP-DocLayout: A Unified Document Layout Detection Model to Accelerate Large-Scale Data Construction

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:17.241605Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:17.241605Z digest=sha256:7b60360052aca6cb1886c6d1514fb5bf87b5f52b84b3ccf3c9be522a861772cd

Observation 14057ee5-a041-4c7e-9f4f-bd9396fddfaa · outbound

This paper cites An end-to-end formula recognition method integrated attention mechanism.Mathematics, 11(1):177, 2022.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding An end-to-end formula recognition method integrated attention mechanism.Mathematics, 11(1):177, 2022

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:44:18.677116Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T15:44:17.247926Z digest=sha256:e1f06278b0c7ed3069747e3f972350934eebb8b7080931104a800e978a7f8b73

Observation 212bb3d0-dab3-411c-8a86-64b1792ab6bb · outbound

This paper cites Deeptabstr: Deep learning based table structure recognition.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding Deeptabstr: Deep learning based table structure recognition

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:44:18.655210Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T15:44:17.252678Z digest=sha256:01299796e0311ed225ff09032481d44824774746789740408787b79f4c989f42

Observation bd2b6cff-7e5e-4652-98ed-4fea655bfc79 · outbound

This paper cites Tsrformer: Table structure recognition with transformers.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding Tsrformer: Table structure recognition with transformers

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:44:18.636070Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T15:44:17.257609Z digest=sha256:e555b186548b6cc7fe6531f6616354991226019634db8d34cc35f0c98fc189d7

Observation b595229f-2acc-4f99-9f13-5109d8aec9c5 · outbound

This paper cites LayoutReader: Pre-training of Text and Layout for Reading Order Detection.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding LayoutReader: Pre-training of Text and Layout for Reading Order Detection

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:17.262939Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:17.262939Z digest=sha256:68e46a0ef3592380fbf102901246d181967cfb001939d89808ce0f7fe0130861

Observation 4fcce5de-9d68-4abf-8115-01f9643393c4 · outbound

This paper cites Internvl: Scaling up vision foundation models and aligning for generic visual-linguistic tasks.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding Internvl: Scaling up vision foundation models and aligning for generic visual-linguistic tasks

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:17.268807Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:17.268807Z digest=sha256:c762903f47ea3ceeeb46e4cd837c3e68c24828209f022860d340f34c7ad6f770

Observation 2e234018-d416-4158-b7d6-81971bb6d02a · outbound

This paper cites How Far Are We to GPT-4V? Closing the Gap to Commercial Multimodal Models with Open-Source Suites.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding How Far Are We to GPT-4V? Closing the Gap to Commercial Multimodal Models with Open-Source Suites

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:17.273912Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:17.273912Z digest=sha256:3ccf3e81e2525bc2871eaae83bc60604b0b2c4ac046d5c561bd04f8b74a16b1a

Observation f5e87724-d7ac-4577-a0f1-a91f5910cb38 · outbound

This paper cites Enhancing the Reasoning Ability of Multimodal Large Language Models via Mixed Preference Optimization.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding Enhancing the Reasoning Ability of Multimodal Large Language Models via Mixed Preference Optimization

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:17.281082Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:17.281082Z digest=sha256:5e47fa2eb6cf6262f8013aebfae9b2b80a38d67819d8a7a7ce0b657dc5e625b0

Observation 750da50b-33ec-435e-af67-dd38fc66f501 · outbound

This paper cites Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:17.286792Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:17.286792Z digest=sha256:c7c91ee67f04a1cef31d4dfc07fbbaea094b55158d95324ac31ef0a7acc46d9a

Observation 00200826-4b3e-483b-b981-c59b066324db · outbound

This paper cites InternVL3.5: Advancing Open-Source Multimodal Models in Versatility, Reasoning, and Efficiency.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding InternVL3.5: Advancing Open-Source Multimodal Models in Versatility, Reasoning, and Efficiency

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:17.292277Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:17.292277Z digest=sha256:1efd08cd4a758b54a586d698db2f7e3f6ea904ccff9051a18d891b2a0f2c8378

Observation 083a24fa-adc3-4c80-b1a8-0e3dde525d6b · outbound

This paper cites Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:17.297182Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:17.297182Z digest=sha256:39fda68c5fb3000ee7c90cf54df8526c8e3125f1207891d6442111df0105d560

Observation 08c702be-99a7-495c-a07a-1c7fdcfdcdce · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-15T15:44:17.303248Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:44:17.303248Z digest=sha256:cf39cce5b851f2d7bf0ef16f6fb6ed7bfd2f2698ad9952810cadc631cafb290e

Pith citing papers

Observation 61e47448-70da-49d5-b4cd-2472b57d0468 · inbound

Parser-Oriented Structural Refinement for a Stable Layout Interface in Document Parsing cites this paper.

Parser-Oriented Structural Refinement for a Stable Layout Interface in Document Parsing Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-07-14T03:21:27.323268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-13T21:07:44.073353Z digest=sha256:ca3440dda711d2ce0c5b5e6e4867bcf9b197fecf1ec54a97ced2866220df6f7b

Observation 4d1387c1-c444-482a-916d-cd38102058a3 · inbound

MinerU2.5-Pro: Pushing the Limits of Data-Centric Document Parsing at Scale cites this paper.

MinerU2.5-Pro: Pushing the Limits of Data-Centric Document Parsing at Scale Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-07-14T03:21:27.323268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-10T18:58:41.377996Z digest=sha256:753f430fa7b0bfe868a164362fc8f9283afc6446fe6484a043ba4a8c57ba7138

Observation 14021d59-1781-42f0-a013-106d9dd5b04b · inbound

How Far Is Document Parsing from Solved? PureDocBench: A Source-TraceableBenchmark across Clean, Degraded, and Real-World Settings cites this paper.

How Far Is Document Parsing from Solved? PureDocBench: A Source-TraceableBenchmark across Clean, Degraded, and Real-World Settings Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-07-14T03:21:27.323268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-11T02:28:02.152600Z digest=sha256:e72450440265e41c346a50c3f9ed4b6afcdd62c67739f554dfc118d3416a5203

Observation 612e5ec8-f733-4f38-9cd7-72c37d730f45 · inbound

MPDocBench-Parse: Benchmarking Practical Multi-page Document Parsing cites this paper.

MPDocBench-Parse: Benchmarking Practical Multi-page Document Parsing Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding

Reference 59

Resolution
verified exact
arxiv_id, observed 2026-07-14T03:21:27.323268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-22T05:58:04.055855Z digest=sha256:e43cecdb7d31da78c1b3f5cd3286e697b33a1cf0aaf595bf915d309a0a8c82c2

Observation 37772d44-a626-4a8f-a5ea-4be6664445b9 · inbound

MPDocBench-Parse: Benchmarking Practical Multi-page Document Parsing cites this paper.

MPDocBench-Parse: Benchmarking Practical Multi-page Document Parsing Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding

Reference 59

Resolution
verified exact
arxiv_id, observed 2026-07-14T03:21:27.323268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-06-30T17:37:33.750306Z digest=sha256:5fd556a095ca92467d62b6622291590fafb3f88e27133a4f7c42e4c8f9355027

Observation 762ba601-85d9-4e7e-b9f2-a7bf8e3e97f4 · inbound

ABot-OCR Technical Report cites this paper.

ABot-OCR Technical Report Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-07-14T03:21:27.323268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-06-29T13:29:17.221676Z digest=sha256:79e1374712d294a76ed31b0ddc75521dc7ce14decddb1fe056088b7dd1bbb472

Observation 4de339ab-37c2-4da1-a298-ef31081c387f · inbound

PaddleOCR-VL-1.6: Expanding the Frontier of Document Parsing with Under-Optimized Region Refinement and Progressive Post-Training cites this paper.

PaddleOCR-VL-1.6: Expanding the Frontier of Document Parsing with Under-Optimized Region Refinement and Progressive Post-Training Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-07-14T03:21:27.323268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-06-28T10:39:14.444486Z digest=sha256:efe8a8b90a726869ce2d08ca0c4c5c8f1fe8002f904be67d617f19f95e3baae5

Observation deaf8128-b58e-4a3a-8ac9-94642f1e2005 · inbound

P-MTP: Efficient Document Parsing via Multi-Token Prediction with Progressive Depth Scaling cites this paper.

P-MTP: Efficient Document Parsing via Multi-Token Prediction with Progressive Depth Scaling Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-07-14T03:21:27.323268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-06-26T00:18:07.278889Z digest=sha256:eb9d02c0b621ed49bc73ee6d05ca176ca53cb8c80929875eef2ca09b25731287

Observation 647786e6-5233-4b77-b337-d1e8c995a3c0 · inbound

OvisOCR2 Technical Report cites this paper.

OvisOCR2 Technical Report Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-02T04:40:04.566211Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T04:40:04.566211Z digest=sha256:31eb6ec51f368576f5fcff85a27d84cd12a6d41b558cf5ac4f184a52bc9a0298

Observation b96afe7f-78f9-4bf8-98cb-7c7211102b1c · inbound

HPD-Parsing: Hierarchical Parallel Document Parsing cites this paper.

HPD-Parsing: Hierarchical Parallel Document Parsing Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-01T14:16:08.107522Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:16:08.107522Z digest=sha256:3bfab7fbf55fa0cac84f23e713af00b2c0989cda4f92ca0df7c0da25fcb6689e

Observation df24d726-cb6d-4f2f-a3ab-c6b587dd6f62 · inbound

NaviDC-OCR: Navigating Document Parsing Across Digital and Camera-Captured Documents cites this paper.

NaviDC-OCR: Navigating Document Parsing Across Digital and Camera-Captured Documents Youtu-Parsing: Perception, Structuring and Recognition via High-Parallelism Decoding

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-15T21:04:30.042864Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T21:04:30.042864Z digest=sha256:f3fad466a7e6e86ef9b388a869cbf9cbbf4badf5137c101547076fb4db997899