Pith. sign in

Paper Citation Record · LEDGER

Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning

As of 17 August 2026, this Paper Citation Record lists 41 of 41 outbound references and 7 inbound Pith citation observations for arXiv:2509.09731.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2509.09731 v1

Coverage vector

measured 41 of 41 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-04T20:28:58.667157Z

measured 48 of 48 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 7 of 7 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-03T10:43:50.289411Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-30T07:44:22.118843Z

Reference resolution

41 of 41 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved41
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation f05893e5-57e7-48f1-94d7-59d5ba16ed80 · outbound

This paper cites , " * write output.state after.block = add.period write newline.

Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning , " * write output.state after.block = add.period write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-04T20:28:55.488191Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:28:55.488191Z digest=sha256:ec290022422a52b62c52fcaed35e96eb4ffd9d94519351688244974fc91ffef2

Observation 9c830c9c-e00f-480a-aa2f-9052dcf7e671 · outbound

This paper cites write newline.

Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-04T20:28:55.595187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:28:55.595187Z digest=sha256:29a85e0c772b0b94d59e1daa3e550c38e2d519200b45bb7a40e11a527db376c4

Observation 911156ce-463a-4dba-bc3f-d66365587edb · outbound

This paper cites an unresolved cited work.

Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning Unresolved cited work

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-04T20:28:55.684666Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:28:55.684666Z digest=sha256:eca41fbeea529720ae09582df1afca59d7864a1282b918722ea153f802d5c9a0

Observation 2e3ce02e-f44c-4212-b5a4-927873b0f9a2 · outbound

This paper cites Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond.

Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-04T20:28:55.790100Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:28:55.790100Z digest=sha256:aabb31421dc9f2f781fd9c4502b98812039edbd2f3f17a41deaa39686be10345

Observation e85b5855-540f-4cbd-ad86-e59174dcc118 · outbound

This paper cites Qwen2.5-VL Technical Report.

Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning Qwen2.5-VL Technical Report

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-04T20:28:55.948935Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:28:55.948935Z digest=sha256:b35381ca0c588725bb5a019d45a9145d3773dd09ca5a4c08e7ca901a585a7a33

Observation 6834c168-d5bc-4d4f-bfcf-8ac73ba9974d · outbound

This paper cites an unresolved cited work.

Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning Unresolved cited work

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-04T20:28:56.073670Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:28:56.073670Z digest=sha256:5903cfbdbb8a3f2c9d8f478757857372987d779485dee32e6f952bb139c946b6

Observation ecce6003-8a7a-455a-aa1b-4041f0ba2e20 · outbound

This paper cites Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling.

Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-04T20:28:56.209162Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:28:56.209162Z digest=sha256:84c305e29aad43e1d3d209560bdee0b66536adaf57d91fe805cd303574d07a84

Observation 7d5ee3e1-6496-4354-9378-596ae27ace96 · outbound

This paper cites Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities.

Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-04T20:28:56.327267Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:28:56.327267Z digest=sha256:ea3e7e962e8efe718442501b8df801d775f6e95b8ba80330b13c528f7017a204

Observation ce9c5351-e408-496c-9530-fecc5e903803 · outbound

This paper cites an unresolved cited work.

Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning Unresolved cited work

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-04T20:28:56.531953Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:28:56.531953Z digest=sha256:922a0c28f92615a235831e6957a0c4d564e5cb715d255634b22ae5840d538167

Observation 68bcb4c3-60dc-4df5-8581-fb6c5d484b7d · outbound

This paper cites an unresolved cited work.

Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning Unresolved cited work

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-04T20:28:56.669105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:28:56.669105Z digest=sha256:e66453f3a72e385a8b2dab719ef32274ab7630269e93a8e803a6ff044f64da48

Observation 0f233e32-b037-4401-a55e-61e8655d5b7d · outbound

This paper cites PDF-MVQA: A Dataset for Multimodal Information Retrieval in PDF-based Visual Question Answering.

Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning PDF-MVQA: A Dataset for Multimodal Information Retrieval in PDF-based Visual Question Answering

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-04T20:28:56.764648Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:28:56.764648Z digest=sha256:a3204c9b13d41f383b939b9f339373c25e834e13b159b4dfcf2a200c46a4f99f

Observation 84b965d6-31f2-41d0-8c12-0fd77429e975 · outbound

This paper cites OCRBench v2: An Improved Benchmark for Evaluating Large Multimodal Models on Visual Text Localization and Reasoning.

Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning OCRBench v2: An Improved Benchmark for Evaluating Large Multimodal Models on Visual Text Localization and Reasoning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-04T20:28:56.875999Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:28:56.875999Z digest=sha256:94ffaeae94a090d19bf89a89acad9aabc7ba5def3601061ba2d9a2c1ce67696a

Observation ba7ee8dd-4604-4a9c-98d6-3baa29446f36 · outbound

This paper cites an unresolved cited work.

Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning Unresolved cited work

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-04T20:28:56.997145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:28:56.997145Z digest=sha256:1a7bb14a9f7c5ed62154c8a85eb9ad42a30f270cf2467b7bfe9df0a7a3f42c85

Observation dc4c2ace-bcc7-4d7e-bfab-26327dd3bb54 · outbound

This paper cites an unresolved cited work.

Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning Unresolved cited work

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-04T20:28:57.072509Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:28:57.072509Z digest=sha256:fd1ec56fb5a667313477eb1217f61d779f6de1e61401fd57bc34c8c58a54b290

Observation 42173e17-e42e-42e5-b848-cb63742da6a0 · outbound

This paper cites GPT-4o System Card.

Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning GPT-4o System Card

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-04T20:28:57.160446Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:28:57.160446Z digest=sha256:e92ac4e8fefed4f6583c9fb7461e1f92d11cd373dd2ae271b502420ef7318852

Observation b5f7687b-1782-48cb-9d62-ba4b4368a11a · outbound

This paper cites R.; Fujii, Y.; Deselaers, T.; Baccash, J.; and Popat, A.

Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning R.; Fujii, Y.; Deselaers, T.; Baccash, J.; and Popat, A

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-04T20:28:57.312037Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:28:57.312037Z digest=sha256:2a4622a5423aecdb6fea71391b07387b9a53b8186ff28bde0184b4227224969e

Observation d8a52f64-a8c7-4f46-957b-134fb5139a23 · outbound

This paper cites K.; and Thiran, J.-P.

Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning K.; and Thiran, J.-P

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-04T20:28:57.424293Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:28:57.424293Z digest=sha256:f26df6d140710e7c40c8bfae8615b8374f44719a2484794197904d21c6f8ff3e

Observation 557e4200-db46-4c2c-9ca7-ecdaffba2050 · outbound

This paper cites an unresolved cited work.

Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-04T20:28:57.529652Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:28:57.529652Z digest=sha256:6719f7745586ef83dbbf7befc7c131da067de73c00a46b758cc8b04d41b20a70

Observation e9721f0c-8083-44ed-9efa-6785e0f783e6 · outbound

This paper cites an unresolved cited work.

Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning Unresolved cited work

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-04T20:28:57.596554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:28:57.596554Z digest=sha256:001976bf02521500fb6124de0e73b5cf71a9b199c85ac686a1a9e540f65cb39d

Observation 2dea98c2-cd7f-40d1-a31d-8ab4466ca338 · outbound

This paper cites LLaVA-OneVision: Easy Visual Task Transfer.

Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning LLaVA-OneVision: Easy Visual Task Transfer

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-04T20:28:57.652002Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:28:57.652002Z digest=sha256:b41faa9514bb72d4f4b96a2beab0da17fa7263a95939d411dcd48db5b962270b

Observation e238eea3-6f32-4f64-bf99-eb5048bc4d12 · outbound

This paper cites an unresolved cited work.

Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning Unresolved cited work

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-04T20:28:57.683646Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:28:57.683646Z digest=sha256:8d5f3214dae7f785e6bcef1c5007d77d7017032a88cb1f031fe0677d34bc77e6

Observation afc78959-9777-4895-b889-565758f6bf31 · outbound

This paper cites an unresolved cited work.

Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning Unresolved cited work

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-04T20:28:57.710848Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:28:57.710848Z digest=sha256:132e673e867b9e773b0e3e1c9081a197105c205269b2d52ade6ca9131e72248a

Observation 6313de6a-6945-4398-91eb-2c5e59d25685 · outbound

This paper cites an unresolved cited work.

Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning Unresolved cited work

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-04T20:28:57.813522Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:28:57.813522Z digest=sha256:63d16b84813d0a2a5da9bee611759e83e8c2cf30f16ff3afa500a608a08442dd

Observation 444421fb-3811-4a36-9131-5a4d9e8f1d4d · outbound

This paper cites an unresolved cited work.

Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning Unresolved cited work

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-04T20:28:57.959749Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:28:57.959749Z digest=sha256:94ad0e48b649714ee08b3c3daeb332c2ec2b484331d57fcfdf25a490d6e73f4e

Observation 923980f4-301f-4b64-a468-19ced63e1a27 · outbound

This paper cites K.; and Chakraborty, A.

Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning K.; and Chakraborty, A

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-04T20:28:58.058222Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:28:58.058222Z digest=sha256:fbbd57e264029be3ce058ca449ca90a353b6818fc29a1f9f668bc94290098f4c

Observation b4a61e2b-1906-41e8-bd58-0084f410b66e · outbound

This paper cites an unresolved cited work.

Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning Unresolved cited work

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-04T20:28:58.157387Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:28:58.157387Z digest=sha256:71de0a2293298a24c1ee9238d7efa227814d9ea36bfe2b7052a1f91c1dd365b4

Observation 5c05b445-3443-469b-9358-f8e16c6f8c64 · outbound

This paper cites Qwen2.5 Technical Report.

Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning Qwen2.5 Technical Report

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-04T20:28:58.325357Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:28:58.325357Z digest=sha256:b1d9a28170b3a874b59bf1334489492bebba458e3c834905b1258c6459ce22f0

Observation 0f80da44-369f-4492-8a0a-4290147fa0e3 · outbound

This paper cites an unresolved cited work.

Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning Unresolved cited work

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-04T20:28:58.539509Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:28:58.539509Z digest=sha256:b4127fea4bdae4a42f4b300d41025fb396dae47f1cd887b05fce3a82f3c04066

Observation 9ae68a05-9103-4eaf-8053-81e0dc4ad9bc · outbound

This paper cites Seed1.5-VL Technical Report.

Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning Seed1.5-VL Technical Report

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-04T20:28:58.597765Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:28:58.597765Z digest=sha256:e84405e461c96508ab1d754b0fcd7665835470796428225d0ee1fdc8ab5a286f

Observation 91e7c22a-d528-4915-896f-12537a872ee0 · outbound

This paper cites an unresolved cited work.

Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning Unresolved cited work

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-04T20:28:58.618224Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:28:58.618224Z digest=sha256:03ad940de3fe58e1c3546e2ea3d0574c575d4bc78f9f7ce353bb72c45933db4b

Observation 2f82990a-7b0e-4f6f-9248-e855639de8e4 · outbound

This paper cites Document-Level Machine Translation with Large Language Models.

Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning Document-Level Machine Translation with Large Language Models

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-04T20:28:58.622466Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:28:58.622466Z digest=sha256:f1dd57fac9a61e714f5aa3ed2b51fd022d26bda03ffbb0a680a0c4a4e3eb15af

Observation b0370a9e-7487-4a8c-95a0-23490b382d37 · outbound

This paper cites DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding.

Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-04T20:28:58.627438Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:28:58.627438Z digest=sha256:e4f071174045de46a5ed371cf52faa880e0bfc555f503d0f3c34acbab7371aa0

Observation dc243383-6c26-4150-adad-655e754655e4 · outbound

This paper cites Qwen2.5-Omni Technical Report.

Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning Qwen2.5-Omni Technical Report

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-04T20:28:58.632055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:28:58.632055Z digest=sha256:84b9e7a3fde3eb3ea557c1adc0301f8d4cb4d296a8760b9f8ca56c29b8ce0f4f

Observation 39b8b67e-0687-4bd8-bcbb-e42f47b4ee30 · outbound

This paper cites an unresolved cited work.

Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning Unresolved cited work

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-04T20:28:58.636639Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:28:58.636639Z digest=sha256:8d0f4f08b81d17c9cbf09d05b995b250a8567161da82cf2268c10410deda4018

Observation feb55710-8259-4352-a362-ac138cbf8833 · outbound

This paper cites an unresolved cited work.

Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning Unresolved cited work

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-04T20:28:58.640806Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:28:58.640806Z digest=sha256:356e3a2286c4487aeda361abc0e23a76e2ad16e37d2ca7cacb92304c06b555be

Observation ec2d7cb7-7569-40e7-8bc1-a9848774673e · outbound

This paper cites Scene Text Recognition with Sliding Convolutional Character Models.

Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning Scene Text Recognition with Sliding Convolutional Character Models

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-04T20:28:58.645516Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:28:58.645516Z digest=sha256:6db318c67b3ca20a5998f1b3c420bcb8f21aab79376ece7dd6590fa05496ef0c

Observation 66b223fc-2b55-4828-9264-a360d38b7796 · outbound

This paper cites Improving the Transformer Translation Model with Document-Level Context.

Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning Improving the Transformer Translation Model with Document-Level Context

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-04T20:28:58.650005Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:28:58.650005Z digest=sha256:62d37a3a78944d9cfcafa4293fb55e22de032b6f627cdbde0106f7bac599c7a2

Observation 57cf7a75-b14d-410c-b564-3b60a47697cc · outbound

This paper cites an unresolved cited work.

Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning Unresolved cited work

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-04T20:28:58.654420Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:28:58.654420Z digest=sha256:58c2484ca30809804e5124d9da0a6bbfbbe103b7d5eb0f322fb2bd2537945d9e

Observation d9796f94-fb58-4486-ac5f-04c0ca76b3ba · outbound

This paper cites BERTScore: Evaluating Text Generation with BERT.

Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning BERTScore: Evaluating Text Generation with BERT

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-04T20:28:58.658554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:28:58.658554Z digest=sha256:08287217a0f30e715c833dc5d56225ee71fd0e3ae03d1076ff18d39c5a76f28c

Observation 2ba5b515-1068-40d9-bb03-eff27398cdb4 · outbound

This paper cites an unresolved cited work.

Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning Unresolved cited work

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-04T20:28:58.663051Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:28:58.663051Z digest=sha256:34f2ea73f195b99910ff1fea25b145de2e73c296539578ac0928e9fe3d6a1d55

Observation c59d43ff-edc1-49a5-93b1-705d9c032d02 · outbound

This paper cites InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models.

Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-04T20:28:58.667157Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:28:58.667157Z digest=sha256:8da1ab758c49904d5fc242e5708ebcadba85e0bfc1d381eea45f0d4a33d3f023

Pith citing papers

Observation c9610636-bfdd-465c-899e-ffaf66302045 · inbound

LPCAN: Lightweight Pyramid Cross-Attention Network for Rail Surface Defect Detection Using RGB-D Data cites this paper.

LPCAN: Lightweight Pyramid Cross-Attention Network for Rail Surface Defect Detection Using RGB-D Data Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-03T10:43:50.289411Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:43:50.289411Z digest=sha256:fdf9bcd08bb24e757872cb68a10ce9be44fc02f390e6116b55f38f397f5fbade

Observation 21d2f413-1c2c-4719-80dc-18887a1dfc29 · inbound

Knowledge-Embedded and Hypernetwork-Guided Few-Shot Substation Meter Defect Image Generation Method cites this paper.

Knowledge-Embedded and Hypernetwork-Guided Few-Shot Substation Meter Defect Image Generation Method Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-03T10:43:42.247993Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:43:42.247993Z digest=sha256:9e815def6312fe039c09dc4f39928f9b28551578ec46f4a28575b812922a3bee

Observation ef8ec59e-2475-4d72-9873-5ec42160fd27 · inbound

Cognitive Mismatch in Multimodal Large Language Models for Discrete Symbol Understanding cites this paper.

Cognitive Mismatch in Multimodal Large Language Models for Discrete Symbol Understanding Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning

Reference 83

Resolution
verified exact
arxiv_id, observed 2026-05-15T09:15:21.120295Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-15T09:11:31.870441Z digest=sha256:6d0c40fe633c6d3766108a78533fcc9b6b94fadd7588c3c50af55777423ff0a6

Observation cbb588df-907a-4b40-b2b3-94f4a0190531 · inbound

Hierarchical Awareness Adapters with Hybrid Pyramid Feature Fusion for Dense Depth Prediction cites this paper.

Hierarchical Awareness Adapters with Hybrid Pyramid Feature Fusion for Dense Depth Prediction Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning

Reference 64

Resolution
verified exact
arxiv_id, observed 2026-05-13T20:03:12.218516Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-13T20:01:57.029143Z digest=sha256:41aca15d7a51bf5e8827fab6b334c8863ffbc65f98313b14c6dcad3fecce6a24

Observation 2e45042a-e34d-459a-a9c2-1c5fa0724e2a · inbound

Do You Need Text Rectification? Soft Attention Mask Embedding for Rectification-Free Scene Text Spotting cites this paper.

Do You Need Text Rectification? Soft Attention Mask Embedding for Rectification-Free Scene Text Spotting Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-05-20T11:33:14.587866Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-20T11:28:51.711892Z digest=sha256:53be606c1939d7f7aa1b9b12105a5c4fb0bc0cb2af7df297458e8e757e568ff1

Observation 9712d40b-4612-4798-8402-1b1288ccb31d · inbound

Cross-Temporal Sinhala OCR: Page-Level Adaptation and Diachronic Analysis cites this paper.

Cross-Temporal Sinhala OCR: Page-Level Adaptation and Diachronic Analysis Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-06-30T07:44:22.120541Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-06-30T07:37:05.806389Z digest=sha256:3b8d59956db9e8a27ec614336404d9340569b29c8a63174a7d5143516350bba1

Observation 83074fb5-d41f-4cdf-9a24-a1817977b1fe · inbound

HCSU: A Dataset and Benchmark for Fine-Grained Historical Calligraphy Style Understanding cites this paper.

HCSU: A Dataset and Benchmark for Fine-Grained Historical Calligraphy Style Understanding Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning

Reference 28

Resolution
unresolved
no resolver link, observed 2026-07-11T21:21:23.781984Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T21:21:23.781984Z digest=sha256:7235738fee9b75d7ac49267c9df8c98a8cab04604b321cafd9274a2a2f4b985f