Pith. sign in

Paper Citation Record · LEDGER

Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning

As of 8 August 2026, this Paper Citation Record lists 41 of 41 outbound references and 7 inbound Pith citation observations for arXiv:2509.09731.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2509.09731 v1

Coverage vector

measured 41 of 41 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-04T20:28:58.667157Z

measured 48 of 48 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 7 of 7 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-03T10:43:50.289411Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-30T07:44:22.118843Z

Reference resolution

41 of 41 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved41
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation f05893e5-57e7-48f1-94d7-59d5ba16ed80 · outbound

This paper cites , " * write output.state after.block = add.period write newline.

Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning , " * write output.state after.block = add.period write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-04T20:28:55.488191Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:28:55.488191Z digest=sha256:63d70dc720418e995dcf6a9bf4581f94bf29681c317f21d907d65108b2a879d9

Observation 9c830c9c-e00f-480a-aa2f-9052dcf7e671 · outbound

This paper cites write newline.

Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-04T20:28:55.595187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:28:55.595187Z digest=sha256:d665a18f4372c012a9d6822d3c7d853e803df56c904455c4e81743708c10588f

Observation 911156ce-463a-4dba-bc3f-d66365587edb · outbound

This paper cites an unresolved cited work.

Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning Unresolved cited work

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-04T20:28:55.684666Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:28:55.684666Z digest=sha256:222fecdb96519050fb25713ea4760466ff885f958e6100a3bb72f5beb229ab6d

Observation 2e3ce02e-f44c-4212-b5a4-927873b0f9a2 · outbound

This paper cites Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond.

Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-04T20:28:55.790100Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:28:55.790100Z digest=sha256:e63acf7f8babd52fe29c6a2557ebe227f4a1c56e613ce882f9208136fe4ec464

Observation e85b5855-540f-4cbd-ad86-e59174dcc118 · outbound

This paper cites Qwen2.5-VL Technical Report.

Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning Qwen2.5-VL Technical Report

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-04T20:28:55.948935Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:28:55.948935Z digest=sha256:9c64ba4139f092cc586e70cb7f7b81f6d8932b7830e6ec99476a32ec17f29378

Observation 6834c168-d5bc-4d4f-bfcf-8ac73ba9974d · outbound

This paper cites an unresolved cited work.

Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning Unresolved cited work

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-04T20:28:56.073670Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:28:56.073670Z digest=sha256:2c80e4f71fd4b9672083013a7e91eaeee9874b1f98b21136a28e50510a9d706c

Observation ecce6003-8a7a-455a-aa1b-4041f0ba2e20 · outbound

This paper cites Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling.

Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-04T20:28:56.209162Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:28:56.209162Z digest=sha256:efe0a6ccb6bafe69431fab5a40e8902b68305319f0a25f11715b3778baa54f84

Observation 7d5ee3e1-6496-4354-9378-596ae27ace96 · outbound

This paper cites Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities.

Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-04T20:28:56.327267Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:28:56.327267Z digest=sha256:8d51dc929860bee425b29d0d6c890ec90b3739069f94987d4efc1e666f3bdcfb

Observation ce9c5351-e408-496c-9530-fecc5e903803 · outbound

This paper cites an unresolved cited work.

Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning Unresolved cited work

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-04T20:28:56.531953Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:28:56.531953Z digest=sha256:50e7f64e9f4535c271fecb47dbf01e1f0bb66baa0c6cc0f0503aec7137c2113a

Observation 68bcb4c3-60dc-4df5-8581-fb6c5d484b7d · outbound

This paper cites an unresolved cited work.

Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning Unresolved cited work

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-04T20:28:56.669105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:28:56.669105Z digest=sha256:a2138dced4abc7b07be4759150278a86c2410c16e18a2a7a8e5756a683979a60

Observation 0f233e32-b037-4401-a55e-61e8655d5b7d · outbound

This paper cites PDF-MVQA: A Dataset for Multimodal Information Retrieval in PDF-based Visual Question Answering.

Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning PDF-MVQA: A Dataset for Multimodal Information Retrieval in PDF-based Visual Question Answering

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-04T20:28:56.764648Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:28:56.764648Z digest=sha256:d964f51928486e583acbeb7234c02186fe78355b70bfa906d61019177e25ba06

Observation 84b965d6-31f2-41d0-8c12-0fd77429e975 · outbound

This paper cites OCRBench v2: An Improved Benchmark for Evaluating Large Multimodal Models on Visual Text Localization and Reasoning.

Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning OCRBench v2: An Improved Benchmark for Evaluating Large Multimodal Models on Visual Text Localization and Reasoning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-04T20:28:56.875999Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:28:56.875999Z digest=sha256:12e7f2d5ed4d059be3c9b2932672a3ae89eadc46aaa5c81963eeae6cfc4e96ac

Observation ba7ee8dd-4604-4a9c-98d6-3baa29446f36 · outbound

This paper cites an unresolved cited work.

Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning Unresolved cited work

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-04T20:28:56.997145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:28:56.997145Z digest=sha256:67256624779d2f6482a44e2136982e7092d62c7312fd35daf71c1514a9a0e68f

Observation dc4c2ace-bcc7-4d7e-bfab-26327dd3bb54 · outbound

This paper cites an unresolved cited work.

Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning Unresolved cited work

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-04T20:28:57.072509Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:28:57.072509Z digest=sha256:c4ced86748702ae3f9d695896c9829a33371484b5abcfa599dc81d69b8adfff3

Observation 42173e17-e42e-42e5-b848-cb63742da6a0 · outbound

This paper cites GPT-4o System Card.

Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning GPT-4o System Card

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-04T20:28:57.160446Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:28:57.160446Z digest=sha256:23622f3ccbc018715e6c6a48d308c792d1fb8d19cff916ce3ecd2b4e56cdd8b8

Observation b5f7687b-1782-48cb-9d62-ba4b4368a11a · outbound

This paper cites R.; Fujii, Y.; Deselaers, T.; Baccash, J.; and Popat, A.

Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning R.; Fujii, Y.; Deselaers, T.; Baccash, J.; and Popat, A

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-04T20:28:57.312037Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:28:57.312037Z digest=sha256:7d20f832b3b8b48727812bc96af0a2758fd4378e70b9b20bba2f45be0abc85ad

Observation d8a52f64-a8c7-4f46-957b-134fb5139a23 · outbound

This paper cites K.; and Thiran, J.-P.

Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning K.; and Thiran, J.-P

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-04T20:28:57.424293Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:28:57.424293Z digest=sha256:e4129cf38ef48b920d730f4ab40b0beb2f14c8021b23473d09d2d76bc2e05647

Observation 557e4200-db46-4c2c-9ca7-ecdaffba2050 · outbound

This paper cites an unresolved cited work.

Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-04T20:28:57.529652Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:28:57.529652Z digest=sha256:cd76a7f82d66aff6febd984711e3d725745244e785c4580285592a762b8c380d

Observation e9721f0c-8083-44ed-9efa-6785e0f783e6 · outbound

This paper cites an unresolved cited work.

Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning Unresolved cited work

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-04T20:28:57.596554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:28:57.596554Z digest=sha256:66025207fff230d85f54690d68563e81931af12f3c6eb79166a78967986256fe

Observation 2dea98c2-cd7f-40d1-a31d-8ab4466ca338 · outbound

This paper cites LLaVA-OneVision: Easy Visual Task Transfer.

Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning LLaVA-OneVision: Easy Visual Task Transfer

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-04T20:28:57.652002Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:28:57.652002Z digest=sha256:c5f583d073eb54c9ad65b899c4e9a3cc76925784ae2e15d77406301aa93a7aa0

Observation e238eea3-6f32-4f64-bf99-eb5048bc4d12 · outbound

This paper cites an unresolved cited work.

Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning Unresolved cited work

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-04T20:28:57.683646Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:28:57.683646Z digest=sha256:9722ed83c79fd6fea16da4318c2ab713592eeca0328d97dfd3924b80aba09893

Observation afc78959-9777-4895-b889-565758f6bf31 · outbound

This paper cites an unresolved cited work.

Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning Unresolved cited work

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-04T20:28:57.710848Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:28:57.710848Z digest=sha256:f66ccb018a967552d9c2cc8d48144e66eb6274184ebe931b2205c08a4c43e753

Observation 6313de6a-6945-4398-91eb-2c5e59d25685 · outbound

This paper cites an unresolved cited work.

Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning Unresolved cited work

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-04T20:28:57.813522Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:28:57.813522Z digest=sha256:73e3479d2ab4c088b16122a2c9d94994f70e7e065515381ca53855769bf53d83

Observation 444421fb-3811-4a36-9131-5a4d9e8f1d4d · outbound

This paper cites an unresolved cited work.

Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning Unresolved cited work

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-04T20:28:57.959749Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:28:57.959749Z digest=sha256:61cb6a8b571ce45283b8e60b607979ad109c31329e142f11dea8b5025df0fa99

Observation 923980f4-301f-4b64-a468-19ced63e1a27 · outbound

This paper cites K.; and Chakraborty, A.

Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning K.; and Chakraborty, A

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-04T20:28:58.058222Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:28:58.058222Z digest=sha256:d4f96419a682fcce6fe816616a91f405831b0f77420550a9757173109291fd5d

Observation b4a61e2b-1906-41e8-bd58-0084f410b66e · outbound

This paper cites an unresolved cited work.

Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning Unresolved cited work

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-04T20:28:58.157387Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:28:58.157387Z digest=sha256:125c93585b488bb7525aa353e532129e72f122f3ea48fa9d4f6b2e8792a26e74

Observation 5c05b445-3443-469b-9358-f8e16c6f8c64 · outbound

This paper cites Qwen2.5 Technical Report.

Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning Qwen2.5 Technical Report

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-04T20:28:58.325357Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:28:58.325357Z digest=sha256:cadb3d574e8c79d33a88d32845697fabce941b5940e56924dc7f4db337e05506

Observation 0f80da44-369f-4492-8a0a-4290147fa0e3 · outbound

This paper cites an unresolved cited work.

Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning Unresolved cited work

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-04T20:28:58.539509Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:28:58.539509Z digest=sha256:8c191a06cdda2eea993b40fb2e09690a3e13d0277c71f95a06e9c9f9d47f47e9

Observation 9ae68a05-9103-4eaf-8053-81e0dc4ad9bc · outbound

This paper cites Seed1.5-VL Technical Report.

Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning Seed1.5-VL Technical Report

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-04T20:28:58.597765Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:28:58.597765Z digest=sha256:71650e69c8d5327be5cceac69b854b096236d00ca47b870753e8028fbb612cfa

Observation 91e7c22a-d528-4915-896f-12537a872ee0 · outbound

This paper cites an unresolved cited work.

Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning Unresolved cited work

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-04T20:28:58.618224Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:28:58.618224Z digest=sha256:fa23ee897a10ae8f23713d94d3537aa926ebe2023217942085860fec1faa2fa7

Observation 2f82990a-7b0e-4f6f-9248-e855639de8e4 · outbound

This paper cites Document-Level Machine Translation with Large Language Models.

Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning Document-Level Machine Translation with Large Language Models

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-04T20:28:58.622466Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:28:58.622466Z digest=sha256:59e02226472d4e3519f2e218bd8912b323aef0b93abcd36947b108d6b6c68214

Observation b0370a9e-7487-4a8c-95a0-23490b382d37 · outbound

This paper cites DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding.

Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-04T20:28:58.627438Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:28:58.627438Z digest=sha256:5627871002ad39694587931441bee47a173e8e0ddcc987f6f942d1f4d0ab5ef2

Observation dc243383-6c26-4150-adad-655e754655e4 · outbound

This paper cites Qwen2.5-Omni Technical Report.

Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning Qwen2.5-Omni Technical Report

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-04T20:28:58.632055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:28:58.632055Z digest=sha256:3ae2223220205d96d19f8151c80993ce7444745704de25db7450f78606e1e742

Observation 39b8b67e-0687-4bd8-bcbb-e42f47b4ee30 · outbound

This paper cites an unresolved cited work.

Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning Unresolved cited work

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-04T20:28:58.636639Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:28:58.636639Z digest=sha256:5646733848eeddb74474e36b3a997464df7095f0577d0755e9a17bef282cb635

Observation feb55710-8259-4352-a362-ac138cbf8833 · outbound

This paper cites an unresolved cited work.

Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning Unresolved cited work

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-04T20:28:58.640806Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:28:58.640806Z digest=sha256:9843b9031629d2938ec7ede184c87749cd7a43ec90fd3132408a09b672b2f9a5

Observation ec2d7cb7-7569-40e7-8bc1-a9848774673e · outbound

This paper cites Scene Text Recognition with Sliding Convolutional Character Models.

Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning Scene Text Recognition with Sliding Convolutional Character Models

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-04T20:28:58.645516Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:28:58.645516Z digest=sha256:a7015878fc37c07e7aeb894e1f51edf8e9ba10e4789499160db125c04363636e

Observation 66b223fc-2b55-4828-9264-a360d38b7796 · outbound

This paper cites Improving the Transformer Translation Model with Document-Level Context.

Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning Improving the Transformer Translation Model with Document-Level Context

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-04T20:28:58.650005Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:28:58.650005Z digest=sha256:647b82ed20740e05efd4675b7e407bc2db32ce057dcfa919781276d94acd1c56

Observation 57cf7a75-b14d-410c-b564-3b60a47697cc · outbound

This paper cites an unresolved cited work.

Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning Unresolved cited work

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-04T20:28:58.654420Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:28:58.654420Z digest=sha256:c74d900ac2fe58ba90a5d3a88def75318bfd203b1dd7a875910dc4e66760f80d

Observation d9796f94-fb58-4486-ac5f-04c0ca76b3ba · outbound

This paper cites BERTScore: Evaluating Text Generation with BERT.

Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning BERTScore: Evaluating Text Generation with BERT

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-04T20:28:58.658554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:28:58.658554Z digest=sha256:267a74d3090e2c8fec03131a16d5295fde193b700bca0a4b01630a3ae0f2108d

Observation 2ba5b515-1068-40d9-bb03-eff27398cdb4 · outbound

This paper cites an unresolved cited work.

Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning Unresolved cited work

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-04T20:28:58.663051Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:28:58.663051Z digest=sha256:67cec75e83e5dbc8c4e8326cc8eb7ce1e988686d4ad559d536c83d31a098cfab

Observation c59d43ff-edc1-49a5-93b1-705d9c032d02 · outbound

This paper cites InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models.

Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-04T20:28:58.667157Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:28:58.667157Z digest=sha256:9b7dbf163983260eb38043875a259e316cba2a5e45f02799d5b22f82b1819e7b

Pith citing papers

Observation c9610636-bfdd-465c-899e-ffaf66302045 · inbound

LPCAN: Lightweight Pyramid Cross-Attention Network for Rail Surface Defect Detection Using RGB-D Data cites this paper.

LPCAN: Lightweight Pyramid Cross-Attention Network for Rail Surface Defect Detection Using RGB-D Data Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-03T10:43:50.289411Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:43:50.289411Z digest=sha256:becedea0b5dc483c672fcc99d143a2c659d03f981f5d7c25dc917997c4a19bd8

Observation 21d2f413-1c2c-4719-80dc-18887a1dfc29 · inbound

Knowledge-Embedded and Hypernetwork-Guided Few-Shot Substation Meter Defect Image Generation Method cites this paper.

Knowledge-Embedded and Hypernetwork-Guided Few-Shot Substation Meter Defect Image Generation Method Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-03T10:43:42.247993Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:43:42.247993Z digest=sha256:2e66478904dbc11b6a3f9ce45cc307a8bc59d31bbceb23a24651071b5beb8415

Observation ef8ec59e-2475-4d72-9873-5ec42160fd27 · inbound

Cognitive Mismatch in Multimodal Large Language Models for Discrete Symbol Understanding cites this paper.

Cognitive Mismatch in Multimodal Large Language Models for Discrete Symbol Understanding Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning

Reference 83

Resolution
verified exact
arxiv_id, observed 2026-05-15T09:15:21.120295Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T09:11:31.870441Z digest=sha256:6c58860198b202e778b53f3f4e9eb3db96216bed2e4a91e69fd0749b79694e70

Observation cbb588df-907a-4b40-b2b3-94f4a0190531 · inbound

Hierarchical Awareness Adapters with Hybrid Pyramid Feature Fusion for Dense Depth Prediction cites this paper.

Hierarchical Awareness Adapters with Hybrid Pyramid Feature Fusion for Dense Depth Prediction Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning

Reference 64

Resolution
verified exact
arxiv_id, observed 2026-05-13T20:03:12.218516Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-13T20:01:57.029143Z digest=sha256:afa132f812c17884644f204b634efbdec2a145871039a7c47d2e803b0229098b

Observation 2e45042a-e34d-459a-a9c2-1c5fa0724e2a · inbound

Do You Need Text Rectification? Soft Attention Mask Embedding for Rectification-Free Scene Text Spotting cites this paper.

Do You Need Text Rectification? Soft Attention Mask Embedding for Rectification-Free Scene Text Spotting Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-05-20T11:33:14.587866Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-20T11:28:51.711892Z digest=sha256:7f46cc1343492e4066923e7fb2b308910ad421e598ceb04e101e51e5bed3b035

Observation 9712d40b-4612-4798-8402-1b1288ccb31d · inbound

Cross-Temporal Sinhala OCR: Page-Level Adaptation and Diachronic Analysis cites this paper.

Cross-Temporal Sinhala OCR: Page-Level Adaptation and Diachronic Analysis Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-06-30T07:44:22.120541Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-30T07:37:05.806389Z digest=sha256:769af8c5a429008b9d2d788e9a4c9a35edf287a683f2789fef0a9eb24d54b105

Observation 83074fb5-d41f-4cdf-9a24-a1817977b1fe · inbound

HCSU: A Dataset and Benchmark for Fine-Grained Historical Calligraphy Style Understanding cites this paper.

HCSU: A Dataset and Benchmark for Fine-Grained Historical Calligraphy Style Understanding Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning

Reference 28

Resolution
unresolved
no resolver link, observed 2026-07-11T21:21:23.781984Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T21:21:23.781984Z digest=sha256:10e2a97627e8263ed5b29332978d028affcf0658661d7930eef3a7979fc0f921