Pith. sign in

Paper Citation Record · LEDGER

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models

As of 10 August 2026, this Paper Citation Record lists 57 of 57 outbound references and 0 inbound Pith citation observations for arXiv:2607.21617.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.21617 v1

Coverage vector

measured 57 of 57 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-02T12:46:12.113877Z

measured 57 of 57 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

57 of 57 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved57
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 2506e149-4c5d-4c23-b8c7-5e1bec7c55c7 · outbound

This paper cites an unresolved cited work.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models Unresolved cited work

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:05.563971Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:05.563971Z digest=sha256:431d0a07d5226ac18323c2ba94dfe3fd28922195c2250aa9b20823975609db93

Observation 12ce943d-fa6f-4641-bbe1-d67a27e3d30d · outbound

This paper cites an unresolved cited work.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models Unresolved cited work

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:05.652444Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:05.652444Z digest=sha256:851bc1c7fd9a420509f4978a878ad01b3981d091789cfb152d467fca8224fae2

Observation aa56fc46-2ecc-4925-83b7-d8274caf0537 · outbound

This paper cites Journal of Foo , volume = 13, number = 1, pages =.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models Journal of Foo , volume = 13, number = 1, pages =

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:05.817702Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:05.817702Z digest=sha256:c9464c71ac0dba8227c5d2db0a8fdaa45b6e96fe258c39b6d090b5034b7e783b

Observation 3feaa8fe-548c-4cb2-ae68-fd73b28f048a · outbound

This paper cites Journal of Foo , volume = 14, number = 1, pages =.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models Journal of Foo , volume = 14, number = 1, pages =

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:05.895645Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:05.895645Z digest=sha256:7066c035bedd24823c696f18e22ecfa965587c4dd6f5989136e780056c06af44

Observation 4d39162e-ad9d-4bcd-a501-f0a62e46ae3c · outbound

This paper cites Trends in Cognitive Sciences , volume =.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models Trends in Cognitive Sciences , volume =

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:06.004527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:06.004527Z digest=sha256:84629f6712fddf19fbc75fccbc3578ec41a0db8e4dc96498b3926b1bf52c91d8

Observation e75dceea-c1bf-4d75-be06-48461e49af6c · outbound

This paper cites an unresolved cited work.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models Unresolved cited work

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:06.101706Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:06.101706Z digest=sha256:0936069b1db53c4466ffa62cdfad9a610768fe51606cbf6e6be88b50055da240

Observation 81ad02b1-c20c-4302-9581-236f3c276028 · outbound

This paper cites arXiv preprint arXiv:2601.20552 , year=.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models arXiv preprint arXiv:2601.20552 , year=

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:06.210404Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:06.210404Z digest=sha256:69352cd246a2366be0cb40dd273d86e4e6f6072cc4d73365061be042e473efb6

Observation 495cdd24-5398-410a-bef2-65f8dd430fe7 · outbound

This paper cites an unresolved cited work.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models Unresolved cited work

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:06.415503Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:06.415503Z digest=sha256:2cf51a11469762ce733092a29b8ad4803beedbfb3f7e13047774ee2368398315

Observation e5209757-c56d-42c5-b2ce-b6d55e9b39a2 · outbound

This paper cites 5: A decoupled vision-language model for efficient high-resolution document parsing , author=.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models 5: A decoupled vision-language model for efficient high-resolution document parsing , author=

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:06.591204Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:06.591204Z digest=sha256:aee51c711667337e8d4631f908e9d32fdd1cafbc3f9dcae1d4a1acf15da09688

Observation 00aff762-07d1-4c8a-b231-06e30cb5a17e · outbound

This paper cites READoc: A Unified Benchmark for Realistic Document Structured Extraction.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models READoc: A Unified Benchmark for Realistic Document Structured Extraction

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:06.801715Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:06.801715Z digest=sha256:dff8768d22153afc6f799e33dea11c67ef1bd594b645558ce8b9941d1810fadb

Observation e8fb02da-0a11-4c58-b63e-8f2d9080f8f1 · outbound

This paper cites Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:06.931296Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:06.931296Z digest=sha256:3e1e3abe5c470f60f97a9ef7530870926a58ed02acf1a6986bb6b05583bb5537

Observation b1c0290a-9c00-43fa-9366-08f455c5770f · outbound

This paper cites Qwen3-VL Technical Report.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models Qwen3-VL Technical Report

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:07.069203Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:07.069203Z digest=sha256:c891dd6267de55454204fa0c50fd7f520d7a83bc61a572ff4624b6123a937b03

Observation 58f3bfe7-80bf-4b5e-ad40-49f0033ef59f · outbound

This paper cites an unresolved cited work.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models Unresolved cited work

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:07.238033Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:07.238033Z digest=sha256:f3c20231b3434d6dbe082abb1ea98eb1d75ccea0ead5005e7625681dd1831f41

Observation f62387f8-0529-431a-a407-c6c2a6b700bb · outbound

This paper cites 2026 , eprint=.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models 2026 , eprint=

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:07.417593Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:07.417593Z digest=sha256:61437b26d783526c02056eda16c2000a297c4642530ced31fb13eafe4e5c1037

Observation a3473b54-a9c2-43a3-8d74-158a70e4d2fe · outbound

This paper cites Proceedings of the IEEE international conference on computer vision , pages=.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models Proceedings of the IEEE international conference on computer vision , pages=

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:07.559805Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:07.559805Z digest=sha256:72068135f053fe7cabeb090549db76f095538ddf43f09e23380e6247220cab5e

Observation 117d116e-3141-421d-96f4-800bea718c1f · outbound

This paper cites 2021 , eprint=.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models 2021 , eprint=

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:07.675539Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:07.675539Z digest=sha256:b8837962b9e176e5dd271f0e872d106093fd6fe4a1c8b438956bb39008186b1a

Observation 9fa0babf-6708-4460-93d5-4728f1b8b621 · outbound

This paper cites 2024 , eprint=.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models 2024 , eprint=

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:07.737098Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:07.737098Z digest=sha256:50dd42d77a3a5b2d339fded416278714f32652612fd300d6d9ccc50bee3a1139

Observation d095c590-24f6-470f-a039-d5219e9a0152 · outbound

This paper cites InternVL3.5: Advancing Open-Source Multimodal Models in Versatility, Reasoning, and Efficiency.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models InternVL3.5: Advancing Open-Source Multimodal Models in Versatility, Reasoning, and Efficiency

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:07.800275Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:07.800275Z digest=sha256:bf0345a68137ea50aeb21e954b27fae3a1fd36bbdc50f4753c1683d966c263a3

Observation f1f44dd4-73b9-4bac-9200-c35f04826d3f · outbound

This paper cites Visual Instruction Tuning , url =.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models Visual Instruction Tuning , url =

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:07.865094Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:07.865094Z digest=sha256:9886feeb61e6c0ba5ea4b0514a7f7a725a83ef67b06b98eaa7c41e6d456ee732

Observation a88e607f-c268-407b-b433-a6ce145eb330 · outbound

This paper cites GPT-4 Technical Report.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models GPT-4 Technical Report

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:07.918581Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:07.918581Z digest=sha256:2f292b97e49a000734cf71c4d44ea8bed9a6fd03c9bab9f3a5d96e1f42aeb0f2

Observation 94eab9b6-08c6-4345-8d31-fe76ff7847b2 · outbound

This paper cites 2023 , eprint=.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models 2023 , eprint=

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:08.056765Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:08.056765Z digest=sha256:9fd807c72a4c593e86e5e9f9e94080e6af1579411b68967c7bbc80ab5b7142fa

Observation cf9bbd9e-5b13-4e88-852f-78888a90c54a · outbound

This paper cites OCRBench: on the hidden mystery of OCR in large multimodal models , volume=.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models OCRBench: on the hidden mystery of OCR in large multimodal models , volume=

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:08.181936Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:08.181936Z digest=sha256:36f4b1850928b9e5ec3f008e66d16dea8ecec78cfa8e603ce7b4f3d3eac0583c

Observation 3266d655-2118-45e4-8058-452e9f17999f · outbound

This paper cites 2025 , eprint=.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models 2025 , eprint=

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:08.377781Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:08.377781Z digest=sha256:7d8af4d7c49d96f63e03590c99e18409e8424f58ab36ea0ec8fc85248b29a6c4

Observation 1b2d3719-4388-4112-93f2-d4aed75c1894 · outbound

This paper cites 2021 , eprint=.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models 2021 , eprint=

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:08.539837Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:08.539837Z digest=sha256:d03a84c42b99edcb316df9db4b6a1b5289bdac5dff0a2f82558142836aa16596

Observation 96228a5e-dbda-4b47-b000-a1a800e4b1d1 · outbound

This paper cites 2024 , eprint=.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models 2024 , eprint=

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:08.733743Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:08.733743Z digest=sha256:8a0c84398b92d5588f125feaa5574e311bdfa9a0eca6c48646530ce12e98cc44

Observation acdba914-2f72-4139-b604-3dcb96c3bd33 · outbound

This paper cites 2020 , eprint=.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models 2020 , eprint=

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:08.880444Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:08.880444Z digest=sha256:a38cc511dc07b37b2a92a0c5d809260012f8a918d0674b762ad6f391a6c3694c

Observation 83be50c9-17ef-4e33-92d1-383f440ca962 · outbound

This paper cites Survey of Hallucination in Natural Language Generation , volume=.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models Survey of Hallucination in Natural Language Generation , volume=

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:09.009817Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:09.009817Z digest=sha256:5f4ad81b02817ced5111437503bf0f5bfaa86b1ebdae94727b912c8aa52bddae

Observation cc7a3d58-72f4-4f7a-9a39-c66ac86c1eb2 · outbound

This paper cites 2023 , eprint=.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models 2023 , eprint=

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:09.137858Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:09.137858Z digest=sha256:74be6867924f87c4e5ade73128b14a691cdd4a2532de1aee941634c2c6c5e9ec

Observation 7a509ac8-3a29-45a8-88a2-9077eb0490e1 · outbound

This paper cites Forty-second International Conference on Machine Learning , year=.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models Forty-second International Conference on Machine Learning , year=

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:09.278951Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:09.278951Z digest=sha256:00e1bc26fc9a10316e42be19392f81ce59df4d26b31756ea03dfa8e00ab437ae

Observation 6fa70d81-a0e3-4978-a9d0-b120f9b03f9d · outbound

This paper cites Linux Journal , volume=.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models Linux Journal , volume=

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:09.381601Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:09.381601Z digest=sha256:76efe23f01de50d366061af8687a6d8b54e536f1030b520e9f3be4489552a5cf

Observation e67333de-d374-4cc8-bede-b4727713c7b8 · outbound

This paper cites National Science Review , volume=.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models National Science Review , volume=

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:09.504605Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:09.504605Z digest=sha256:eabbac687e40f3cfebf453670e1d073f8c032abe4a35ad8286cd8b3ab06f2a47

Observation ccaf5634-fbf2-4422-9539-f4bee3b06dcf · outbound

This paper cites IEEE transactions on pattern analysis and machine intelligence , volume=.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models IEEE transactions on pattern analysis and machine intelligence , volume=

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:09.576749Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:09.576749Z digest=sha256:278ba6636bc08f8f8028f46fe28414be08eab58c891fc9eb97d23c4d9c05315a

Observation f0a64ca0-7550-4def-924a-95cf1737f7f5 · outbound

This paper cites LayoutLM: Pre-training of Text and Layout for Document Image Understanding , url=.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models LayoutLM: Pre-training of Text and Layout for Document Image Understanding , url=

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:09.648762Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:09.648762Z digest=sha256:c1d5e4a63304debca1383bf4a9c0d9497f640fe660b3a9a731cd7454850d05e4

Observation 350f618f-059a-41b7-ba55-954416cb2379 · outbound

This paper cites Ocean-OCR: Towards General OCR Application via a Vision-Language Model.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models Ocean-OCR: Towards General OCR Application via a Vision-Language Model

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:09.712264Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:09.712264Z digest=sha256:76240b2b299fab870401e46161c896209a133ef9182eb0ddb24021c1a57dfc65

Observation f5dcf2d0-eb38-4d89-90a1-76b8b8a6d2bd · outbound

This paper cites arXiv preprint arXiv:2502.18443 , year=.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models arXiv preprint arXiv:2502.18443 , year=

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:09.786583Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:09.786583Z digest=sha256:3043200dfe9ce3c750efa0fb4f05fce112bc0181b69e0a21905564a4fda745d9

Observation a1d801ed-00cc-44b4-a674-a62160a0b3f4 · outbound

This paper cites MinerU: An Open-Source Solution for Precise Document Content Extraction.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models MinerU: An Open-Source Solution for Precise Document Content Extraction

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:09.834605Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:09.834605Z digest=sha256:5ca27293e13beae3b5204ffc85c92ee615207870746eb1787680387ff2243ec3

Observation 6cf654f6-b8e2-46b7-bfd4-dc8472285628 · outbound

This paper cites PP-OCR: A Practical Ultra Lightweight OCR System.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models PP-OCR: A Practical Ultra Lightweight OCR System

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:09.903700Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:09.903700Z digest=sha256:987ca00862cff4e297ec855415c1c445af884441727f560a088ea9ebaac7ee03

Observation 3208d8a3-ba0b-4477-a22a-a040458141d6 · outbound

This paper cites Ninth international conference on document analysis and recognition (ICDAR 2007) , volume=.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models Ninth international conference on document analysis and recognition (ICDAR 2007) , volume=

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:09.982617Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:09.982617Z digest=sha256:ab59c33f3d109490547c908b49f1ccfaea6233ef3c77ef7b219acfb5765b3ac7

Observation c6a80f1e-fd10-4168-98e4-4b78c9ac3e91 · outbound

This paper cites Proceedings of the 33rd ACM International Conference on Multimedia , pages=.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models Proceedings of the 33rd ACM International Conference on Multimedia , pages=

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:10.083752Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:10.083752Z digest=sha256:0ff19ea68a4009fdfcb44de00061de10b75301f66e7f82f1aac74ae4d240b6c1

Observation 7fbb042b-3d26-450e-a81c-df8ea532cd38 · outbound

This paper cites Journal of machine learning research , volume=.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models Journal of machine learning research , volume=

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:10.148535Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:10.148535Z digest=sha256:cb95e8c06befded4e8464f4659d6831a4ea07bc6086e7223cce77dc357766adb

Observation 1ff74ba2-0d7b-4154-9a18-52ad8d5f86df · outbound

This paper cites 2021 , publisher =.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models 2021 , publisher =

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:10.215369Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:10.215369Z digest=sha256:ddc0c60c6ffdc9cceeb12ff82f5f8e707d0cb0b15b6101df14f4472aa1e1c0a0

Observation d790d216-96c4-4b5a-a2b9-a2aa1327bccd · outbound

This paper cites 2025 , howpublished =.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models 2025 , howpublished =

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:10.467072Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:10.467072Z digest=sha256:e2cb2451b7745b4243f23249a6163a7569853153918ce58e30e1d66d72fddeca

Observation 4ff9148a-d2d0-41b5-8e2f-77fde8d1b489 · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models Gemini: A Family of Highly Capable Multimodal Models

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:10.631021Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:10.631021Z digest=sha256:de895fe1d21478bf7848ecb98b293c08e9ef1783a005645b73ae34509083f556

Observation 4be17acc-ce43-4ce5-a1d3-191c13f5a822 · outbound

This paper cites ArXiv , year=.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models ArXiv , year=

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:10.833031Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:10.833031Z digest=sha256:0f7bb7b02f58b05ee2aea3352b62a71ea905a58d9d6f37f1895ffc56a54762ac

Observation 7e7b6a7c-e5c9-4bd9-b3dd-fc45d1c41fbb · outbound

This paper cites Synthetic and Natural Noise Both Break Neural Machine Translation.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models Synthetic and Natural Noise Both Break Neural Machine Translation

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:10.997571Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:10.997571Z digest=sha256:5d5687a8e0fb4fe94ac18ecb7df20724c95082f756af2e1c23c2c9e91f483ac6

Observation 830efd6d-2cf9-4606-8a6b-4d1427ef2019 · outbound

This paper cites Proceedings of the 56th Annual Meeting of the Association for Computational Linguistics (Volume 2: Short Papers) , pages=.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models Proceedings of the 56th Annual Meeting of the Association for Computational Linguistics (Volume 2: Short Papers) , pages=

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:11.160831Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:11.160831Z digest=sha256:dfdb0fef9934a90e50e2e1dd27a4aaa8eb90c2354f4251746078cf795fb6d791

Observation 06c4e05f-6f1c-4428-84b8-3618773331e7 · outbound

This paper cites Proceedings of the 57th Annual Meeting of the Association for Computational Linguistics , pages=.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models Proceedings of the 57th Annual Meeting of the Association for Computational Linguistics , pages=

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:11.255823Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:11.255823Z digest=sha256:5bb2cc49e574312e705db113cfac79244abae9c656af07beaa04adeb2db1e8ab

Observation 3bc4cac9-7383-4f29-b767-246a0b6046e6 · outbound

This paper cites Proceedings of the 2023 Conference on Empirical Methods in Natural Language Processing , pages=.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models Proceedings of the 2023 Conference on Empirical Methods in Natural Language Processing , pages=

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:11.345358Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:11.345358Z digest=sha256:7678281b9877a696cf1d83b2de1060bc5c07d30f2e63e45bf9f6c836f385ac29

Observation 35c88fa4-4807-4b80-a29a-539699e3a66e · outbound

This paper cites Qwen3.5-Omni Technical Report.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models Qwen3.5-Omni Technical Report

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:11.456166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:11.456166Z digest=sha256:95bc5de9ce4546693ce1956028ae89c63ffe51f76815991c4b72e423c3be684d

Observation bdbc8cab-a8f3-4595-baf1-9cc026a839ad · outbound

This paper cites arXiv preprint arXiv:2510.17771 , year=.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models arXiv preprint arXiv:2510.17771 , year=

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:11.588458Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:11.588458Z digest=sha256:ef2cc71e80489311ba9dd7bedb8c28cf295b78ccd733a6d17f5c2625bcc82a26

Observation 90c8079e-6d5d-48f6-a6c6-372bf49b2598 · outbound

This paper cites Vision Language Models are Biased.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models Vision Language Models are Biased

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:11.638518Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:11.638518Z digest=sha256:68543b28c8176f058e352de23c79b4a111decf723a1cdc95e23b6a829961ced4

Observation d01c5297-61e0-4901-bdd7-ef70a5c2b097 · outbound

This paper cites Findings of the Association for Computational Linguistics: NAACL 2025 , pages=.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models Findings of the Association for Computational Linguistics: NAACL 2025 , pages=

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:11.740750Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:11.740750Z digest=sha256:85f2e903f156a8f9c8865e46a5519a02d10229fbbd467247ec9efebd1cea1d66

Observation 02b638c6-b96b-4c5b-b6dd-8f5c1690261e · outbound

This paper cites Proceedings of the 2023 conference on empirical methods in natural language processing , pages=.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models Proceedings of the 2023 conference on empirical methods in natural language processing , pages=

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:11.787638Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:11.787638Z digest=sha256:3a75ec4a2bde4101d320e06595c40f78093a09f513bc2afc32d926d22b03a16f

Observation 922e6cd1-b60e-4253-ad2c-c0c94c0dc56c · outbound

This paper cites Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:11.856663Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:11.856663Z digest=sha256:a2bb71647a5718dba6c5c884071ae6678505ac708995a35cfd1019ffb2534db4

Observation 8f392fc2-ebd7-4c9d-b058-17b934684485 · outbound

This paper cites AMBER: An LLM-free Multi-dimensional Benchmark for MLLMs Hallucination Evaluation.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models AMBER: An LLM-free Multi-dimensional Benchmark for MLLMs Hallucination Evaluation

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:11.926310Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:11.926310Z digest=sha256:000490c1ea13ab6ebd28d52de70760257f0c773fd5bad30ccb0854ccdfb52cf3

Observation 29d1489b-8bfe-4abc-8a9d-3c27e3256531 · outbound

This paper cites 2026 , howpublished=.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models 2026 , howpublished=

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:11.997842Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:11.997842Z digest=sha256:bc7cfef495991040e4ef895bab110f4dde01895c42a584e211fc356bbff1d268

Observation f9e2377e-e515-40a3-92e3-f7617384f504 · outbound

This paper cites OpenAI GPT-5 System Card.

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models OpenAI GPT-5 System Card

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-02T12:46:12.113877Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:46:12.113877Z digest=sha256:95b0f687094b664b1a7457c788a7f6637e24af007e7e45081bcdd79235098b45

Pith citing papers

No inbound Pith citation observations are available.