Pith. sign in

Paper Citation Record · LEDGER

LLaVA-Read: Enhancing Reading Ability of Multimodal Language Models

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:2407.19185.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2407.19185 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 5 of 5 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T12:54:56.375610Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 03acbac3-de00-4c32-926f-a2c60e18e971 · inbound

EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering cites this paper.

EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering LLaVA-Read: Enhancing Reading Ability of Multimodal Language Models

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-08T12:54:56.375610Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T12:54:56.375610Z digest=sha256:f0c28f194ac335dd5156ac51e24051320cb71ed809ef07b6596703ef1de23feb

Observation 97993c1c-9a6d-4154-9aff-ae7d5edad3c3 · inbound

MusiXQA: Advancing Visual Music Understanding in Multimodal Large Language Models cites this paper.

MusiXQA: Advancing Visual Music Understanding in Multimodal Large Language Models LLaVA-Read: Enhancing Reading Ability of Multimodal Language Models

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T21:57:07.053168Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:57:07.053168Z digest=sha256:1962ba14987874bbe8258aa7e9772086b2374d3d68b64cf191380fd3c01b4e04

Observation 2cb5e786-46b7-4cdb-aa97-67a2c5b36239 · inbound

A Survey on MLLM-based Visually Rich Document Understanding: Methods, Challenges, and Emerging Trends cites this paper.

A Survey on MLLM-based Visually Rich Document Understanding: Methods, Challenges, and Emerging Trends LLaVA-Read: Enhancing Reading Ability of Multimodal Language Models

Reference 72

Resolution
verified exact
arxiv_id, observed 2026-05-19T04:42:04.077947Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-19T04:38:49.512293Z digest=sha256:c87f94b744a458fd73df9ea15b6c40f34b03fe6dca0793fe72cf7ef69fe11bda

Observation 994391db-86c6-4abe-b1a3-4383ced3ff08 · inbound

ReforMe: Re-Shaping Documents with Contextual Prompting and Layout-Aware Propagation cites this paper.

ReforMe: Re-Shaping Documents with Contextual Prompting and Layout-Aware Propagation LLaVA-Read: Enhancing Reading Ability of Multimodal Language Models

Reference 55

Resolution
verified exact
arxiv_id, observed 2026-07-02T04:56:38.956998Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-28T08:43:27.671508Z digest=sha256:566a12dd7d5b35b53a9d1851fad95c032750343dc147ed6b773d311b874ac072

Observation 94de8d19-2ad0-4711-b9bc-c59ded39dc22 · inbound

DocOCR-Eval: A Correction-Based Framework for OCR Tool Selection Without Ground Truth cites this paper.

DocOCR-Eval: A Correction-Based Framework for OCR Tool Selection Without Ground Truth LLaVA-Read: Enhancing Reading Ability of Multimodal Language Models

Reference 161

Resolution
unresolved
no resolver link, observed 2026-08-02T14:55:37.457570Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T14:55:37.457570Z digest=sha256:d263e32eeb9436e1bdf7b6c98b0af1447700819faa749c6af21a96d69f6d34f1