Pith. sign in

Paper Citation Record · LEDGER

Vision-Language Models for Medical Report Generation and Visual Question Answering: A Review

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:2403.02469.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2403.02469 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 5 of 5 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:35:07.407511Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-18T18:22:46.845434Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 9cd5f3d5-eaab-4130-8cbc-14b0bebc64d6 · inbound

DrVD-Bench: Do Vision-Language Models Reason Like Human Doctors in Medical Image Diagnosis? cites this paper.

DrVD-Bench: Do Vision-Language Models Reason Like Human Doctors in Medical Image Diagnosis? Vision-Language Models for Medical Report Generation and Visual Question Answering: A Review

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:07.407511Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:35:07.407511Z digest=sha256:532b33fcf0fb4f91d0495983d0c5bc0a7be6d8caa125418e3fcb005d0b9f1997

Observation 6eca3ffc-47e8-4153-aeaf-5e1b8e09c441 · inbound

Analysis of Blood Report Images Using General Purpose Vision-Language Models cites this paper.

Analysis of Blood Report Images Using General Purpose Vision-Language Models Vision-Language Models for Medical Report Generation and Visual Question Answering: A Review

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-18T18:22:46.849703Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-18T18:22:10.226073Z digest=sha256:315e3f122f5e69d08baf3ac943f4f4f7830286e48affcbb314f1cf295441ad11

Observation c3f916ae-5c86-4f4a-91d2-83e5e98408f9 · inbound

Medical Report Generation: A Hierarchical Task Structure-Based Cross-Modal Causal Intervention Framework cites this paper.

Medical Report Generation: A Hierarchical Task Structure-Based Cross-Modal Causal Intervention Framework Vision-Language Models for Medical Report Generation and Visual Question Answering: A Review

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-18T01:45:36.968527Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-18T01:45:29.400398Z digest=sha256:b43544617b590e52a9bcf9446e0f21505630891ea0dc62ad3fd9d72156bba775

Observation 8a39552c-019c-4f81-acd3-883678661fbc · inbound

DentiAsk: A VQA Benchmark for Multimodal Reasoning in Panoramic Dental Radiographs cites this paper.

DentiAsk: A VQA Benchmark for Multimodal Reasoning in Panoramic Dental Radiographs Vision-Language Models for Medical Report Generation and Visual Question Answering: A Review

Reference 42

Resolution
unresolved
no resolver link, observed 2026-07-13T07:15:40.242481Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-13T07:15:40.242481Z digest=sha256:667e86a006edfb15ec92c02d8f65fd6905ca5b1b746226d1206adbe52926164b

Observation 91bb6565-e0dc-45f2-ac94-45ccabee905a · inbound

DobicVLM: Aligning Chest X-Ray Report Generation with Clinically-Grounded Programmatic Rewards via Group Relative Policy Optimization cites this paper.

DobicVLM: Aligning Chest X-Ray Report Generation with Clinically-Grounded Programmatic Rewards via Group Relative Policy Optimization Vision-Language Models for Medical Report Generation and Visual Question Answering: A Review

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-01T13:52:40.407120Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T13:52:40.407120Z digest=sha256:534667fc3f1153f015ba3d1fd42cd7cce742d6c47abd80505a5b521aa6f27d76