Pith. sign in

Paper Citation Record · LEDGER

How Far Have Medical Vision-Language Models Come? A Comprehensive Benchmarking Study

As of 11 August 2026, this Paper Citation Record lists 8 of 8 outbound references and 2 inbound Pith citation observations for arXiv:2507.11200.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.11200 v2

Coverage vector

measured 8 of 8 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T17:19:47.183398Z

measured 10 of 10 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-12T13:30:26.488826Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T21:18:59.467672Z

Reference resolution

8 of 8 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved8
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 740feadf-96d1-44be-8a17-9966a2751240 · outbound

This paper cites , " * write output.state after.block = add.period write.

How Far Have Medical Vision-Language Models Come? A Comprehensive Benchmarking Study , " * write output.state after.block = add.period write

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T17:19:46.780825Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:19:46.780825Z digest=sha256:d848694419a6ef038542253ec535b559a07afb4ba571208224df87a034b03d50

Observation 1d22e2f0-6239-43fb-b5d2-e812226b0145 · outbound

This paper cites write newline.

How Far Have Medical Vision-Language Models Come? A Comprehensive Benchmarking Study write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T17:19:46.808585Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:19:46.808585Z digest=sha256:233c8a65dc89b7d1ed7b960ea332b2956f016f36889c3815e49a3e298b4174c7

Observation 2364e215-df6b-4c97-ad7a-176ea07231e5 · outbound

This paper cites Qwen2.5-VL Technical Report.

How Far Have Medical Vision-Language Models Come? A Comprehensive Benchmarking Study Qwen2.5-VL Technical Report

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T17:19:46.853396Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:19:46.853396Z digest=sha256:159fcdbe3aeab599f46d51a09e9b1cfecb7e068a3475c8d2a5db82f1a39e72b2

Observation faf3aa55-f2f1-43e7-b1d3-3c55f7592289 · outbound

This paper cites HuatuoGPT-Vision, Towards Injecting Medical Visual Knowledge into Multimodal LLMs at Scale.

How Far Have Medical Vision-Language Models Come? A Comprehensive Benchmarking Study HuatuoGPT-Vision, Towards Injecting Medical Visual Knowledge into Multimodal LLMs at Scale

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T17:19:46.890585Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:19:46.890585Z digest=sha256:1493cd2807a7838317f3150f47214b36afe42d62c4a9f3446b8cc3b8a87f2745

Observation 9c0eb45d-968e-4b77-abab-ae40351bd9f6 · outbound

This paper cites MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning.

How Far Have Medical Vision-Language Models Come? A Comprehensive Benchmarking Study MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T17:19:46.982626Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:19:46.982626Z digest=sha256:f4ccea583c76253ba43bccf396e1896bff26ca6605edf88f82263646092342fc

Observation 46e0d4a7-b639-4306-8547-51f65d2169d2 · outbound

This paper cites MiMo-VL Technical Report.

How Far Have Medical Vision-Language Models Come? A Comprehensive Benchmarking Study MiMo-VL Technical Report

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T17:19:47.032597Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:19:47.032597Z digest=sha256:42620edd8d78fbe0d4f867d1185d5846e7f57929ee79aa43656e2221d54fa5cd

Observation a2fd42b1-9267-40cb-a7c9-dcdd005dd195 · outbound

This paper cites Disentangling Reasoning and Knowledge in Medical Large Language Models.

How Far Have Medical Vision-Language Models Come? A Comprehensive Benchmarking Study Disentangling Reasoning and Knowledge in Medical Large Language Models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T17:19:47.083790Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:19:47.083790Z digest=sha256:a90aa967866494e5eca8e90f8a20e718a8c04938d05210f21ccfc1f23d246269

Observation 369e548c-b35f-4172-a194-0dc66d14b44b · outbound

This paper cites Lingshu: A Generalist Foundation Model for Unified Multimodal Medical Understanding and Reasoning.

How Far Have Medical Vision-Language Models Come? A Comprehensive Benchmarking Study Lingshu: A Generalist Foundation Model for Unified Multimodal Medical Understanding and Reasoning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T17:19:47.183398Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:19:47.183398Z digest=sha256:1521f4aeb9f17a64b1fe972d5c1b4e2ff7976524b8dd14cc762ac1b988eb32d6

Pith citing papers

Observation a4d1559b-01ff-4c57-80ba-495f6bd1e003 · inbound

Evaluation of Medical Vision Language Models HuluMed and MedGemma, and general purpose chatbots Gemma 3, ChatGPT Plus, and Claude Pro on real previously unseen wound images cites this paper.

Evaluation of Medical Vision Language Models HuluMed and MedGemma, and general purpose chatbots Gemma 3, ChatGPT Plus, and Claude Pro on real previously unseen wound images How Far Have Medical Vision-Language Models Come? A Comprehensive Benchmarking Study

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-07-03T21:18:59.469375Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-27T00:42:29.542446Z digest=sha256:3bde652d1b62d4e269b48080ab6a23d19eb8dae8987cf3cbb70213fc80c531c3

Observation 0e712b36-152f-4be7-bb8a-d510af778c50 · inbound

Evaluation of Medical Vision Language Models HuluMed and MedGemma, and general purpose chatbots Gemma 3, ChatGPT Plus, and Claude Pro on real previously unseen wound images cites this paper.

Evaluation of Medical Vision Language Models HuluMed and MedGemma, and general purpose chatbots Gemma 3, ChatGPT Plus, and Claude Pro on real previously unseen wound images How Far Have Medical Vision-Language Models Come? A Comprehensive Benchmarking Study

Reference 9

Resolution
unresolved
no resolver link, observed 2026-07-12T13:30:26.488826Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T13:30:26.488826Z digest=sha256:7c6e71500309d37f5ac08104387dc55c89c32e54b7d4e9d81afb187fc1c48076