Pith. sign in

Paper Citation Record · LEDGER

MedBLINK: Probing Basic Perception in Multimodal Language Models for Medicine

As of 8 August 2026, this Paper Citation Record lists 60 of 60 outbound references and 1 inbound Pith citation observation for arXiv:2508.02951.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.02951 v1

Coverage vector

measured 60 of 60 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T04:50:45.382132Z

measured 61 of 61 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T04:50:45.050116Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-06T04:50:45.808704Z

Reference resolution

60 of 60 outbound references displayed

  • verified exact3
  • verified fuzzy29
  • unresolved23
  • parse uncertain4
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 47c153a4-45ee-4a12-91f4-2dc59ba0543a · outbound

This paper cites MedBLINK: Probing Basic Perception in Multimodal Language Models for Medicine.

MedBLINK: Probing Basic Perception in Multimodal Language Models for Medicine MedBLINK: Probing Basic Perception in Multimodal Language Models for Medicine

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-08-06T04:50:45.814234Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T04:50:45.050116Z digest=sha256:5f90e96d36dd1e0959d86102c37455b8e8f16d8071824573ab3ad3e380f4e21e

Observation 710b94de-7a56-4593-a1ca-35d63f90de9a · outbound

This paper cites These models combine text and image understanding and are typ- ically evaluated using visual question answering (VQA) tasks.

MedBLINK: Probing Basic Perception in Multimodal Language Models for Medicine These models combine text and image understanding and are typ- ically evaluated using visual question answering (VQA) tasks

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T04:50:46.557554Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T04:50:45.056467Z digest=sha256:b69277a1632a0e86619e2bf08fca77795ac0c0f89d8a7fd4c5136049b94947d2

Observation 6af40122-e672-4103-8449-829ae1cbafde · outbound

This paper cites Perception enables clinicians to ex- tract key visual features before engaging in more complex 2 Figure 2.

MedBLINK: Probing Basic Perception in Multimodal Language Models for Medicine Perception enables clinicians to ex- tract key visual features before engaging in more complex 2 Figure 2

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T04:50:46.527645Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T04:50:45.067950Z digest=sha256:e63b9f8b101ce3a821127c574d1678f0db85b56f2e35b7f89baf1fcc0e7c44b0

Observation e4d88d05-d7f5-42a3-9f80-f09ff5e32940 · outbound

This paper cites an unresolved cited work.

MedBLINK: Probing Basic Perception in Multimodal Language Models for Medicine Unresolved cited work

Reference 4

Resolution
malformed identifier
raw_fallback, observed 2026-08-06T04:50:46.452654Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T04:50:45.090608Z digest=sha256:5f3963bb28b1d7029eed2f2446cdb70fd8f4b108102bd8037e26765dbdc2f767

Observation 847fb887-7ac6-46d8-bfa2-1993dbbd1d71 · outbound

This paper cites an unresolved cited work.

MedBLINK: Probing Basic Perception in Multimodal Language Models for Medicine Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-06T04:50:46.510041Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T04:50:45.073600Z digest=sha256:eefc6c48a4a58ef155efec5da6a06c10dbf9aa047816828c229016210437b64d

Observation 7f86ac3c-fbac-475f-8c64-101e529991c7 · outbound

This paper cites an unresolved cited work.

MedBLINK: Probing Basic Perception in Multimodal Language Models for Medicine Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-06T04:50:46.491579Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T04:50:45.079017Z digest=sha256:00f1294346737a09312730503a3852dfac3e4b2b2f8210a7f84c9c9ec1fb121b

Observation 54d22850-cdf2-4baa-ae46-bf76d3454179 · outbound

This paper cites We mimic this form of prompting, by leveraging visual cues like points/dots to spatially prompt the models when answering the specified question [43, 50].

MedBLINK: Probing Basic Perception in Multimodal Language Models for Medicine We mimic this form of prompting, by leveraging visual cues like points/dots to spatially prompt the models when answering the specified question [43, 50]

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T04:50:46.474122Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T04:50:45.084623Z digest=sha256:e5d295e6ab20533ce43ffbbbd2aa010f82f42f63003017659ee051635fd28207

Observation f0ade7de-781a-417f-b728-1541bb262b80 · outbound

This paper cites Current models including leading generalist and domain-specialized systems perform far below human levels on perceptual tasks that clinicians solve effortlessly (best: 65% vs.

MedBLINK: Probing Basic Perception in Multimodal Language Models for Medicine Current models including leading generalist and domain-specialized systems perform far below human levels on perceptual tasks that clinicians solve effortlessly (best: 65% vs

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T04:50:46.435678Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T04:50:45.096640Z digest=sha256:b2437420747aa8a5c7c3c2898edf1f302ba05faa16fb50b108e5453f510c9dd8

Observation 6280ff95-d0ab-431c-bc7b-415cae1e46b7 · outbound

This paper cites an unresolved cited work.

MedBLINK: Probing Basic Perception in Multimodal Language Models for Medicine Unresolved cited work

Reference 11

Resolution
parse uncertain
raw_fallback, observed 2026-08-06T04:50:45.865994Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T04:50:45.371172Z digest=sha256:312a9ab758f97e72cfa50c3dfbb34ed9a03e22f0af9a9e87dc59b603d98727d8

Observation d1fcc25d-bb44-4297-9949-9cef024da10d · outbound

This paper cites an unresolved cited work.

MedBLINK: Probing Basic Perception in Multimodal Language Models for Medicine Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-06T04:50:45.848416Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T04:50:45.376206Z digest=sha256:1e472ca7413f465501a2c02b172388d66b0ab400b02c538e3e2c348b1766d4c1

Observation 382d428a-6816-4345-9815-1370576706e9 · outbound

This paper cites Small Specialized models We train small specialized models for some of the tasks with sizeable train sets from the original dataset used to construct the task.

MedBLINK: Probing Basic Perception in Multimodal Language Models for Medicine Small Specialized models We train small specialized models for some of the tasks with sizeable train sets from the original dataset used to construct the task

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T04:50:45.832023Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T04:50:45.382132Z digest=sha256:fa01dc1d724a4c67e2cd1de76cde222925117c52feadbb7bf8b07bc0fe36067c

Observation 29967050-21e7-49b3-b916-a48d40f29a5d · outbound

This paper cites Omnimedvqa: A new large- scale comprehensive evaluation benchmark for medical lvlm.

MedBLINK: Probing Basic Perception in Multimodal Language Models for Medicine Omnimedvqa: A new large- scale comprehensive evaluation benchmark for medical lvlm

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T04:50:46.417870Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T04:50:45.101847Z digest=sha256:b2d6d3c2c6857915823dca840728f6212339f3852c556cb89c346f0ed74ade06

Observation 3fd89b40-4785-4f9c-9255-30b339a8010b · outbound

This paper cites GPT-4o System Card.

MedBLINK: Probing Basic Perception in Multimodal Language Models for Medicine GPT-4o System Card

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T04:50:45.107165Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T04:50:45.107165Z digest=sha256:3ca07014225aca2be0a53d26e20c7ff17a5cde21a7534ddfe150d533e843d852

Observation f3429a48-d7ad-4bf8-9555-7e26f08a7c10 · outbound

This paper cites The diagnostic impact of contrast-enhanced computed tomography (cect) in evaluating lymph node involvement in colorectal cancer: a comprehen- sive review.

MedBLINK: Probing Basic Perception in Multimodal Language Models for Medicine The diagnostic impact of contrast-enhanced computed tomography (cect) in evaluating lymph node involvement in colorectal cancer: a comprehen- sive review

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T04:50:46.401814Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T04:50:45.113559Z digest=sha256:1cd351d8216702957798b51a1e694c35ce92d918044a950f8ef11b28893fbe19

Observation 4fa8b32c-08b4-4b95-8ee6-b032d571f42c · outbound

This paper cites Chexpert: A large chest radiograph dataset with uncertainty labels and expert comparison.

MedBLINK: Probing Basic Perception in Multimodal Language Models for Medicine Chexpert: A large chest radiograph dataset with uncertainty labels and expert comparison

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T04:50:46.369421Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T04:50:45.119177Z digest=sha256:39915c44d82169439b8ee418e16aaee6c9796649d082a32e8328a6b91ee5918d

Observation 4702e604-2bdc-4114-abe0-c795ffa19983 · outbound

This paper cites A dataset of clinically generated visual questions and answers about radiology images.

MedBLINK: Probing Basic Perception in Multimodal Language Models for Medicine A dataset of clinically generated visual questions and answers about radiology images

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T04:50:46.346514Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T04:50:45.124916Z digest=sha256:5520c2a19e4e38a821f5b8b948330c51715833fd1f8bdfbe2502f6a7f5d101c9

Observation 42ef0246-05b1-41a5-90c1-25d056987509 · outbound

This paper cites Llava-onevision: Easy visual task transfer, 2024.

MedBLINK: Probing Basic Perception in Multimodal Language Models for Medicine Llava-onevision: Easy visual task transfer, 2024

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T04:50:46.328216Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T04:50:45.130611Z digest=sha256:9caa70cf83e1142040a444f9fc61f44037a996d502cf29e31b61d55444bee836

Observation 8008b3fc-6a3b-4fff-87cd-a44e45d15971 · outbound

This paper cites LLaVA-OneVision: Easy Visual Task Transfer.

MedBLINK: Probing Basic Perception in Multimodal Language Models for Medicine LLaVA-OneVision: Easy Visual Task Transfer

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T04:50:45.135890Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T04:50:45.135890Z digest=sha256:88cd74add5199ee60db3c66c8dc0b11f7c1d1b9b45d4684d1321e247475a1ae4

Observation 80da6ea6-1f71-466c-9550-417b77da3a1f · outbound

This paper cites Llava-med: Training a large language- and-vision assistant for biomedicine in one day.

MedBLINK: Probing Basic Perception in Multimodal Language Models for Medicine Llava-med: Training a large language- and-vision assistant for biomedicine in one day

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T04:50:45.141848Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T04:50:45.141848Z digest=sha256:0f81d059d870e2713f34491c568e01d8ebaa01e00f232a0d533a1c691b567fe0

Observation 0f1634c8-b201-43e2-94db-5372c0df43d5 · outbound

This paper cites Slake: A semantically-labeled knowledge- enhanced dataset for medical visual question answering.

MedBLINK: Probing Basic Perception in Multimodal Language Models for Medicine Slake: A semantically-labeled knowledge- enhanced dataset for medical visual question answering

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T04:50:46.300661Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T04:50:45.146966Z digest=sha256:0490723fc018fd247514dfc2648a2dddf45a7ebe7486b88ff8d1bd28c5737426

Observation 3039d259-df27-4083-9eb5-38244175a009 · outbound

This paper cites Improved Baselines with Visual Instruction Tuning.

MedBLINK: Probing Basic Perception in Multimodal Language Models for Medicine Improved Baselines with Visual Instruction Tuning

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T04:50:45.152773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T04:50:45.152773Z digest=sha256:3ee97f1a0d33f1ac0629528aa0fb23dd6515568e8e218278b3d9e194979a55b4

Observation 503ad637-3270-4f14-acbd-35c7b1d9855a · outbound

This paper cites A Spectrum Evaluation Benchmark for Medical Multi-Modal Large Language Models.

MedBLINK: Probing Basic Perception in Multimodal Language Models for Medicine A Spectrum Evaluation Benchmark for Medical Multi-Modal Large Language Models

Reference 36

Resolution
verified exact
local_arxiv, observed 2026-08-06T04:50:45.737508Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T04:50:45.158675Z digest=sha256:7781c176a8092dc17f2eb904b76bc48d2aa7a142eefdee7781e4161b15c06cca

Observation be03e375-8617-4b08-8dc6-5727ce900fe7 · outbound

This paper cites The coding of roentgen images for computer analysis as applied to lung cancer.

MedBLINK: Probing Basic Perception in Multimodal Language Models for Medicine The coding of roentgen images for computer analysis as applied to lung cancer

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T04:50:46.282228Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T04:50:45.164214Z digest=sha256:44992f8e4b31fbc1e3bb438f1963b7bad3bb5b6fb7d423bb34ee13558386ecf2

Observation b9d79b9d-06b9-4d12-a317-a8d6f17014b8 · outbound

This paper cites Lu, Bowen Chen, Drew F.

MedBLINK: Probing Basic Perception in Multimodal Language Models for Medicine Lu, Bowen Chen, Drew F

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T04:50:46.265144Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T04:50:45.170000Z digest=sha256:eb0123724f0793b9b1400bfad41c7eed20119e5754900802d58f54ddc57af89b

Observation 3fa64eb0-1fd8-4a7c-a042-bc6278f9868b · outbound

This paper cites Lung-rads: pushing the limits.

MedBLINK: Probing Basic Perception in Multimodal Language Models for Medicine Lung-rads: pushing the limits

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T04:50:46.246274Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T04:50:45.175275Z digest=sha256:dad219a050530b8263f2e28338a12f63451f14d2b7fe54da9230b925240d4329

Observation dca4a1ca-1a4c-4c61-89a8-22c858a68100 · outbound

This paper cites Towards Accurate Differential Diagnosis with Large Language Models.

MedBLINK: Probing Basic Perception in Multimodal Language Models for Medicine Towards Accurate Differential Diagnosis with Large Language Models

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T04:50:45.181592Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T04:50:45.181592Z digest=sha256:863195d74192fa41b3d29db72678cc9e3964ec158fecfcf2a61ab2e307a8dc50

Observation 4614b4aa-68e4-4136-8e44-eb087144fc66 · outbound

This paper cites Med-flamingo: a multimodal medical few-shot learner.

MedBLINK: Probing Basic Perception in Multimodal Language Models for Medicine Med-flamingo: a multimodal medical few-shot learner

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T04:50:46.229278Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T04:50:45.188471Z digest=sha256:ebe75f921784bb21cc8ffb04c78dab615ae36e287979d0c71c0d14f7bc0cdf6e

Observation 7a1be3ce-2ba3-40d0-9881-1351627d55e9 · outbound

This paper cites Interactions of perceptual and conceptual pro- cessing: Expertise in medical image diagnosis.

MedBLINK: Probing Basic Perception in Multimodal Language Models for Medicine Interactions of perceptual and conceptual pro- cessing: Expertise in medical image diagnosis

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T04:50:46.212954Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T04:50:45.194078Z digest=sha256:06a410817880e4fb75be14af2348d1a99cc522f17d30fe3c92da87f9685eab98

Observation 0910e13e-2cab-4d99-92ef-3dbdd9f931a7 · outbound

This paper cites Beyond the Hype: A dispassionate look at vision-language models in medical scenario.

MedBLINK: Probing Basic Perception in Multimodal Language Models for Medicine Beyond the Hype: A dispassionate look at vision-language models in medical scenario

Reference 43

Resolution
verified exact
local_arxiv, observed 2026-08-06T04:50:45.691969Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T04:50:45.199029Z digest=sha256:9ec7995657723121683800b29e7d36c1ec28fae7cf8d574bd24ea9b7ee68a3b0

Observation 3dbd8851-1a05-4103-8386-467712ece06c · outbound

This paper cites Video-based ai for beat-to-beat assessment of cardiac function.

MedBLINK: Probing Basic Perception in Multimodal Language Models for Medicine Video-based ai for beat-to-beat assessment of cardiac function

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T04:50:46.196722Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T04:50:45.205271Z digest=sha256:18b096067e5ce0177d7fb7a9fa20f234dbae2555f3cc1ec9400a078fa8671e8b

Observation bb599bb3-104d-4f89-b8eb-12d27c251d4f · outbound

This paper cites Kvasir: A multi-class image dataset for computer aided gastrointestinal disease detection.

MedBLINK: Probing Basic Perception in Multimodal Language Models for Medicine Kvasir: A multi-class image dataset for computer aided gastrointestinal disease detection

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T04:50:46.177764Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T04:50:45.210922Z digest=sha256:1a0ed9f6e4273e9c56484cee539e2e245a7728f92efae63072e2557e5844c577

Observation 4c63d650-f3ba-46a8-9471-4d7e92744df8 · outbound

This paper cites Learning transferable visual models from natural language supervision, 2021.

MedBLINK: Probing Basic Perception in Multimodal Language Models for Medicine Learning transferable visual models from natural language supervision, 2021

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-06T04:50:45.215891Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T04:50:45.215891Z digest=sha256:78d93672cec2ffad39be2ce00bea62e1d5a8608825ed631f6c9d79a684a0e034

Observation 95b409ea-16bb-41c5-9bbe-3c8adf40784d · outbound

This paper cites Mul- timedeval: A benchmark and a toolkit for evaluating medical vision-language models.

MedBLINK: Probing Basic Perception in Multimodal Language Models for Medicine Mul- timedeval: A benchmark and a toolkit for evaluating medical vision-language models

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-06T04:50:45.220710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T04:50:45.220710Z digest=sha256:48162830e8769cc8e116a4196e65c2abf8a22bb46db314d86ee28532c7e618bd

Observation dbf45989-4f88-43f9-bbfb-004d4022d6f9 · outbound

This paper cites Hematoxylin and eosin-stained whole slide image dataset annotated for skin tissue segmentation.

MedBLINK: Probing Basic Perception in Multimodal Language Models for Medicine Hematoxylin and eosin-stained whole slide image dataset annotated for skin tissue segmentation

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T04:50:46.142357Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T04:50:45.225802Z digest=sha256:0158a93663b4a89aa42ad758b559698c2f2149fa1a739f00bc71dcdf2edadcdf

Observation cfd50d5a-86c3-4101-adee-4ff98f625024 · outbound

This paper cites MediConfusion: Can you trust your AI radiologist? Probing the reliability of multimodal medical foundation models.

MedBLINK: Probing Basic Perception in Multimodal Language Models for Medicine MediConfusion: Can you trust your AI radiologist? Probing the reliability of multimodal medical foundation models

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-06T04:50:45.231663Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T04:50:45.231663Z digest=sha256:059be1b4e56b03b188a0b42801394810d72557595ff3c90eda20eb6cad35fe67

Observation 46cd4533-fb17-4e19-bb77-767cfc68408f · outbound

This paper cites Quilt-llava: Visual instruction tuning by extracting localized narratives from open-source histopathology videos.

MedBLINK: Probing Basic Perception in Multimodal Language Models for Medicine Quilt-llava: Visual instruction tuning by extracting localized narratives from open-source histopathology videos

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T04:50:46.121877Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T04:50:45.237379Z digest=sha256:760c6d24207e101b3c8c0d5bca8e14abaaa17c3a52135683688548e23219bb29

Observation 72043f7c-db68-416f-b52f-7337a3562546 · outbound

This paper cites Diagnostic ultrasound imaging: inside out.

MedBLINK: Probing Basic Perception in Multimodal Language Models for Medicine Diagnostic ultrasound imaging: inside out

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T04:50:46.102569Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T04:50:45.242452Z digest=sha256:d86bfe6f87151912ad2a1a1c9f9820b999a58c468f96aa8e567a5269b8c6728a

Observation db8c5afb-1320-454f-9fe9-b77cbb2332d1 · outbound

This paper cites Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context.

MedBLINK: Probing Basic Perception in Multimodal Language Models for Medicine Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-06T04:50:45.249202Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T04:50:45.249202Z digest=sha256:af505ef0ed00afaf807cde7a37c63410013d2fedb582abf7078f56ba28ee23d3

Observation 66423372-b8a5-4212-9344-495965152d86 · outbound

This paper cites What clinicians want: contextualizing explainable machine learning for clinical end use.

MedBLINK: Probing Basic Perception in Multimodal Language Models for Medicine What clinicians want: contextualizing explainable machine learning for clinical end use

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T04:50:46.084094Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T04:50:45.255259Z digest=sha256:6f7f05fb918a494a3f67a6bc22d63c8fa8d2f24b224f555054b5870c06e248a7

Observation 92afd454-277d-4b91-8d57-a9cbcf723d8f · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

MedBLINK: Probing Basic Perception in Multimodal Language Models for Medicine Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-06T04:50:45.261256Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T04:50:45.261256Z digest=sha256:b4a70911e4dd6b0e2a381c13f96ae652b907c7cc682a1cdae2eefc239971b9c0

Observation f998d0a9-5378-44a1-8740-88c68dbde60a · outbound

This paper cites Corrado, Yossi Matias, Karan Singhal, Pete Florence, Alan Karthikesalingam, and Vivek Natarajan.

MedBLINK: Probing Basic Perception in Multimodal Language Models for Medicine Corrado, Yossi Matias, Karan Singhal, Pete Florence, Alan Karthikesalingam, and Vivek Natarajan

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T04:50:46.067523Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T04:50:45.268906Z digest=sha256:2d97d6c2dbcb01d6ea15744752a03e1f0495d88c8b36ed8d80f338c7780477e0

Observation 00bfc95d-4598-4802-a203-d9aea1a245c1 · outbound

This paper cites Baichuan-M1: Pushing the Medical Capability of Large Language Models.

MedBLINK: Probing Basic Perception in Multimodal Language Models for Medicine Baichuan-M1: Pushing the Medical Capability of Large Language Models

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-06T04:50:45.273861Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T04:50:45.273861Z digest=sha256:167b21c6717fdb9fe1d10d8742abf2aebd9eec40814062895f6a0979d7eaf275

Observation 0a69827b-6308-4eb3-9760-2852c2d11ef6 · outbound

This paper cites Chestx- ray8: Hospital-scale chest x-ray database and benchmarks on weakly-supervised classification and localization of common thorax diseases.

MedBLINK: Probing Basic Perception in Multimodal Language Models for Medicine Chestx- ray8: Hospital-scale chest x-ray database and benchmarks on weakly-supervised classification and localization of common thorax diseases

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-06T04:50:45.279006Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T04:50:45.279006Z digest=sha256:bbb1c37b2f98ba10537372f58505fb05ebd119f71e1189b19b6104916456de94

Observation fd19d5cf-bcf1-4ea1-a4d5-e45f44bcb0ff · outbound

This paper cites Detection of radiographic abnor- malities in mammograms by means of optical scanning and computer analysis.

MedBLINK: Probing Basic Perception in Multimodal Language Models for Medicine Detection of radiographic abnor- malities in mammograms by means of optical scanning and computer analysis

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T04:50:46.040026Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T04:50:45.283634Z digest=sha256:ce75cb3f5afb1dddca6c697a547aa8865ac8438befefe503eea662585df8e7bb

Observation 12078155-60aa-4a78-b8e6-edd367152b85 · outbound

This paper cites Can GPT-4V(ision) Serve Medical Applications? Case Studies on GPT-4V for Multimodal Medical Diagnosis.

MedBLINK: Probing Basic Perception in Multimodal Language Models for Medicine Can GPT-4V(ision) Serve Medical Applications? Case Studies on GPT-4V for Multimodal Medical Diagnosis

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-06T04:50:45.288838Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T04:50:45.288838Z digest=sha256:f7b2583647b6e101c6b328f2164057017b9d07d4caa5efeebaaae9fe23c392c2

Observation 360772cb-1472-4e1e-934d-18f1316792ab · outbound

This paper cites Towards Generalist Foundation Model for Radiology by Leveraging Web-scale 2D&3D Medical Data.

MedBLINK: Probing Basic Perception in Multimodal Language Models for Medicine Towards Generalist Foundation Model for Radiology by Leveraging Web-scale 2D&3D Medical Data

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-06T04:50:45.295335Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T04:50:45.295335Z digest=sha256:08caf1e3abccbd54baeeaf320d5c83a148a33b75fe0479ba238e148a9eb8c8db

Observation a26b6eee-0fca-48a4-a7db-b5e9c62f8bbf · outbound

This paper cites MediConfusion [ 49] probes failure modes on visually dissimilar image pairs.

MedBLINK: Probing Basic Perception in Multimodal Language Models for Medicine MediConfusion [ 49] probes failure modes on visually dissimilar image pairs

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T04:50:46.543033Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T04:50:45.061737Z digest=sha256:3558969892347fd6fbae9bb571067c27eeb1e44e5f6c0c61ab39ac037935acc8

Observation ec15513e-c7ba-42c5-a280-4db773bee185 · outbound

This paper cites CARES: A Comprehensive Benchmark of Trustworthiness in Medical Vision Language Models.

MedBLINK: Probing Basic Perception in Multimodal Language Models for Medicine CARES: A Comprehensive Benchmark of Trustworthiness in Medical Vision Language Models

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-06T04:50:45.300767Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T04:50:45.300767Z digest=sha256:3bece02c4442627ff91321529209cebb7e3f3522676c2341b0458e019f6b9eb2

Observation 42125ddc-4570-45f7-b84c-5dba5e666f68 · outbound

This paper cites MedTrinity-25M: A Large-scale Multimodal Dataset with Multigranular Annotations for Medicine.

MedBLINK: Probing Basic Perception in Multimodal Language Models for Medicine MedTrinity-25M: A Large-scale Multimodal Dataset with Multigranular Annotations for Medicine

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-06T04:50:45.306820Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T04:50:45.306820Z digest=sha256:17a653624fdeddbfb9f7165b5e10493e66caed60350186516e41e8e8a26828cf

Observation ea407e41-f8ca-4185-a42f-eee8bff03f30 · outbound

This paper cites Depth anything: Unleashing the power of large-scale unlabeled data.

MedBLINK: Probing Basic Perception in Multimodal Language Models for Medicine Depth anything: Unleashing the power of large-scale unlabeled data

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T04:50:46.022706Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T04:50:45.312871Z digest=sha256:d97b6898ddbd61772f6ae8542cc0b3d438e4c604e74d0aba3d45eedf817063c5

Observation d4185d3c-b5de-49c8-9e27-4923254fee06 · outbound

This paper cites Advancing Multimodal Medical Capabilities of Gemini.

MedBLINK: Probing Basic Perception in Multimodal Language Models for Medicine Advancing Multimodal Medical Capabilities of Gemini

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-06T04:50:45.317935Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T04:50:45.317935Z digest=sha256:51245510b5bc4fb3e8bc08d5e0b1cafcac7a82a2ce6a819a94a3075941a41783

Observation 82cbc2ec-e34a-479a-ac03-8f78de3111f6 · outbound

This paper cites Gmai-mmbench: A comprehensive multimodal evaluation benchmark towards general medical ai.

MedBLINK: Probing Basic Perception in Multimodal Language Models for Medicine Gmai-mmbench: A comprehensive multimodal evaluation benchmark towards general medical ai

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T04:50:46.002495Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T04:50:45.323714Z digest=sha256:5f1c389816a1256cbb7417d079a1fc65841a0be1d085f675aa6b38fc53f896d2

Observation d7768017-fca9-49a6-b8dd-6a6de030ffa8 · outbound

This paper cites PMC-VQA: Visual Instruction Tuning for Medical Visual Question Answering.

MedBLINK: Probing Basic Perception in Multimodal Language Models for Medicine PMC-VQA: Visual Instruction Tuning for Medical Visual Question Answering

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-06T04:50:45.328920Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T04:50:45.328920Z digest=sha256:984e9917abd1eb19a200651ddc2d6b0d1c5bfa7682c34913504ba6ecde9e4a5c

Observation 8cfcdeb0-8fa5-4ce3-9f83-fe4581342e22 · outbound

This paper cites an unresolved cited work.

MedBLINK: Probing Basic Perception in Multimodal Language Models for Medicine Unresolved cited work

Reference 68

Resolution
parse uncertain
raw_fallback, observed 2026-08-06T04:50:45.983909Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T04:50:45.334499Z digest=sha256:6ac374b29fc6f70021c4623fd73d700cf02f8c693375356689c90983e374979e

Observation 22370e71-3979-4e4d-8cf4-dbd23db8f2b5 · outbound

This paper cites an unresolved cited work.

MedBLINK: Probing Basic Perception in Multimodal Language Models for Medicine Unresolved cited work

Reference 69

Resolution
parse uncertain
raw_fallback, observed 2026-08-06T04:50:45.968069Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T04:50:45.339961Z digest=sha256:64e7f983f8a837e983f663c9dfaab14496e10db6f7c81b929593d89a35973af8

Observation 34fc5d89-f78b-4671-aa47-456e4ddfc969 · outbound

This paper cites an unresolved cited work.

MedBLINK: Probing Basic Perception in Multimodal Language Models for Medicine Unresolved cited work

Reference 70

Resolution
parse uncertain
raw_fallback, observed 2026-08-06T04:50:45.951271Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T04:50:45.346086Z digest=sha256:407cddd7c8d409091d41c4a1f187a65c1af5999b2d2c5e1471c317334d62a61c

Observation 04848663-0141-4fc7-b3ed-9f935c04c9e9 · outbound

This paper cites an unresolved cited work.

MedBLINK: Probing Basic Perception in Multimodal Language Models for Medicine Unresolved cited work

Reference 71

Resolution
unresolved
raw_fallback, observed 2026-08-06T04:50:45.935521Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T04:50:45.351161Z digest=sha256:49245e28bca13da6a4c6739b464bd7454b0c0c92db0264be430a61ae51281799

Observation 710a4900-c0a2-41d9-bbfd-5c4fa6eb15e1 · outbound

This paper cites Here we use two versions as well, the 0.5B parameterized model, and 7B parameter- ized models.

MedBLINK: Probing Basic Perception in Multimodal Language Models for Medicine Here we use two versions as well, the 0.5B parameterized model, and 7B parameter- ized models

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T04:50:45.918541Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T04:50:45.356130Z digest=sha256:6ef7e6a984ae2772178a136c685eeaf9111e25af58c11ad4ddf1be73c16b571a

Observation b2e27d68-1e17-4fcc-bf01-fa012d90f9a5 · outbound

This paper cites Specifically, we prompt it with three questions and answers from PMC-VQA benchmark.

MedBLINK: Probing Basic Perception in Multimodal Language Models for Medicine Specifically, we prompt it with three questions and answers from PMC-VQA benchmark

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T04:50:45.900362Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T04:50:45.361108Z digest=sha256:87de81806d3886f52709b5a8befa7ab43bab16193b633cbde3be1ba1e582ca5e

Observation dbc38cef-16cb-40c4-93ca-4cd9ceffcb02 · outbound

This paper cites an unresolved cited work.

MedBLINK: Probing Basic Perception in Multimodal Language Models for Medicine Unresolved cited work

Reference 74

Resolution
unresolved
raw_fallback, observed 2026-08-06T04:50:45.882897Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T04:50:45.366141Z digest=sha256:5a60911645f173a43478b9daba896527ab2919f9bb545f15d3cf530aa8721b38

Pith citing papers

Observation 47c153a4-45ee-4a12-91f4-2dc59ba0543a · inbound

MedBLINK: Probing Basic Perception in Multimodal Language Models for Medicine cites this paper.

MedBLINK: Probing Basic Perception in Multimodal Language Models for Medicine MedBLINK: Probing Basic Perception in Multimodal Language Models for Medicine

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-08-06T04:50:45.814234Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T04:50:45.050116Z digest=sha256:5f90e96d36dd1e0959d86102c37455b8e8f16d8071824573ab3ad3e380f4e21e