Pith. sign in

Paper Citation Record · LEDGER

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable?

As of 13 August 2026, this Paper Citation Record lists 30 of 30 outbound references and 1 inbound Pith citation observation for arXiv:2505.15425.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.15425 v2

Coverage vector

measured 30 of 30 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:21:21.068140Z

measured 31 of 31 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-04T19:13:53.947203Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

30 of 30 outbound references displayed

  • verified exact0
  • verified fuzzy16
  • unresolved10
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch3

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation f9830d4a-77cd-49dc-bb44-3a02a7bea89d · outbound

This paper cites https://pmc.ncbi.nlm.nih.gov/, accessed: 2025-04-05 9.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? https://pmc.ncbi.nlm.nih.gov/, accessed: 2025-04-05 9

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:21:25.441969Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T15:21:17.569003Z digest=sha256:5fdc5d7c6c2d8034b209032b026723e243d04f8f1ed1a88ece38479f75a372b3

Observation ba554940-a11b-4daa-adf3-e009f5bc99d8 · outbound

This paper cites A Survey of Medical Vision-and-Language Applications and Their Techniques.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? A Survey of Medical Vision-and-Language Applications and Their Techniques

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T15:21:17.662259Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:21:17.662259Z digest=sha256:204eda4a2989baa3caab741929b00b982e0b0a766b31af2d8dd310a02effeb6b

Observation bf83a1f4-5215-408a-b498-db43f58fb80e · outbound

This paper cites In: ICML Workshop on Computational Biology (2021) 2, 9, 20.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? In: ICML Workshop on Computational Biology (2021) 2, 9, 20

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:21:25.260674Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T15:21:17.808789Z digest=sha256:5cf1021bce0c5e50e0ab0d9e3a0a9dd4c991ede2e2fba7e6b45c6382efe78e6b

Observation 915abbaa-6cd8-407d-94a5-815b401de255 · outbound

This paper cites Benchmarking Robustness of Contrastive Learning Models for Medical Image-Report Retrieval.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? Benchmarking Robustness of Contrastive Learning Models for Medical Image-Report Retrieval

Reference 4

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T15:21:21.850405Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T15:21:17.873767Z digest=sha256:6f0dad6be547c44b911223fa8bb478e93cb69db9d947a37433366bd7ac1738e6

Observation 33df9c7a-f727-447c-b8b0-3f4c76bdfce1 · outbound

This paper cites MedMNIST-C: Comprehensive benchmark and improved classifier robustness by simulating realistic image corruptions.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? MedMNIST-C: Comprehensive benchmark and improved classifier robustness by simulating realistic image corruptions

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T15:21:17.967185Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:21:17.967185Z digest=sha256:5d5dc16d525298a03702c6b228af4ae4d42bb8afd0827de5d34e55aea84f42ab

Observation 43739308-acd1-498c-aec4-5ee90fcd9201 · outbound

This paper cites In: 2025 IEEE 22nd In- ternational Symposium on Biomedical Imaging (ISBI).

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? In: 2025 IEEE 22nd In- ternational Symposium on Biomedical Imaging (ISBI)

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:21:25.046793Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T15:21:18.102723Z digest=sha256:c23644bb36365eec65536415f5fd44f10469dce27f2f1416857b6f73fb9e4037

Observation 90a6e254-9ef8-4767-a418-54fd3f10fecd · outbound

This paper cites In: International Conference on Medical Image Computing and Computer-Assisted Intervention.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? In: International Conference on Medical Image Computing and Computer-Assisted Intervention

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:21:24.860908Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T15:21:18.204355Z digest=sha256:3dfa38181d9cfb77a13ffe75815a982734009cc6b91b421fc801a95c6bcefd42

Observation 956ac906-776d-4e26-aaf0-f4fd2f77d46c · outbound

This paper cites In: International Conference on Medical Image Computing and Computer-Assisted Intervention.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? In: International Conference on Medical Image Computing and Computer-Assisted Intervention

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:21:24.653457Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T15:21:18.321901Z digest=sha256:4709264c36afb422792dc3866d93029d2507ea5109aebc70f2c0389989e45a0b

Observation c53fdf87-3b86-48f8-91ef-d0c64306771a · outbound

This paper cites MedFuse: Multi-modal fusion with clinical time-series data and chest X-ray images.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? MedFuse: Multi-modal fusion with clinical time-series data and chest X-ray images

Reference 9

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T15:21:21.675948Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T15:21:18.499866Z digest=sha256:f61cdfd41a437bbdaaf7410f59ad92bd572d73cc2bdc9f0308de8152f9dbbb0b

Observation 58959005-40ed-403b-83fa-087010bf6644 · outbound

This paper cites In: ICLR (2019) 2.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? In: ICLR (2019) 2

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:21:24.434967Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T15:21:18.635490Z digest=sha256:8c7a026fe05755068de04e5cefc16f2775a44ce611a8e9fbed2998ba39f5b354

Observation c6332889-c3c1-4d21-a701-e48ba1998e42 · outbound

This paper cites ICLR1(2), 3 (2022) 8.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? ICLR1(2), 3 (2022) 8

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:21:24.253117Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T15:21:18.776022Z digest=sha256:9728eed012b583a897c32f347027db96d901148412b0c6de5bec507adb49cef6

Observation 4ef916d3-dec1-4414-b183-18395aa5422f · outbound

This paper cites Test-Time Low Rank Adaptation via Confidence Maximization for Zero-Shot Generalization of Vision-Language Models.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? Test-Time Low Rank Adaptation via Confidence Maximization for Zero-Shot Generalization of Vision-Language Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T15:21:18.969671Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:21:18.969671Z digest=sha256:9d4b35e834785e09aa39aa19c5d69240f1b719dd404e474ea4d4cefe32ebaaf5

Observation bd4549b0-3e2a-421f-bf14-7c5a85841c67 · outbound

This paper cites Noise is an Efficient Learner for Zero-Shot Vision-Language Models.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? Noise is an Efficient Learner for Zero-Shot Vision-Language Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T15:21:19.132444Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:21:19.132444Z digest=sha256:f19f87c68274ea76fe020a49317b5656f1661912a2d2d9f3dc5534f700ce19ba

Observation f675dad8-f92c-4e60-b223-63bef8bb45ef · outbound

This paper cites Proceedings of the AAAI Conference on Artificial Intelligence (2019) 2.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? Proceedings of the AAAI Conference on Artificial Intelligence (2019) 2

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:21:23.961842Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T15:21:19.292556Z digest=sha256:d0d819d115c579d44d853ed32341cd1b229c5bcb6f38e936c5005d6178582fed

Observation dc13bbbe-e59a-4483-92ac-88d0c26f5846 · outbound

This paper cites MIMIC-CXR-JPG, a large publicly available database of labeled chest radiographs.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? MIMIC-CXR-JPG, a large publicly available database of labeled chest radiographs

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T15:21:19.434663Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:21:19.434663Z digest=sha256:56d548c8c8132f95c42af02a15c3dfe8085ef85e76117abf93dabaa779e78e2e

Observation 1ef37400-0521-4c12-a2e3-0e6c5992fbd1 · outbound

This paper cites Radiology p.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? Radiology p

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:21:23.735407Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T15:21:19.547418Z digest=sha256:a391fe5ed5237e31723bcbeefa3f8c185d931d405b2e6052b7ece2ef93705fc0

Observation add91f25-6bb1-405f-bc53-8dc00554c113 · outbound

This paper cites IEEE Reviews in Biomedical Engineering (2025) 2.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? IEEE Reviews in Biomedical Engineering (2025) 2

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:21:23.484419Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T15:21:19.662098Z digest=sha256:3294b288bbeea42ac2fc89de5dd6266a0ab22abe681eebe1471ada48f2a83ae3

Observation 85331b14-0378-4f8b-b6ba-70500ef184b5 · outbound

This paper cites UniMed-CLIP: Towards a Unified Image-Text Pretraining Paradigm for Diverse Medical Imaging Modalities.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? UniMed-CLIP: Towards a Unified Image-Text Pretraining Paradigm for Diverse Medical Imaging Modalities

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T15:21:19.892988Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:21:19.892988Z digest=sha256:ca222747f7414c434055715c99e3a63c69bedfde5831c22a034b1f5daab8b7aa

Observation fcfb872b-e090-4485-911c-a27521c16616 · outbound

This paper cites Journal of Biomedical Informatics135, 104234 (2022) 4.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? Journal of Biomedical Informatics135, 104234 (2022) 4

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:21:23.210950Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T15:21:19.989329Z digest=sha256:73428eba79bd762e9d0e74e89f140dc48370de30d9053f83638d47362eee892c

Observation d36f2d04-e14a-40b2-96e8-927b071b9a66 · outbound

This paper cites In: International Conference on Medical Image Computing and Computer-Assisted Intervention.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? In: International Conference on Medical Image Computing and Computer-Assisted Intervention

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:21:22.936325Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T15:21:20.112534Z digest=sha256:1474200f3d70fe6751446053d73707ad7a4098c9ab0d8b1c2fbd41fdf9e6a914

Observation f70c6676-9a38-4cb1-9d3d-3cce8a6d8c35 · outbound

This paper cites Medical Image Analysis 58, 101562 (2019).

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? Medical Image Analysis 58, 101562 (2019)

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T15:21:20.205494Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:21:20.205494Z digest=sha256:c54d80742653e6155d5c47eabd3bd561822fe9655c12d6159fd108160243e111

Observation 78361df4-daa6-42ef-b0aa-f166aacf199c · outbound

This paper cites On Evaluating Adversarial Robustness of Volumetric Medical Segmentation Models.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? On Evaluating Adversarial Robustness of Volumetric Medical Segmentation Models

Reference 22

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T15:21:21.410077Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T15:21:20.305953Z digest=sha256:143bb7f3c78696e617d69cf0f325acbed95769d13aef669b8abe6fc755d1292b

Observation 2ee9eb48-4ba0-4738-8125-fca4412923ce · outbound

This paper cites In: ICML (2021) 4, 9, 19.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? In: ICML (2021) 4, 9, 19

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:21:22.715955Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T15:21:20.429736Z digest=sha256:43840faa3a5d9552f9837ce933a6bbd2062dbc3b3fdbd5169f28250fd4b31d19

Observation 245362ca-8fb4-49ea-a018-1636c0a2a0eb · outbound

This paper cites Seco de Herrera, A., et al.: Rocov2: Radiol- ogy objects in context version 2, an updated multimodal image dataset.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? Seco de Herrera, A., et al.: Rocov2: Radiol- ogy objects in context version 2, an updated multimodal image dataset

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:21:22.529248Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T15:21:20.528599Z digest=sha256:32ff0a13d7cab525e2796ea304680721b9bbd51b9b571747aa0ae86b703736a4

Observation 0e4732d4-e1a2-4b31-a9eb-378084a98ea3 · outbound

This paper cites In: Proceedings of the IEEE/CVF conference on computer vision and pattern recognition.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? In: Proceedings of the IEEE/CVF conference on computer vision and pattern recognition

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:21:22.339789Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T15:21:20.597170Z digest=sha256:beb40e4f536337392d86f7e1aaa21201351b43d9d9bbe71b046f688bd764f496

Observation 6fdd2f16-5734-48ce-a493-ac6b85ea37e4 · outbound

This paper cites MedCLIP: Contrastive Learning from Unpaired Medical Images and Text.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? MedCLIP: Contrastive Learning from Unpaired Medical Images and Text

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T15:21:20.695492Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:21:20.695492Z digest=sha256:7b43db3c2daebe338b0bf1b1d06451ee5ae783ed7d23106c29beb16541e54734

Observation e30919d9-7ff2-4ffb-9a2e-df4ebc6656f8 · outbound

This paper cites an unresolved cited work.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? Unresolved cited work

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T15:21:20.803512Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:21:20.803512Z digest=sha256:6e230595f4b66bd85d13ba9c69e12162a421650428708073e783dfab9079fa8a

Observation 8f4f9e26-01f8-41f2-83ee-40643a67c91c · outbound

This paper cites Frontiers in neuroinformatics13, 46 (2019) 2, 4, 5.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? Frontiers in neuroinformatics13, 46 (2019) 2, 4, 5

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:21:22.173875Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T15:21:20.895536Z digest=sha256:5389bd8e63a21c8451884e336619856e69340b1b6c184af3012911c3327add1a

Observation 69be5949-9051-4277-b855-3fa3e2acfa69 · outbound

This paper cites BiomedCLIP: a multimodal biomedical foundation model pretrained from fifteen million scientific image-text pairs.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? BiomedCLIP: a multimodal biomedical foundation model pretrained from fifteen million scientific image-text pairs

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T15:21:20.976948Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:21:20.976948Z digest=sha256:67605f34c372574c81f22658601fe49276557b8dddd575889535d18ee6dc5e6c

Observation f2ecfa22-c5ac-4eb4-9fb4-792d2754a503 · outbound

This paper cites In-Distribution.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? In-Distribution

Reference 30

Resolution
malformed identifier
raw_fallback, observed 2026-08-07T15:21:22.036992Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T15:21:21.068140Z digest=sha256:59c577432f94860ad8a0bec57e0b1ae1dfca6500b3759bb590d9d9d94828cd5b

Pith citing papers

Observation e131d344-9f9e-46f7-9fe1-dff80d68d595 · inbound

Decoupling Clinical and Class-Agnostic Features for Reliable Few-Shot Adaptation under Shift cites this paper.

Decoupling Clinical and Class-Agnostic Features for Reliable Few-Shot Adaptation under Shift On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable?

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-04T19:13:53.947203Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:13:53.947203Z digest=sha256:453728ca4d8056d6a51546f98ba7f1a39e825e82d7bc600d4e122342c9abd5ae