Pith. sign in

Paper Citation Record · LEDGER

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models

As of 21 August 2026, this Paper Citation Record lists 56 of 56 outbound references and 0 inbound Pith citation observations for arXiv:2506.08990.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.08990 v1

Coverage vector

measured 56 of 56 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T05:02:21.165503Z

measured 56 of 56 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

56 of 56 outbound references displayed

  • verified exact0
  • verified fuzzy27
  • unresolved29
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 9fd25426-9bcc-4eb4-bee7-b13f40e3ed74 · outbound

This paper cites Imagenet: A large-scale hierarchical image database,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Imagenet: A large-scale hierarchical image database,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:20.773766Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:20.773766Z digest=sha256:b7e0c0d35b3f5e6a3cde26a4f02ae9cc0166144d4866bf62abd866ca2b250357

Observation 247791b4-f9dd-4f94-8849-25c385551f2c · outbound

This paper cites MIMIC-CXR-JPG, a large publicly available database of labeled chest radiographs.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models MIMIC-CXR-JPG, a large publicly available database of labeled chest radiographs

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:20.823905Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:20.823905Z digest=sha256:833a1011c1ab4b6b9bc23c3a974d81b789b0c0084e939db7e8b6ef19856521c0

Observation 9de23ca4-edcd-4f77-a29c-69e2200a2b5a · outbound

This paper cites Chestx-ray8: Hospital-scale chest x-ray database and benchmarks on weakly-supervised classification and localization of common thorax diseases,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Chestx-ray8: Hospital-scale chest x-ray database and benchmarks on weakly-supervised classification and localization of common thorax diseases,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:22.003020Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T05:02:20.873856Z digest=sha256:6717f01d0a15abb919dc383323b74b3d627eb9f41a657f788ff508f47658c976

Observation dad709e1-77fc-482f-9a7b-e01b3d739726 · outbound

This paper cites Chexpert: A large chest radiograph dataset with uncertainty labels and expert comparison,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Chexpert: A large chest radiograph dataset with uncertainty labels and expert comparison,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:21.986416Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T05:02:20.929274Z digest=sha256:9d885a2d683e137f4c347fb20f081806568c1d88dcdd8183b0876d1558f8c02d

Observation e8d136c1-2ca1-422c-9dac-c5d1403ff169 · outbound

This paper cites Learning transferable visual models from natural language supervision,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Learning transferable visual models from natural language supervision,

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:20.949952Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:20.949952Z digest=sha256:5af834956f954a976d74dc683d708ae450f97ec14063269c37878882b154b758

Observation 76f20c04-0ca6-4e17-98a2-cea9eab6695c · outbound

This paper cites Contrastive learning of medical visual representations from paired images and text,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Contrastive learning of medical visual representations from paired images and text,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:21.961508Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T05:02:20.954272Z digest=sha256:87d77df8d078e57fa2b9c5286be58c1d570b776548ed8bb8d51a5307ec081a7b

Observation 72dc1b6d-e7a1-497c-824a-7adabb02216a · outbound

This paper cites Gloria: A multimodal global-local representation learning framework for label- efficient medical image recognition,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Gloria: A multimodal global-local representation learning framework for label- efficient medical image recognition,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:21.948018Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T05:02:20.959477Z digest=sha256:6e3c6015c55eeada6f214ac8ef5f0a6039b6812a52bc5f050b1316df60b70db2

Observation 45fb1f3a-ccf5-48ad-995d-7bbbca04886e · outbound

This paper cites Making the most of text semantics to improve biomedical vision– language processing,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Making the most of text semantics to improve biomedical vision– language processing,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:20.963272Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:20.963272Z digest=sha256:ecf26311b2fad584b533bc0c5893619746c3ec0a500aa67c04ebb9de8d4ccffc

Observation e3ef6427-e5b4-4135-9c46-d39e5e22dc4f · outbound

This paper cites A multimodal biomedical foundation model trained from fifteen million image–text pairs,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models A multimodal biomedical foundation model trained from fifteen million image–text pairs,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:21.923279Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T05:02:20.967837Z digest=sha256:5e9870ff98d98158e21e7de7abeb2059760ade92b504cf1e6263e2c75652afa2

Observation ac4bc219-aaac-4734-8b41-47db5935ee65 · outbound

This paper cites Advancing radiograph representation learning with masked record modeling,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Advancing radiograph representation learning with masked record modeling,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:21.909308Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T05:02:20.971917Z digest=sha256:d01d306feb7b4aec678729f47406615e33285f9f91cebdd190627cb6859ba688

Observation 1944ea8f-beef-46cf-a372-fdb3f8bab1c2 · outbound

This paper cites Exploring scalable medical image encoders beyond text supervision.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Exploring scalable medical image encoders beyond text supervision

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:20.975873Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:20.975873Z digest=sha256:9bd865d54bc493bb399532e4af3143a42f8bc589f5539719e492ae487048cb4e

Observation cf92113c-2d18-48dd-ab23-d108399a51d8 · outbound

This paper cites Medical Vision Language Pretraining: A survey.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Medical Vision Language Pretraining: A survey

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:20.981003Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:20.981003Z digest=sha256:aac1074d29d58965d9b3674ec4747347169b9d729eb9955457de9bb765cd29ae

Observation 1e88a606-8554-46d1-b82d-162313e4065c · outbound

This paper cites Unichest: Conquer-and-divide pre-training for multi-source chest x-ray classifica- tion,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Unichest: Conquer-and-divide pre-training for multi-source chest x-ray classifica- tion,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:21.895563Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T05:02:20.985555Z digest=sha256:8038de1f4febe4c1de46f0c360240ff33d3c391a60b085f790698ce179e3b434

Observation 36067afd-f3b0-47bd-aa4f-44169d599346 · outbound

This paper cites Im- proving medical vision-language contrastive pretraining with semantics- aware triage,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Im- proving medical vision-language contrastive pretraining with semantics- aware triage,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:21.881471Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T05:02:20.989976Z digest=sha256:91fecf731f79fc4b633a505e7a4c180e8ccecdd6a59ca6ad6897a3a318074844

Observation 6d72bb59-ce72-44a4-9a8c-fad5114fa252 · outbound

This paper cites Exploring scalable medical image encoders beyond text supervision,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Exploring scalable medical image encoders beyond text supervision,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:21.867278Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T05:02:20.993776Z digest=sha256:80bc21fbb540fb6e40c18ea385ac62330f43b0e6054e48a75304f25048054f03

Observation 9d5e6252-eb2a-4858-b259-dda8df479418 · outbound

This paper cites Learning to exploit temporal structure for biomedical vision-language processing,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Learning to exploit temporal structure for biomedical vision-language processing,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:20.998341Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:20.998341Z digest=sha256:7058e0bf3248ef0cce77391a06fe46e3d9c54cbcb8b61834dbc387c017bfb3c3

Observation 984b26ad-05cb-41e8-b24a-31d5bad53fe8 · outbound

This paper cites Gener- alized radiograph representation learning via cross-supervision between images and free-text radiology reports,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Gener- alized radiograph representation learning via cross-supervision between images and free-text radiology reports,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:21.842083Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T05:02:21.002149Z digest=sha256:4c8ec97494e9b55561e3a1c248890e0c143a33ac1ba09bac6cc2da2ba7024d07

Observation 331a124b-a0f9-40ab-be09-e1410669366b · outbound

This paper cites Chexrelnet: An anatomy-aware model for tracking longi- tudinal relationships between chest x-rays,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Chexrelnet: An anatomy-aware model for tracking longi- tudinal relationships between chest x-rays,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:21.826179Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T05:02:21.006246Z digest=sha256:06f52e43f52731a55873ef407b25d02e6eec705361585e616a4830eef6378fc9

Observation 2390c2ce-d423-4c67-9e85-c087b41cd959 · outbound

This paper cites Chexfusion: Effective fusion of multi-view features using transformers for long-tailed chest x-ray classification,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Chexfusion: Effective fusion of multi-view features using transformers for long-tailed chest x-ray classification,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:21.809530Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T05:02:21.010139Z digest=sha256:9a7d385bc2e373a50137cdcf4c3d482d98d4c841cdcc75de5b99cdd7365602a2

Observation 9a6e642a-3f72-4cf6-89db-62c0af89edad · outbound

This paper cites Act like a radiologist: towards reliable multi-view correspondence reasoning for mammogram mass detection,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Act like a radiologist: towards reliable multi-view correspondence reasoning for mammogram mass detection,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:21.793117Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T05:02:21.014390Z digest=sha256:6d0ce6b2b13784fbd95ab15063514fe8ae937a81ef853e26f11bb3e56e72a98a

Observation 97f3da43-ea3f-477b-a633-d5d0676e25b3 · outbound

This paper cites Deltanet: Conditional medical report generation for covid- 19 diagnosis,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Deltanet: Conditional medical report generation for covid- 19 diagnosis,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:21.775166Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T05:02:21.019327Z digest=sha256:92973eaf70f668b8159b528a220a592bfe9ba5ba072d021fb530a9f57c8909e0

Observation 67f3d0d5-0c5d-4ed6-ab47-b094e5d444cc · outbound

This paper cites Self-supervised learning for medical image analysis using image context restoration,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Self-supervised learning for medical image analysis using image context restoration,

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:21.023679Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:21.023679Z digest=sha256:4aa8a954f918179d761c3ca2a806c989de64dff09b27d242cad0f180522fa85e

Observation 9266745b-4665-4de7-9dab-56ec3fdc579f · outbound

This paper cites Models genesis,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Models genesis,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:21.745864Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T05:02:21.027902Z digest=sha256:74ce96eae852b5feb359061197b03d6801686dbc9ddf2abbb4e49b6a8a6fb37b

Observation 64b13a44-3f32-4af8-b3c9-a5ec1d543910 · outbound

This paper cites Comparing to learn: Surpassing imagenet pretraining on radiographs by comparing image representations,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Comparing to learn: Surpassing imagenet pretraining on radiographs by comparing image representations,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:21.728548Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T05:02:21.032152Z digest=sha256:fefb9054a161d775dfb1b593ca262496a3f22a814ebb61b067b06bd95263818f

Observation 092cee43-8f44-407b-85d3-08a0d796bb1e · outbound

This paper cites Blip: Bootstrapping language-image pre-training for unified vision-language understanding and generation,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Blip: Bootstrapping language-image pre-training for unified vision-language understanding and generation,

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:21.035993Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:21.035993Z digest=sha256:f8bba80673509f63999826dd4272c599c8b144340dc74b1c5d581b804edd9d3c

Observation 825e5557-90db-47c7-8405-3d29a0b86a88 · outbound

This paper cites Vilt: Vision-and-language transformer without convolution or region supervision,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Vilt: Vision-and-language transformer without convolution or region supervision,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:21.700519Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T05:02:21.040046Z digest=sha256:e2162b3da10c127dec11ada825f7f4b09da5f7074f3a21e0acb51ea77e89c557

Observation da009c55-246b-49ff-88d3-e3ca112e52dc · outbound

This paper cites Align before fuse: Vision and language representation learning with momentum distillation,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Align before fuse: Vision and language representation learning with momentum distillation,

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:21.044391Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:21.044391Z digest=sha256:18f34443eabac11805569346dfb149709b1546b2c5c89c6938221e067befe62a

Observation 31c42f06-88bd-4af9-8adb-30ca8ec081aa · outbound

This paper cites Image as a Foreign Language: BEiT Pretraining for All Vision and Vision-Language Tasks.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Image as a Foreign Language: BEiT Pretraining for All Vision and Vision-Language Tasks

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:21.048709Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:21.048709Z digest=sha256:fec75a6126ced2c64c9100ed7d04bacc6fc45f38932b4c3ed4238ba5b59b6171

Observation 6c061e64-054d-4aa8-a068-8b045d806f30 · outbound

This paper cites Knowledge- enhanced visual-language pre-training on chest radiology images,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Knowledge- enhanced visual-language pre-training on chest radiology images,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:21.672655Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T05:02:21.052954Z digest=sha256:39b128fa12fde430f5aec0b72b401fbc3299e08a5d95d82fb2985e46d31a8688

Observation 82213d81-baba-48b2-bfa3-c8891ca4ba0b · outbound

This paper cites Carzero: Cross-attention alignment for radiology zero-shot classifica- tion,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Carzero: Cross-attention alignment for radiology zero-shot classifica- tion,

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:21.056753Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:21.056753Z digest=sha256:595139e2df921e8350e316eb676d4d2a2afb356b75a87e43ffa27fcf83ebc4e9

Observation f7a18f76-d20c-4853-af0a-549b7b906163 · outbound

This paper cites Enhancing representation in radiography- reports foundation model: A granular alignment algorithm using masked contrastive learning,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Enhancing representation in radiography- reports foundation model: A granular alignment algorithm using masked contrastive learning,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:21.643981Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T05:02:21.060464Z digest=sha256:80233c58dcdec3138c361e0c73d9a4a45c1f5023ac5cc890bd2112340117518f

Observation cf44f368-257a-4fed-8d31-577fdf9d7fc8 · outbound

This paper cites Medklip: Medical knowledge enhanced language-image pre-training for x-ray diagnosis,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Medklip: Medical knowledge enhanced language-image pre-training for x-ray diagnosis,

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:21.064175Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:21.064175Z digest=sha256:1f2040ae89dc30e4e108a7886726f8fbf1fdf96d373dd9a9213d8bda46cd8a4e

Observation e597b183-f5b7-4d40-becd-542f72198e17 · outbound

This paper cites Parameter-efficient transfer learning for nlp,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Parameter-efficient transfer learning for nlp,

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:21.068498Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:21.068498Z digest=sha256:b04844c47664629982454e29f2521dcdf619df2fbea9758fb124200203bf2f90

Observation 4dd1dd1f-afbf-4c0f-9aed-886cfd4e1571 · outbound

This paper cites LoRA: Low-Rank Adaptation of Large Language Models.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models LoRA: Low-Rank Adaptation of Large Language Models

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:21.072896Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:21.072896Z digest=sha256:8ee13bada2cd390c6bbabfc3ef1a1810272673a293f4088f60e861b0558459a1

Observation 2414083d-5b26-41e4-b267-929771e16a86 · outbound

This paper cites Few-shot parameter-efficient fine-tuning is better and cheaper than in-context learning,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Few-shot parameter-efficient fine-tuning is better and cheaper than in-context learning,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:21.600504Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T05:02:21.077066Z digest=sha256:e67c0bf49d640839e686aca7f27007458fed370d2d81aac29cbc21fac5901a52

Observation adf02300-4504-4571-be33-5c3b7c460cf1 · outbound

This paper cites Bitfit: Simple parameter- efficient fine-tuning for transformer-based masked language-models,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Bitfit: Simple parameter- efficient fine-tuning for transformer-based masked language-models,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:21.582459Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T05:02:21.081187Z digest=sha256:09ed612da70146a258e519edad35df49e76e88b60ad81a85a7956b55da369ffe

Observation 582f3bc9-43a9-4b30-80ad-3c3bb71c53fa · outbound

This paper cites Towards a Unified View of Parameter-Efficient Transfer Learning.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Towards a Unified View of Parameter-Efficient Transfer Learning

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:21.085120Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:21.085120Z digest=sha256:d469f2387fe495e4614ab66aea1c2e71b21eaebdf6894aab3b567929c0e308ed

Observation 6d08116e-f5cd-4dc2-9991-c6c6cbcf7a3d · outbound

This paper cites Visual prompt tuning,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Visual prompt tuning,

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:21.089229Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:21.089229Z digest=sha256:bd2849cb2238489c2db1fd0913d590ad3b70cf984bffb38a1d885ddfc300200e

Observation 9cb90014-3299-4f7d-bbe1-e5906504b314 · outbound

This paper cites Vl-adapter: Parameter-efficient transfer learning for vision-and-language tasks,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Vl-adapter: Parameter-efficient transfer learning for vision-and-language tasks,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:21.557861Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T05:02:21.093214Z digest=sha256:732c14bbdf14f3d63be29ac9c127844b8c52a74a8383578db64754882a0cbc2b

Observation cd240235-3b9b-41ee-b783-60e85f662059 · outbound

This paper cites AIM: Adapting Image Models for Efficient Video Action Recognition.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models AIM: Adapting Image Models for Efficient Video Action Recognition

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:21.096816Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:21.096816Z digest=sha256:298d34894d4f64d26134eb08cf77fa53db8e17919bd16e0200e2154beb82bcdd

Observation b1b5eee8-d624-4a21-8933-82b1f20b639e · outbound

This paper cites Vision Transformer Adapter for Dense Predictions.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Vision Transformer Adapter for Dense Predictions

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:21.100923Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:21.100923Z digest=sha256:ed3c135d9aa14dc2e28992407dbaf45a560957176ee75e279a686d077a66900e

Observation f0189b50-bdb5-4c9c-b43f-64fff44af3f5 · outbound

This paper cites Parameter-Efficient Fine-Tuning for Medical Image Analysis: The Missed Opportunity.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Parameter-Efficient Fine-Tuning for Medical Image Analysis: The Missed Opportunity

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:21.105209Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:21.105209Z digest=sha256:59473d338673ba66ada7553310f259833fdb0c10090c9c1024304ff1473ee8d6

Observation 3cf57145-50b6-4376-b42d-feb2aa78d11f · outbound

This paper cites MeLo: Low-rank Adaptation is Better than Fine-tuning for Medical Image Diagnosis.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models MeLo: Low-rank Adaptation is Better than Fine-tuning for Medical Image Diagnosis

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:21.109848Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:21.109848Z digest=sha256:a57410693401401bccb3a69512f5c32a3c266b4f4ad25fbd2d27e9ad05fd885e

Observation 940aa1de-6d6a-46ff-a5c0-aaa6968df2bc · outbound

This paper cites Less Could Be Better: Parameter-efficient Fine-tuning Advances Medical Vision Foundation Models.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Less Could Be Better: Parameter-efficient Fine-tuning Advances Medical Vision Foundation Models

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:21.115036Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:21.115036Z digest=sha256:dc5299b86e982065b172b3ccf70deef3555aab97f5b0bf2b5b3fb6ce002574f0

Observation 9c12a0b7-ad35-4518-b3c9-b359c796cb09 · outbound

This paper cites Prompt tuning for parameter- efficient medical image segmentation,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Prompt tuning for parameter- efficient medical image segmentation,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:21.542188Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T05:02:21.119838Z digest=sha256:e3c6562324eaa9c9c886cfe7ec56cd38227f277a07719c7c7ffab916798f42dd

Observation 12073882-8e25-47fa-b8f3-13c49385940a · outbound

This paper cites Towards foundation models and few-shot parameter-efficient fine-tuning for volumetric organ seg- mentation,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Towards foundation models and few-shot parameter-efficient fine-tuning for volumetric organ seg- mentation,

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:21.525535Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T05:02:21.123971Z digest=sha256:73f300f27afa06505526c04222e3e4404019a5fcebaae3709e433c65fdf8beac

Observation fd6089af-958e-43a2-98df-a43146bd29da · outbound

This paper cites Contrastive Alignment of Vision to Language Through Parameter-Efficient Transfer Learning.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Contrastive Alignment of Vision to Language Through Parameter-Efficient Transfer Learning

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:21.128128Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:21.128128Z digest=sha256:c118d17e1dfd178e9667ce592729f7339751f25a6e8e21426d7fff387785453e

Observation dfda98a6-9006-4d5c-a902-26e0b1555765 · outbound

This paper cites Scaling language- image pre-training via masking,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Scaling language- image pre-training via masking,

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:21.509160Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T05:02:21.132633Z digest=sha256:a4c2d0093bd465dcbc5a5fee721d5f96dd5b944e1bf84e01e63eee53c9c47b83

Observation 677245e6-b7da-4913-b2a4-b271c50526f2 · outbound

This paper cites Attention is all you need,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Attention is all you need,

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:21.136469Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:21.136469Z digest=sha256:459f3a5d7e1ecb16bc40228f10b23a98bfdfa16e7cf0ce59a72c866c24c2cba3

Observation e36c3afa-c4bb-4bf9-af8b-6324a2c96d89 · outbound

This paper cites Gaussian Error Linear Units (GELUs).

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Gaussian Error Linear Units (GELUs)

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:21.140339Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:21.140339Z digest=sha256:a9ab612b8e1ffbf5a5b3ceba45bfe417d3ba5abe15ad682a155d8375c3df0fbb

Observation 56ec48d9-6ab1-429d-8320-7828089505f6 · outbound

This paper cites Mvco- dot: Multi-view contrastive domain transfer network for medical report generation,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Mvco- dot: Multi-view contrastive domain transfer network for medical report generation,

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:21.485757Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T05:02:21.144569Z digest=sha256:2bbc2f57d055771ed8a0e4beaea7f6a6fb56c8aa1a30406d1aad747e7bfe7306

Observation 88e534ca-45b7-4aef-b061-c2390706973e · outbound

This paper cites Representation Learning with Contrastive Predictive Coding.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Representation Learning with Contrastive Predictive Coding

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:21.149064Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:21.149064Z digest=sha256:674d9efa68991a736340414cce7af1910c07367016d98071d214a0dc249d2afa

Observation ed74d849-4d17-493c-90d0-3a5ca81c09aa · outbound

This paper cites Augmenting the national institutes of health chest radiograph dataset with expert annotations of possible pneumonia,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Augmenting the national institutes of health chest radiograph dataset with expert annotations of possible pneumonia,

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:21.470237Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T05:02:21.153308Z digest=sha256:b386d3e8e8b539d5ac4d5df770439a79b516d22e9553ee3cf5faed58dc556048

Observation c3d7a77f-16a5-4936-be96-df9a5d821719 · outbound

This paper cites Improving Factual Completeness and Consistency of Image-to-Text Radiology Report Generation.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Improving Factual Completeness and Consistency of Image-to-Text Radiology Report Generation

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:21.157303Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:21.157303Z digest=sha256:529d189a3dd34fc5742d141db15a51f263e54d7fa20d3eb31c1dbeb119591f4d

Observation 01006229-228d-4ad5-9916-e3da28397fdb · outbound

This paper cites Pytorch: An imperative style, high-performance deep learning library,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Pytorch: An imperative style, high-performance deep learning library,

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:21.161272Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:21.161272Z digest=sha256:8092ec0e0e3dd0cc64e4bf012522df633f29a61a334e3194e58f70b4417bfcff

Observation 75493d61-a82c-408a-8cd7-dfe7fe39055b · outbound

This paper cites Decoupled Weight Decay Regularization.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Decoupled Weight Decay Regularization

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:21.165503Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:21.165503Z digest=sha256:e6162d05ec0796387de83f0aa1a1a7a5397ac0c722d78f4cc77903b5ae0198ad

Pith citing papers

No inbound Pith citation observations are available.