Pith. sign in

Paper Citation Record · LEDGER

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images

As of 9 August 2026, this Paper Citation Record lists 52 of 52 outbound references and 1 inbound Pith citation observation for arXiv:2502.05928.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.05928 v4

Coverage vector

measured 52 of 52 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-08T17:26:28.903390Z

measured 53 of 53 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-27T13:08:02.505489Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T05:47:40.979179Z

Reference resolution

52 of 52 outbound references displayed

  • verified exact2
  • verified fuzzy20
  • unresolved30
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation e94c8803-b6cf-4fd2-9228-7b62e17d2d6e · outbound

This paper cites Hasan, Vivek Datla, Joey Liu, Dina Demner-Fushman, and Henning Müller.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images Hasan, Vivek Datla, Joey Liu, Dina Demner-Fushman, and Henning Müller

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:26:29.828526Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T17:26:28.724936Z digest=sha256:d1099758bb0fecf098f4e8b0f69fdec8333ca4b3ddb7753e5248a9ae7208d1a5

Observation fa800f7b-5c6f-4c94-b913-93a123ec5cae · outbound

This paper cites Spice: Semantic propositional image caption evaluation.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images Spice: Semantic propositional image caption evaluation

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:26:29.818384Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T17:26:28.728940Z digest=sha256:89abacd8c3dd5a8c5fdf9986183cc8a812e75732e507e8bf6ec4a1331901578d

Observation 540e8e91-458f-4f1e-91e6-d3103796147a · outbound

This paper cites Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-08T17:26:28.732983Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:26:28.732983Z digest=sha256:821f3752e920981486fea1ee654458976438578a08b6a51908b8e1bb5d890bb2

Observation 1e2dd7d5-2972-44cf-9b51-239aa2f9f2ac · outbound

This paper cites The revolution of multimodal large language models: A survey.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images The revolution of multimodal large language models: A survey

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:26:29.807623Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T17:26:28.737058Z digest=sha256:aa99b1dec6305176fe978237d86e5ea84ceccdbb5fc5803a61abc4a067a71e36

Observation c5800137-91d9-42a9-98e2-38c9fa5183e6 · outbound

This paper cites R-LLaVA: Improving Med-VQA Understanding through Visual Region of Interest.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images R-LLaVA: Improving Med-VQA Understanding through Visual Region of Interest

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-08T17:26:28.740621Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:26:28.740621Z digest=sha256:c7aaf8df3a52d5ba3621a0a8982a484858d38fe2d863dd15d9281a0958a268b6

Observation e3cbf274-c92d-4935-bcc2-06816c407c19 · outbound

This paper cites Mixed Pseudo Labels for Semi-Supervised Object Detection.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images Mixed Pseudo Labels for Semi-Supervised Object Detection

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-08T17:26:28.744479Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:26:28.744479Z digest=sha256:7102ca10a89fcaa00741e033f2ade8d436d5a1daee08f2c1dbae8f261a8d3b4f

Observation 5ae6695e-eac8-43b9-9496-ea7ad6022937 · outbound

This paper cites SAM-Med2D.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images SAM-Med2D

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-08T17:26:28.748689Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:26:28.748689Z digest=sha256:2c2667025417d85c4cf0e2b01c2a606bb5e57db89bc2d600e16afb3412a123e9

Observation 1ca61755-d128-45b1-8520-0ecce10fa243 · outbound

This paper cites an unresolved cited work.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-08T17:26:29.795526Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T17:26:28.752414Z digest=sha256:1216126ee7628214a6f52784883e49dafe3711af8c9fcc492f2a66ad9ef12dc6

Observation 0acab50f-b9fe-49bc-84cd-98e8c32a5ce6 · outbound

This paper cites Enhancing medical VQA with multimodal determination rationales.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images Enhancing medical VQA with multimodal determination rationales

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:26:29.784785Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T17:26:28.755932Z digest=sha256:72ebe46404d417b3622d4978a85a4e87b6f3324f073dcb548cea13725743788b

Observation 34f346a7-b109-45c5-9f8f-861d1200639a · outbound

This paper cites Hasan, Yuan Ling, Oladimeji Farri, Joey Liu, Henning Müller, and Matthew P.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images Hasan, Yuan Ling, Oladimeji Farri, Joey Liu, Henning Müller, and Matthew P

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:26:29.774228Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T17:26:28.759219Z digest=sha256:c6878741280415134a101d381bb0ef2b2abdfee179170d8036e4f95ca2336df8

Observation d4c6f41c-3374-47b6-8195-7c141732ab98 · outbound

This paper cites Parameter-Efficient Fine-Tuning Medical Multimodal Large Language Models for Medical Visual Grounding.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images Parameter-Efficient Fine-Tuning Medical Multimodal Large Language Models for Medical Visual Grounding

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-08T17:26:28.762817Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:26:28.762817Z digest=sha256:6f73c5a2fc09234d526ffc93962da2267273b609afc76f2c4ae673df3d8bbb2d

Observation 69d37335-c96d-4ef6-98c4-bffdab392d90 · outbound

This paper cites PathVQA: 30000+ Questions for Medical Visual Question Answering.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images PathVQA: 30000+ Questions for Medical Visual Question Answering

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-08T17:26:28.766461Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:26:28.766461Z digest=sha256:c8641558bd4c336298644ef778a638dd772c1570addb6b568ee57f502f988616

Observation db98cd9f-dd53-43a8-a682-9030fbe7127b · outbound

This paper cites Rotary position embedding for vision transformer.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images Rotary position embedding for vision transformer

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:26:29.764002Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T17:26:28.770156Z digest=sha256:1df260d4cd7f7cb51384e8962464f94b66ba2fe559b2f90bb42d433aa32e027d

Observation 23533ab1-26e9-4a57-8ea7-518607cc5574 · outbound

This paper cites Distilling the Knowledge in a Neural Network.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images Distilling the Knowledge in a Neural Network

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-08T17:26:28.773566Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:26:28.773566Z digest=sha256:bab786f1e93b3ee3d7ae49428f9f0f38c785e830d2997102b0101b305822605a

Observation b158e2a1-d394-46b1-a099-d0a7429cd01c · outbound

This paper cites A refer-and-ground multimodal large language model for biomedicine.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images A refer-and-ground multimodal large language model for biomedicine

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:26:29.754097Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T17:26:28.777238Z digest=sha256:48c0da1a33deed711a83f58c6c488b0655446ca78a3f600fc5128cff907cf3c4

Observation 841e7257-7abe-48ce-9a2f-b7a8475f97ab · outbound

This paper cites MedVisionLlama: Leveraging Pre-Trained Large Language Model Layers to Enhance Medical Image Segmentation.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images MedVisionLlama: Leveraging Pre-Trained Large Language Model Layers to Enhance Medical Image Segmentation

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-08-08T17:26:29.468034Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T17:26:28.784274Z digest=sha256:495b029abbcab6e1e31eed6bd01aa38418533b4f51b8a3eab4eb30262b265d37

Observation 32585bb0-b0cc-45e0-bbdc-a68845552b5f · outbound

This paper cites A dataset of clinically generated visual questions and answers about radiology images.Scientific data, 5(1):1–10, 2018.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images A dataset of clinically generated visual questions and answers about radiology images.Scientific data, 5(1):1–10, 2018

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-08T17:26:28.788101Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:26:28.788101Z digest=sha256:73e12dfa1da7e70aa6a9b864fb5b2f8caf17d7bb9ab775d41389aa15f5b4bfb3

Observation 7971a455-4b7a-454b-8ed3-10a139425ea5 · outbound

This paper cites Vision-Language Instruction Tuning: A Review and Analysis.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images Vision-Language Instruction Tuning: A Review and Analysis

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-08T17:26:28.791677Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:26:28.791677Z digest=sha256:7cd9670e98f8743a65a293c9f2b4486da1f5900c806cd7ec8d7dccc93ca676cd

Observation c0c6ff4c-f141-424e-814d-60fd02252925 · outbound

This paper cites Llava-med: Training a large language-and-vision assistant for biomedicine in one day.Advances in Neural Information Processing Systems, 36:28541–28564, 2023.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images Llava-med: Training a large language-and-vision assistant for biomedicine in one day.Advances in Neural Information Processing Systems, 36:28541–28564, 2023

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-08T17:26:28.795486Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:26:28.795486Z digest=sha256:bda52ff8653371c44506293990211cc31bd965d4c7f2f106642024c0d249f1ab

Observation 13f61901-5b5e-486c-92ba-5df82e69fbd1 · outbound

This paper cites Towards visual-prompt temporal answer grounding in instructional video.IEEE Transactions on Pattern Analysis and Machine Intelligence, 2024.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images Towards visual-prompt temporal answer grounding in instructional video.IEEE Transactions on Pattern Analysis and Machine Intelligence, 2024

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:26:29.724444Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T17:26:28.798995Z digest=sha256:36d015e63c3bf5166261ad68e499aa5c156b2f804199849b243f26f7d8bcbe4f

Observation 33438188-404f-4b70-9b29-08705daa8ddf · outbound

This paper cites Pseudo labels for unsupervised domain adaptation: A review.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images Pseudo labels for unsupervised domain adaptation: A review

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:26:29.713921Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T17:26:28.802171Z digest=sha256:ee34b1bd9a85990b667b792391da87232c214a73d082ded88573a96dc714d048

Observation 695be0d4-5b54-4b5f-9a06-ecfb3bfe3028 · outbound

This paper cites A comprehensive survey and guide to multimodal large language models in vision-language tasks.arXiv preprint arXiv:2411.06284, 2024.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images A comprehensive survey and guide to multimodal large language models in vision-language tasks.arXiv preprint arXiv:2411.06284, 2024

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-08T17:26:28.805527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:26:28.805527Z digest=sha256:c8f309a15c4cb27d3bf74301a9164337adc03f03ab64c010dd2626fc0848f451

Observation 48c9db19-c440-426d-ac51-59cfa09cbf05 · outbound

This paper cites Medfilip: Medical fine-grained language-image pre-training.IEEE Journal of Biomedical and Health Informatics, pages 1–11, 2025.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images Medfilip: Medical fine-grained language-image pre-training.IEEE Journal of Biomedical and Health Informatics, pages 1–11, 2025

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:26:29.702840Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T17:26:28.808811Z digest=sha256:9cbc16a542e2a022eae2043e151bc4599d19c56dee83e37e736968cbe763bfbf

Observation 56aafb93-0a05-4117-aa2b-b8717cd5045b · outbound

This paper cites HealthGPT: A Medical Large Vision-Language Model for Unifying Comprehension and Generation via Heterogeneous Knowledge Adaptation.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images HealthGPT: A Medical Large Vision-Language Model for Unifying Comprehension and Generation via Heterogeneous Knowledge Adaptation

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-08T17:26:28.812115Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:26:28.812115Z digest=sha256:ecd27b91766fdf651e72ed7f3692101edc3102359aca6179e95d5c09a427c9cd

Observation b2c74f02-0fc8-47de-861a-15489c8ad1fd · outbound

This paper cites Lawrence Zitnick.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images Lawrence Zitnick

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:26:29.690906Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T17:26:28.815732Z digest=sha256:e2f040d7d531297e8e320f72cfc2191433b65347ee3e3196ee9ce01ff2989200

Observation 031bcc8b-2902-4d29-8b1b-e62b516a752e · outbound

This paper cites Medical visual question answering: A survey.Artificial Intelligence in Medicine, 143:102611, September 2023.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images Medical visual question answering: A survey.Artificial Intelligence in Medicine, 143:102611, September 2023

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:26:29.679656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T17:26:28.819189Z digest=sha256:fb339e74697b9b5f6b1dd0da23132a6beb707287f50cc0fc70c71e6bdffdac1a

Observation 239d602e-9789-4b5e-a6a6-8a165b986033 · outbound

This paper cites Slake: A semantically-labeled knowledge-enhanced dataset for medical visual question answering.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images Slake: A semantically-labeled knowledge-enhanced dataset for medical visual question answering

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-08T17:26:28.822637Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:26:28.822637Z digest=sha256:966b7525eb431c0eb9c43e1ca01276ccb6f5ac2b57e129e89b44a8110f97c69b

Observation 55c9dc88-8717-4f83-8263-97d5b6ce4808 · outbound

This paper cites Improved baselines with visual instruction tuning.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images Improved baselines with visual instruction tuning

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-08T17:26:28.825851Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:26:28.825851Z digest=sha256:2a4f4f0d5fe3f659f2db22c460d4e2921aeeed0ef36a374d567e5fe9d7538427

Observation a10a4ee1-0311-4bfc-89cb-233eb6a60227 · outbound

This paper cites Visual instruction tuning.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images Visual instruction tuning

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-08T17:26:28.829253Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:26:28.829253Z digest=sha256:2dc513c0196cfc5b672c8274faaa97f30e8e265349c42749a3aa0831e817c28a

Observation 160477dd-c428-4f93-b800-d8fef28e03f6 · outbound

This paper cites HC-LLM: Historical-Constrained Large Language Models for Radiology Report Generation.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images HC-LLM: Historical-Constrained Large Language Models for Radiology Report Generation

Reference 30

Resolution
verified exact
local_arxiv, observed 2026-08-08T17:26:29.189728Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T17:26:28.832678Z digest=sha256:1e200d1656bd483c14a567251f9a8a0036e7b14f7fd93f747d74ab2b02cf49fd

Observation a983ef33-30cc-4b66-8310-7c0560b3da1c · outbound

This paper cites Vkd: Improving knowledge distillation using orthogonal projections.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images Vkd: Improving knowledge distillation using orthogonal projections

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:26:29.649505Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T17:26:28.836028Z digest=sha256:0670b6d4b40946f8613b221a5e93662595581dc6660e54cc120a249444942bde

Observation cb4d674a-f012-4b0e-a89a-f91dffee36f4 · outbound

This paper cites Med-flamingo: a multimodal medical few-shot learner.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images Med-flamingo: a multimodal medical few-shot learner

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-08T17:26:28.839204Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:26:28.839204Z digest=sha256:c79f8dae18c534f660529ff9ecc722721d98ad8f43ff85fa7405809a085bafc1

Observation 37f7ff8b-45b8-4db2-907e-ad3c8bfead33 · outbound

This paper cites Learning deep representations with probabilistic knowledge transfer.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images Learning deep representations with probabilistic knowledge transfer

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:26:29.631339Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T17:26:28.842573Z digest=sha256:8c009a220d864a909bb1a04c1819e969687d7fdd8e1f5c13dc884815e950c5bd

Observation d278ba54-4f14-45bc-8044-f97fde37b150 · outbound

This paper cites Roformer: Enhanced transformer with rotary position embedding.Neurocomputing, 568:127063, 2024.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images Roformer: Enhanced transformer with rotary position embedding.Neurocomputing, 568:127063, 2024

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-08T17:26:28.846187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:26:28.846187Z digest=sha256:7ac91399a80a11d769520d09894d650e31f11dbc5c61e9cebfd5bbe043ccd9be

Observation cf8f52fb-56cd-4f2a-9d0e-f3305d27f596 · outbound

This paper cites Similarity-preserving knowledge distillation.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images Similarity-preserving knowledge distillation

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:26:29.612914Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T17:26:28.849454Z digest=sha256:10c0cc6b93cfa6abc54f54172623e0073b4ba04986a80deaff11105232eaddb6

Observation 5623161c-9ea5-4de5-8e2c-877dd240504b · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-08T17:26:28.852580Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:26:28.852580Z digest=sha256:02615fe47e1d92fafedb42137caa85bca307f93c19e06eca4c063f067cface16

Observation 04d82afc-160b-4efb-a9fd-25be3bb7c2ad · outbound

This paper cites ITA: Image-text alignments for multi-modal named entity recognition.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images ITA: Image-text alignments for multi-modal named entity recognition

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:26:29.602771Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T17:26:28.856024Z digest=sha256:816c5df175d45436935977eeae9330a5ee1468c5ccbfbf2a9980b636ba568767

Observation 0c8f2209-5d3c-42fd-ab0c-1e4749873c54 · outbound

This paper cites VideoRoPE: What Makes for Good Video Rotary Position Embedding?.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images VideoRoPE: What Makes for Good Video Rotary Position Embedding?

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-08T17:26:28.859252Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:26:28.859252Z digest=sha256:aa465401c3662bf117a3c7fdbebff8f3f5b5083cde038f50a347f8d6c35b4c14

Observation 0c3aebb5-9e25-4515-a346-e664a6db041b · outbound

This paper cites Model Merging in LLMs, MLLMs, and Beyond: Methods, Theories, Applications and Opportunities.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images Model Merging in LLMs, MLLMs, and Beyond: Methods, Theories, Applications and Opportunities

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-08T17:26:28.862700Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:26:28.862700Z digest=sha256:983d8acfd37df0a2abb275d1f068b5775261bfaaafd037259c386ecbb0405881

Observation 10bb5225-3e6f-45bf-ad53-76867a54c702 · outbound

This paper cites LLM-MedQA: Enhancing Medical Question Answering through Case Studies in Large Language Models.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images LLM-MedQA: Enhancing Medical Question Answering through Case Studies in Large Language Models

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-08T17:26:28.866002Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:26:28.866002Z digest=sha256:d9420aede7854ec511554d0e3fdfa2f8a43717f9ba2dd1374715b5b5ba240426

Observation 4df41cd2-d9da-4d39-9485-c4850eebf467 · outbound

This paper cites SA-Med2D-20M Dataset: Segment Anything in 2D Medical Imaging with 20 Million masks.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images SA-Med2D-20M Dataset: Segment Anything in 2D Medical Imaging with 20 Million masks

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-08T17:26:28.869297Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:26:28.869297Z digest=sha256:1e1fefae1c305dc77966b45280d1e5adec46473fb9f54244e8bf42b4116b0ccd

Observation 607dcf93-e167-40e9-96a6-8bbbf35ff8da · outbound

This paper cites Ferret: Refer and Ground Anything Anywhere at Any Granularity.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images Ferret: Refer and Ground Anything Anywhere at Any Granularity

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-08T17:26:28.873088Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:26:28.873088Z digest=sha256:c56c9ef0b6c01f70b12ac8e4c46f1ace9363e43e46149ac8de262c08c288ea85

Observation 95457c9d-4a1a-4666-b1d1-2adb74e27b9f · outbound

This paper cites Visual-Oriented Fine-Grained Knowledge Editing for MultiModal Large Language Models.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images Visual-Oriented Fine-Grained Knowledge Editing for MultiModal Large Language Models

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-08T17:26:28.876560Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:26:28.876560Z digest=sha256:7384fc432aaa0dd464f96a11ee935c0cd60036b6a10dfa0cca3e191a6dd07212

Observation 30ff817f-1f3f-4dfc-b3db-e497b9323d08 · outbound

This paper cites Ferret-v2: An Improved Baseline for Referring and Grounding with Large Language Models.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images Ferret-v2: An Improved Baseline for Referring and Grounding with Large Language Models

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-08T17:26:28.879917Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:26:28.879917Z digest=sha256:fb38b7552d7607129f2270e1b785fb0a0f827e93dbf84de62409fe15f673626f

Observation 88d15807-218e-4def-bd5c-c2f9e0a3a0b3 · outbound

This paper cites Davison, Hui Ren, Jing Huang, Chen Chen, Yuyin Zhou, Sunyang Fu, Wei Liu, Tianming Liu, Xiang Li, Yong Chen, Lifang He, James Zou, Quanzheng Li, Hongfang Liu, and Lichao Sun.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images Davison, Hui Ren, Jing Huang, Chen Chen, Yuyin Zhou, Sunyang Fu, Wei Liu, Tianming Liu, Xiang Li, Yong Chen, Lifang He, James Zou, Quanzheng Li, Hongfang Liu, and Lichao Sun

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:26:29.591556Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T17:26:28.883459Z digest=sha256:adac4ba91f71ba75091920e59c8b0763d82bd0e81b8d1f92de3e42ebe47ee1ec

Observation 5d9ad6d2-4394-4689-aa94-8848b23c618c · outbound

This paper cites Negative-aware attention framework for image-text matching.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images Negative-aware attention framework for image-text matching

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:26:29.579997Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T17:26:28.886762Z digest=sha256:20b33c53753135dbc8c5cf726d43b8388b6e91cd6c442dd4973d891b8f576c38

Observation 2484f6b7-4569-4f55-bfed-cb73f174fa36 · outbound

This paper cites BiomedCLIP: a multimodal biomedical foundation model pretrained from fifteen million scientific image-text pairs.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images BiomedCLIP: a multimodal biomedical foundation model pretrained from fifteen million scientific image-text pairs

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-08T17:26:28.889887Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:26:28.889887Z digest=sha256:2040eb0ae8f77f0b0e9b91406f42eb4980d44dd17131e297991790e7664abcb2

Observation 46d5a016-04ec-457d-bfbc-899176af1080 · outbound

This paper cites Instruction tuning for large language models: A survey.arXiv preprint arXiv:2308.10792, 2023.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images Instruction tuning for large language models: A survey.arXiv preprint arXiv:2308.10792, 2023

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-08T17:26:28.893406Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:26:28.893406Z digest=sha256:21bca7ef46e75c2f1f36d5f577d13a26e52d41aa2a298d05ce4459a7e38fc93e

Observation 56cda649-f7b4-4bce-8130-274aebc7143d · outbound

This paper cites Consecutive knowledge meta-adaptation learning for unsupervised medical diagnosis.Knowledge-Based Systems, 291:111573, 2024.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images Consecutive knowledge meta-adaptation learning for unsupervised medical diagnosis.Knowledge-Based Systems, 291:111573, 2024

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:26:29.568438Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T17:26:28.896634Z digest=sha256:e527170544da8283462720a4ee0b4365de28b7e17429587d63311a932ccc6a58

Observation 84a7b064-aa80-47ab-bb7a-e8eba4e1d54d · outbound

This paper cites Scene parsing through ade20k dataset.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images Scene parsing through ade20k dataset

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-08T17:26:28.899998Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:26:28.899998Z digest=sha256:e253effd4c0ad8ffd65fd6fe51896fa5810ce5b206ec52e74115f60381345bfe

Observation 48afdfad-3585-4f92-8ba8-faa735393b9a · outbound

This paper cites Semantic understanding of scenes through the ade20k dataset.International Journal of Computer Vision, 127:302–321, 2019.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images Semantic understanding of scenes through the ade20k dataset.International Journal of Computer Vision, 127:302–321, 2019

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:26:29.552018Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T17:26:28.903390Z digest=sha256:28cb5e815b74d8d26ec404a122f62eb22bdc51671c1d0600082eee668ccec09d

Observation 941e323c-6c5a-4ede-b595-6a6d7a3907c5 · outbound

This paper cites an unresolved cited work.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images Unresolved cited work

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-08T17:26:28.780612Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:26:28.780612Z digest=sha256:90d55569edd011bf44269d18b79f7ce83073ef4890ef707d0738cbfea209cf82

Pith citing papers

Observation 2ac4f0d3-c1ee-45d7-8d69-6eadbac8efd2 · inbound

FADA: Accessible fetal ultrasound interpretation and annotation with a selectively distilled unified vision-language model cites this paper.

FADA: Accessible fetal ultrasound interpretation and annotation with a selectively distilled unified vision-language model ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-07-03T05:47:40.980557Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-27T13:08:02.505489Z digest=sha256:8028e2a6f21485216123478639e40b2e8efa51c658b98655e33ab9f23ee8445b