Pith. sign in

Paper Citation Record · LEDGER

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images

As of 8 August 2026, this Paper Citation Record lists 52 of 52 outbound references and 1 inbound Pith citation observation for arXiv:2502.05928.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.05928 v4

Coverage vector

measured 52 of 52 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-08T17:26:28.903390Z

measured 53 of 53 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-27T13:08:02.505489Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T05:47:40.979179Z

Reference resolution

52 of 52 outbound references displayed

  • verified exact2
  • verified fuzzy20
  • unresolved30
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation e94c8803-b6cf-4fd2-9228-7b62e17d2d6e · outbound

This paper cites Hasan, Vivek Datla, Joey Liu, Dina Demner-Fushman, and Henning Müller.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images Hasan, Vivek Datla, Joey Liu, Dina Demner-Fushman, and Henning Müller

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:26:29.828526Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T17:26:28.724936Z digest=sha256:505736d1699ed07d9f46bf61b0a029e117e278c66136b05de3538ea942e81ea6

Observation fa800f7b-5c6f-4c94-b913-93a123ec5cae · outbound

This paper cites Spice: Semantic propositional image caption evaluation.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images Spice: Semantic propositional image caption evaluation

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:26:29.818384Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T17:26:28.728940Z digest=sha256:567ecda9caf1a76d0816a1cf92c727e18880ac5112f7b06c5ff1f13cf58d274d

Observation 540e8e91-458f-4f1e-91e6-d3103796147a · outbound

This paper cites Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-08T17:26:28.732983Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:26:28.732983Z digest=sha256:821f3752e920981486fea1ee654458976438578a08b6a51908b8e1bb5d890bb2

Observation 1e2dd7d5-2972-44cf-9b51-239aa2f9f2ac · outbound

This paper cites The revolution of multimodal large language models: A survey.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images The revolution of multimodal large language models: A survey

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:26:29.807623Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T17:26:28.737058Z digest=sha256:ca47d200c1a8bd1b93033e3e8ac0f1af9100e44fea5839075c7f5eb1d2790ed9

Observation c5800137-91d9-42a9-98e2-38c9fa5183e6 · outbound

This paper cites R-LLaVA: Improving Med-VQA Understanding through Visual Region of Interest.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images R-LLaVA: Improving Med-VQA Understanding through Visual Region of Interest

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-08T17:26:28.740621Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:26:28.740621Z digest=sha256:c7aaf8df3a52d5ba3621a0a8982a484858d38fe2d863dd15d9281a0958a268b6

Observation e3cbf274-c92d-4935-bcc2-06816c407c19 · outbound

This paper cites Mixed Pseudo Labels for Semi-Supervised Object Detection.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images Mixed Pseudo Labels for Semi-Supervised Object Detection

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-08T17:26:28.744479Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:26:28.744479Z digest=sha256:7102ca10a89fcaa00741e033f2ade8d436d5a1daee08f2c1dbae8f261a8d3b4f

Observation 5ae6695e-eac8-43b9-9496-ea7ad6022937 · outbound

This paper cites SAM-Med2D.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images SAM-Med2D

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-08T17:26:28.748689Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:26:28.748689Z digest=sha256:2c2667025417d85c4cf0e2b01c2a606bb5e57db89bc2d600e16afb3412a123e9

Observation 1ca61755-d128-45b1-8520-0ecce10fa243 · outbound

This paper cites an unresolved cited work.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-08T17:26:29.795526Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T17:26:28.752414Z digest=sha256:50bfaf8f4aba4439b611e73c47de1554ec9635b785fa215a5c7db0ffb4b57e4b

Observation 0acab50f-b9fe-49bc-84cd-98e8c32a5ce6 · outbound

This paper cites Enhancing medical VQA with multimodal determination rationales.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images Enhancing medical VQA with multimodal determination rationales

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:26:29.784785Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T17:26:28.755932Z digest=sha256:e94b7e0f995bd0bb579a28ae03b83e76628392e541d832dc3f11dfaf73dbeea4

Observation 34f346a7-b109-45c5-9f8f-861d1200639a · outbound

This paper cites Hasan, Yuan Ling, Oladimeji Farri, Joey Liu, Henning Müller, and Matthew P.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images Hasan, Yuan Ling, Oladimeji Farri, Joey Liu, Henning Müller, and Matthew P

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:26:29.774228Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T17:26:28.759219Z digest=sha256:b1a9ac02d322e79d56d5954c880d0d6b8d894bfc3091c9f8c5dc5548b2bbc7ee

Observation d4c6f41c-3374-47b6-8195-7c141732ab98 · outbound

This paper cites Parameter-Efficient Fine-Tuning Medical Multimodal Large Language Models for Medical Visual Grounding.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images Parameter-Efficient Fine-Tuning Medical Multimodal Large Language Models for Medical Visual Grounding

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-08T17:26:28.762817Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:26:28.762817Z digest=sha256:6f73c5a2fc09234d526ffc93962da2267273b609afc76f2c4ae673df3d8bbb2d

Observation 69d37335-c96d-4ef6-98c4-bffdab392d90 · outbound

This paper cites PathVQA: 30000+ Questions for Medical Visual Question Answering.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images PathVQA: 30000+ Questions for Medical Visual Question Answering

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-08T17:26:28.766461Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:26:28.766461Z digest=sha256:c8641558bd4c336298644ef778a638dd772c1570addb6b568ee57f502f988616

Observation db98cd9f-dd53-43a8-a682-9030fbe7127b · outbound

This paper cites Rotary position embedding for vision transformer.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images Rotary position embedding for vision transformer

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:26:29.764002Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T17:26:28.770156Z digest=sha256:0e8517eb05e6a46e11684f63bbadf60eb5e3d745e7a8d69881c37137f9a7f68c

Observation 23533ab1-26e9-4a57-8ea7-518607cc5574 · outbound

This paper cites Distilling the Knowledge in a Neural Network.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images Distilling the Knowledge in a Neural Network

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-08T17:26:28.773566Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:26:28.773566Z digest=sha256:bab786f1e93b3ee3d7ae49428f9f0f38c785e830d2997102b0101b305822605a

Observation b158e2a1-d394-46b1-a099-d0a7429cd01c · outbound

This paper cites A refer-and-ground multimodal large language model for biomedicine.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images A refer-and-ground multimodal large language model for biomedicine

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:26:29.754097Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T17:26:28.777238Z digest=sha256:b151cf7620393112a845e8c89112de333766a4c16884aeb2e027556ec3a2ed01

Observation 841e7257-7abe-48ce-9a2f-b7a8475f97ab · outbound

This paper cites MedVisionLlama: Leveraging Pre-Trained Large Language Model Layers to Enhance Medical Image Segmentation.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images MedVisionLlama: Leveraging Pre-Trained Large Language Model Layers to Enhance Medical Image Segmentation

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-08-08T17:26:29.468034Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T17:26:28.784274Z digest=sha256:d2cf366873caa8baa7e3cbf1c96ddd4853d461220830fefe3edb724f62746945

Observation 32585bb0-b0cc-45e0-bbdc-a68845552b5f · outbound

This paper cites A dataset of clinically generated visual questions and answers about radiology images.Scientific data, 5(1):1–10, 2018.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images A dataset of clinically generated visual questions and answers about radiology images.Scientific data, 5(1):1–10, 2018

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-08T17:26:28.788101Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:26:28.788101Z digest=sha256:73e12dfa1da7e70aa6a9b864fb5b2f8caf17d7bb9ab775d41389aa15f5b4bfb3

Observation 7971a455-4b7a-454b-8ed3-10a139425ea5 · outbound

This paper cites Vision-Language Instruction Tuning: A Review and Analysis.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images Vision-Language Instruction Tuning: A Review and Analysis

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-08T17:26:28.791677Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:26:28.791677Z digest=sha256:7cd9670e98f8743a65a293c9f2b4486da1f5900c806cd7ec8d7dccc93ca676cd

Observation c0c6ff4c-f141-424e-814d-60fd02252925 · outbound

This paper cites Llava-med: Training a large language-and-vision assistant for biomedicine in one day.Advances in Neural Information Processing Systems, 36:28541–28564, 2023.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images Llava-med: Training a large language-and-vision assistant for biomedicine in one day.Advances in Neural Information Processing Systems, 36:28541–28564, 2023

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-08T17:26:28.795486Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:26:28.795486Z digest=sha256:bda52ff8653371c44506293990211cc31bd965d4c7f2f106642024c0d249f1ab

Observation 13f61901-5b5e-486c-92ba-5df82e69fbd1 · outbound

This paper cites Towards visual-prompt temporal answer grounding in instructional video.IEEE Transactions on Pattern Analysis and Machine Intelligence, 2024.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images Towards visual-prompt temporal answer grounding in instructional video.IEEE Transactions on Pattern Analysis and Machine Intelligence, 2024

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:26:29.724444Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T17:26:28.798995Z digest=sha256:69362d0708e00f64ec6e1a58d8f0c7dad0e6de9924d62498c38b0b949a8d8983

Observation 33438188-404f-4b70-9b29-08705daa8ddf · outbound

This paper cites Pseudo labels for unsupervised domain adaptation: A review.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images Pseudo labels for unsupervised domain adaptation: A review

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:26:29.713921Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T17:26:28.802171Z digest=sha256:0941b0129059d500b66963bd91673a0ab18720e19046471a2d272c35b3dd0604

Observation 695be0d4-5b54-4b5f-9a06-ecfb3bfe3028 · outbound

This paper cites A comprehensive survey and guide to multimodal large language models in vision-language tasks.arXiv preprint arXiv:2411.06284, 2024.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images A comprehensive survey and guide to multimodal large language models in vision-language tasks.arXiv preprint arXiv:2411.06284, 2024

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-08T17:26:28.805527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:26:28.805527Z digest=sha256:c8f309a15c4cb27d3bf74301a9164337adc03f03ab64c010dd2626fc0848f451

Observation 48c9db19-c440-426d-ac51-59cfa09cbf05 · outbound

This paper cites Medfilip: Medical fine-grained language-image pre-training.IEEE Journal of Biomedical and Health Informatics, pages 1–11, 2025.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images Medfilip: Medical fine-grained language-image pre-training.IEEE Journal of Biomedical and Health Informatics, pages 1–11, 2025

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:26:29.702840Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T17:26:28.808811Z digest=sha256:d87d6c880b2b3f3513f88840e9098bbf36f8ddfae1e53772df0a86b40c47c0e8

Observation 56aafb93-0a05-4117-aa2b-b8717cd5045b · outbound

This paper cites HealthGPT: A Medical Large Vision-Language Model for Unifying Comprehension and Generation via Heterogeneous Knowledge Adaptation.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images HealthGPT: A Medical Large Vision-Language Model for Unifying Comprehension and Generation via Heterogeneous Knowledge Adaptation

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-08T17:26:28.812115Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:26:28.812115Z digest=sha256:ecd27b91766fdf651e72ed7f3692101edc3102359aca6179e95d5c09a427c9cd

Observation b2c74f02-0fc8-47de-861a-15489c8ad1fd · outbound

This paper cites Lawrence Zitnick.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images Lawrence Zitnick

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:26:29.690906Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T17:26:28.815732Z digest=sha256:1ff4889088ce3bad77c49ee87a64e5200920d8a9e8d2fb039b79fd6756b98134

Observation 031bcc8b-2902-4d29-8b1b-e62b516a752e · outbound

This paper cites Medical visual question answering: A survey.Artificial Intelligence in Medicine, 143:102611, September 2023.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images Medical visual question answering: A survey.Artificial Intelligence in Medicine, 143:102611, September 2023

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:26:29.679656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T17:26:28.819189Z digest=sha256:e987660098fc44fcd24eae33c05e8e2e5388003d4102b1e4fab942334a082fdc

Observation 239d602e-9789-4b5e-a6a6-8a165b986033 · outbound

This paper cites Slake: A semantically-labeled knowledge-enhanced dataset for medical visual question answering.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images Slake: A semantically-labeled knowledge-enhanced dataset for medical visual question answering

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-08T17:26:28.822637Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:26:28.822637Z digest=sha256:966b7525eb431c0eb9c43e1ca01276ccb6f5ac2b57e129e89b44a8110f97c69b

Observation 55c9dc88-8717-4f83-8263-97d5b6ce4808 · outbound

This paper cites Improved baselines with visual instruction tuning.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images Improved baselines with visual instruction tuning

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-08T17:26:28.825851Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:26:28.825851Z digest=sha256:2a4f4f0d5fe3f659f2db22c460d4e2921aeeed0ef36a374d567e5fe9d7538427

Observation a10a4ee1-0311-4bfc-89cb-233eb6a60227 · outbound

This paper cites Visual instruction tuning.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images Visual instruction tuning

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-08T17:26:28.829253Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:26:28.829253Z digest=sha256:2dc513c0196cfc5b672c8274faaa97f30e8e265349c42749a3aa0831e817c28a

Observation 160477dd-c428-4f93-b800-d8fef28e03f6 · outbound

This paper cites HC-LLM: Historical-Constrained Large Language Models for Radiology Report Generation.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images HC-LLM: Historical-Constrained Large Language Models for Radiology Report Generation

Reference 30

Resolution
verified exact
local_arxiv, observed 2026-08-08T17:26:29.189728Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T17:26:28.832678Z digest=sha256:f9e35179329832349c33181a5ca3c21d1e13c06ab3a7a13f2ef46a1b3e98cf22

Observation a983ef33-30cc-4b66-8310-7c0560b3da1c · outbound

This paper cites Vkd: Improving knowledge distillation using orthogonal projections.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images Vkd: Improving knowledge distillation using orthogonal projections

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:26:29.649505Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T17:26:28.836028Z digest=sha256:af211adcbdb32c6aa8cb0dfce4b3698ba3ae9103de221c0a8f74d9d2f153caa3

Observation cb4d674a-f012-4b0e-a89a-f91dffee36f4 · outbound

This paper cites Med-flamingo: a multimodal medical few-shot learner.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images Med-flamingo: a multimodal medical few-shot learner

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-08T17:26:28.839204Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:26:28.839204Z digest=sha256:c79f8dae18c534f660529ff9ecc722721d98ad8f43ff85fa7405809a085bafc1

Observation 37f7ff8b-45b8-4db2-907e-ad3c8bfead33 · outbound

This paper cites Learning deep representations with probabilistic knowledge transfer.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images Learning deep representations with probabilistic knowledge transfer

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:26:29.631339Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T17:26:28.842573Z digest=sha256:0d93366f45f9778c217fa00a683542b5ed3d3ec4f23ad29872a5ec28aaedf5d9

Observation d278ba54-4f14-45bc-8044-f97fde37b150 · outbound

This paper cites Roformer: Enhanced transformer with rotary position embedding.Neurocomputing, 568:127063, 2024.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images Roformer: Enhanced transformer with rotary position embedding.Neurocomputing, 568:127063, 2024

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-08T17:26:28.846187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:26:28.846187Z digest=sha256:7ac91399a80a11d769520d09894d650e31f11dbc5c61e9cebfd5bbe043ccd9be

Observation cf8f52fb-56cd-4f2a-9d0e-f3305d27f596 · outbound

This paper cites Similarity-preserving knowledge distillation.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images Similarity-preserving knowledge distillation

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:26:29.612914Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T17:26:28.849454Z digest=sha256:49120049009e63698d0e62b145bfe01b7fdb0b0266ae7393b80cc2e423fe738c

Observation 5623161c-9ea5-4de5-8e2c-877dd240504b · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-08T17:26:28.852580Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:26:28.852580Z digest=sha256:02615fe47e1d92fafedb42137caa85bca307f93c19e06eca4c063f067cface16

Observation 04d82afc-160b-4efb-a9fd-25be3bb7c2ad · outbound

This paper cites ITA: Image-text alignments for multi-modal named entity recognition.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images ITA: Image-text alignments for multi-modal named entity recognition

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:26:29.602771Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T17:26:28.856024Z digest=sha256:1b2c0551b235da377dbdbba166a677e96d3c68e5333e434e8622dc3f26fe3928

Observation 0c8f2209-5d3c-42fd-ab0c-1e4749873c54 · outbound

This paper cites VideoRoPE: What Makes for Good Video Rotary Position Embedding?.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images VideoRoPE: What Makes for Good Video Rotary Position Embedding?

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-08T17:26:28.859252Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:26:28.859252Z digest=sha256:aa465401c3662bf117a3c7fdbebff8f3f5b5083cde038f50a347f8d6c35b4c14

Observation 0c3aebb5-9e25-4515-a346-e664a6db041b · outbound

This paper cites Model Merging in LLMs, MLLMs, and Beyond: Methods, Theories, Applications and Opportunities.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images Model Merging in LLMs, MLLMs, and Beyond: Methods, Theories, Applications and Opportunities

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-08T17:26:28.862700Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:26:28.862700Z digest=sha256:983d8acfd37df0a2abb275d1f068b5775261bfaaafd037259c386ecbb0405881

Observation 10bb5225-3e6f-45bf-ad53-76867a54c702 · outbound

This paper cites LLM-MedQA: Enhancing Medical Question Answering through Case Studies in Large Language Models.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images LLM-MedQA: Enhancing Medical Question Answering through Case Studies in Large Language Models

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-08T17:26:28.866002Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:26:28.866002Z digest=sha256:d9420aede7854ec511554d0e3fdfa2f8a43717f9ba2dd1374715b5b5ba240426

Observation 4df41cd2-d9da-4d39-9485-c4850eebf467 · outbound

This paper cites SA-Med2D-20M Dataset: Segment Anything in 2D Medical Imaging with 20 Million masks.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images SA-Med2D-20M Dataset: Segment Anything in 2D Medical Imaging with 20 Million masks

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-08T17:26:28.869297Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:26:28.869297Z digest=sha256:1e1fefae1c305dc77966b45280d1e5adec46473fb9f54244e8bf42b4116b0ccd

Observation 607dcf93-e167-40e9-96a6-8bbbf35ff8da · outbound

This paper cites Ferret: Refer and Ground Anything Anywhere at Any Granularity.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images Ferret: Refer and Ground Anything Anywhere at Any Granularity

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-08T17:26:28.873088Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:26:28.873088Z digest=sha256:c56c9ef0b6c01f70b12ac8e4c46f1ace9363e43e46149ac8de262c08c288ea85

Observation 95457c9d-4a1a-4666-b1d1-2adb74e27b9f · outbound

This paper cites Visual-Oriented Fine-Grained Knowledge Editing for MultiModal Large Language Models.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images Visual-Oriented Fine-Grained Knowledge Editing for MultiModal Large Language Models

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-08T17:26:28.876560Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:26:28.876560Z digest=sha256:7384fc432aaa0dd464f96a11ee935c0cd60036b6a10dfa0cca3e191a6dd07212

Observation 30ff817f-1f3f-4dfc-b3db-e497b9323d08 · outbound

This paper cites Ferret-v2: An Improved Baseline for Referring and Grounding with Large Language Models.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images Ferret-v2: An Improved Baseline for Referring and Grounding with Large Language Models

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-08T17:26:28.879917Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:26:28.879917Z digest=sha256:fb38b7552d7607129f2270e1b785fb0a0f827e93dbf84de62409fe15f673626f

Observation 88d15807-218e-4def-bd5c-c2f9e0a3a0b3 · outbound

This paper cites Davison, Hui Ren, Jing Huang, Chen Chen, Yuyin Zhou, Sunyang Fu, Wei Liu, Tianming Liu, Xiang Li, Yong Chen, Lifang He, James Zou, Quanzheng Li, Hongfang Liu, and Lichao Sun.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images Davison, Hui Ren, Jing Huang, Chen Chen, Yuyin Zhou, Sunyang Fu, Wei Liu, Tianming Liu, Xiang Li, Yong Chen, Lifang He, James Zou, Quanzheng Li, Hongfang Liu, and Lichao Sun

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:26:29.591556Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T17:26:28.883459Z digest=sha256:ade346a34725dd1ac57d30516d264944006ad0ae18475047996b742b6ac9daee

Observation 5d9ad6d2-4394-4689-aa94-8848b23c618c · outbound

This paper cites Negative-aware attention framework for image-text matching.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images Negative-aware attention framework for image-text matching

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:26:29.579997Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T17:26:28.886762Z digest=sha256:5355d51924a93931ee6eb39be58132b2d0fe7a92e6d2bc69de3b0c18eef583a4

Observation 2484f6b7-4569-4f55-bfed-cb73f174fa36 · outbound

This paper cites BiomedCLIP: a multimodal biomedical foundation model pretrained from fifteen million scientific image-text pairs.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images BiomedCLIP: a multimodal biomedical foundation model pretrained from fifteen million scientific image-text pairs

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-08T17:26:28.889887Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:26:28.889887Z digest=sha256:2040eb0ae8f77f0b0e9b91406f42eb4980d44dd17131e297991790e7664abcb2

Observation 46d5a016-04ec-457d-bfbc-899176af1080 · outbound

This paper cites Instruction tuning for large language models: A survey.arXiv preprint arXiv:2308.10792, 2023.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images Instruction tuning for large language models: A survey.arXiv preprint arXiv:2308.10792, 2023

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-08T17:26:28.893406Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:26:28.893406Z digest=sha256:21bca7ef46e75c2f1f36d5f577d13a26e52d41aa2a298d05ce4459a7e38fc93e

Observation 56cda649-f7b4-4bce-8130-274aebc7143d · outbound

This paper cites Consecutive knowledge meta-adaptation learning for unsupervised medical diagnosis.Knowledge-Based Systems, 291:111573, 2024.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images Consecutive knowledge meta-adaptation learning for unsupervised medical diagnosis.Knowledge-Based Systems, 291:111573, 2024

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:26:29.568438Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T17:26:28.896634Z digest=sha256:7220f28bdf132f331390e1131ca48220ecc870db87d6287a5e828dd1373c6475

Observation 84a7b064-aa80-47ab-bb7a-e8eba4e1d54d · outbound

This paper cites Scene parsing through ade20k dataset.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images Scene parsing through ade20k dataset

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-08T17:26:28.899998Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:26:28.899998Z digest=sha256:e253effd4c0ad8ffd65fd6fe51896fa5810ce5b206ec52e74115f60381345bfe

Observation 48afdfad-3585-4f92-8ba8-faa735393b9a · outbound

This paper cites Semantic understanding of scenes through the ade20k dataset.International Journal of Computer Vision, 127:302–321, 2019.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images Semantic understanding of scenes through the ade20k dataset.International Journal of Computer Vision, 127:302–321, 2019

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:26:29.552018Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T17:26:28.903390Z digest=sha256:17ce733805751cc26b250606670159abe86183d9c0b7746b63275750c5573cb5

Observation 941e323c-6c5a-4ede-b595-6a6d7a3907c5 · outbound

This paper cites an unresolved cited work.

ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images Unresolved cited work

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-08T17:26:28.780612Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:26:28.780612Z digest=sha256:90d55569edd011bf44269d18b79f7ce83073ef4890ef707d0738cbfea209cf82

Pith citing papers

Observation 2ac4f0d3-c1ee-45d7-8d69-6eadbac8efd2 · inbound

FADA: Accessible fetal ultrasound interpretation and annotation with a selectively distilled unified vision-language model cites this paper.

FADA: Accessible fetal ultrasound interpretation and annotation with a selectively distilled unified vision-language model ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-07-03T05:47:40.980557Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-27T13:08:02.505489Z digest=sha256:7a8da344bb478ccde758542d2e710c1c138dba6183033a9cbba1185e3c37c226