Pith. sign in

Paper Citation Record · LEDGER

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting

As of 15 August 2026, this Paper Citation Record lists 60 of 60 outbound references and 2 inbound Pith citation observations for arXiv:2509.03800.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2509.03800 v1

Coverage vector

measured 60 of 60 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T10:46:09.996574Z

measured 62 of 62 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-03T19:34:04.544567Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-28T19:22:34.657932Z

Reference resolution

60 of 60 outbound references displayed

  • verified exact1
  • verified fuzzy39
  • unresolved20
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a6fa80d1-f5d6-4f4a-8c11-f7590af040bf · outbound

This paper cites Merlin: A vision language foundation model for 3d computed tomography.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Merlin: A vision language foundation model for 3d computed tomography

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-05T10:46:04.243549Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:46:04.243549Z digest=sha256:ff050499cb93cf3afd4ad97e7389c588079b80528c647699ba8f7feb3751adca

Observation b33a7ab9-9f9f-45d3-8445-97b33ff2b16b · outbound

This paper cites A vision–language foundation model for the generation of realistic chest x-ray images.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting A vision–language foundation model for the generation of realistic chest x-ray images

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.958572Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-05T10:46:04.330865Z digest=sha256:c64870cf4b934d5ece528663861a3e7f49b15f0275a48abf920c20cea2f95306

Observation cb37fff5-0668-4885-9707-e0ec24f3015e · outbound

This paper cites Making the most of text semantics to improve biomedical vision–language processing.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Making the most of text semantics to improve biomedical vision–language processing

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-05T10:46:04.426003Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:46:04.426003Z digest=sha256:077f1d949c61ff702bc0c0feaf50783f43701517fab11df421ed73f162dc747e

Observation d404663f-9670-4008-8297-2b334af143b6 · outbound

This paper cites Understanding and confronting our mis- takes: the epidemiology of error in radiology and strategies for error reduction.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Understanding and confronting our mis- takes: the epidemiology of error in radiology and strategies for error reduction

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.932829Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-05T10:46:04.516487Z digest=sha256:3e94159bf446bf2d458cd632bf36be398b45abf611f33602d35a6119e6dafaaa

Observation c143d303-e11e-4e10-bd16-89cf43f4e210 · outbound

This paper cites Joint modeling of chest radiographs and radiology reports for pulmonary edema assessment.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Joint modeling of chest radiographs and radiology reports for pulmonary edema assessment

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.917747Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-05T10:46:04.585735Z digest=sha256:43e3f195c10b97b7a2d77866810c9ba0f6614bf6270b94f43777a92895ac8c94

Observation 41ed5110-7327-4e3d-bafa-1a1622ef2ee5 · outbound

This paper cites Contrastive Localized Language-Image Pre-Training.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Contrastive Localized Language-Image Pre-Training

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-05T10:46:04.661105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:46:04.661105Z digest=sha256:db11d013d4669a71be9d869a9eb5990b6e0ba6b3ab3d67bd740ef7025f141b54

Observation 3f111173-6cf8-4d07-8f39-b1d2c4ddf619 · outbound

This paper cites A review of medical image data augmentation techniques for deep learning applications.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting A review of medical image data augmentation techniques for deep learning applications

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.900852Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-05T10:46:04.776379Z digest=sha256:ab2a966a208e39d8ced842af0881202dffabd9f3df3ac0498e048185d0c0ef65

Observation f555adc7-97f4-414c-8a7b-5d9ba7d8b2d8 · outbound

This paper cites Machine-learning-based multiple abnormality prediction with large-scale chest computed tomography volumes.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Machine-learning-based multiple abnormality prediction with large-scale chest computed tomography volumes

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.884634Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-05T10:46:04.842478Z digest=sha256:e4a01f361ba32cc7264e69411cd3bd6779f4448edd5c545d607fd263a017d9d4

Observation 85df312c-28df-4027-8a32-0cf6486e38a7 · outbound

This paper cites Multi-View and Multi-Scale Alignment for Contrastive Language-Image Pre-training in Mammography.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Multi-View and Multi-Scale Alignment for Contrastive Language-Image Pre-training in Mammography

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-08-05T10:46:10.536952Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-05T10:46:04.962958Z digest=sha256:24b3eb2a827744ef204b45d0586cc84d78c37ae462c9bc254c9a4a9a5dd8776e

Observation 4feab371-5592-4901-b37e-221704c40765 · outbound

This paper cites The Llama 3 Herd of Models.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting The Llama 3 Herd of Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-05T10:46:05.092590Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:46:05.092590Z digest=sha256:7e2c432f99e4c5375a14b594f59f9fbcde1f3cd183832a434b65a4f01b1422f0

Observation 13ac3f6d-a019-4f0b-825c-d36d7414fa43 · outbound

This paper cites Devel- oping generalist foundation models from a multimodal dataset for 3d computed tomography.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Devel- oping generalist foundation models from a multimodal dataset for 3d computed tomography

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-05T10:46:05.191575Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:46:05.191575Z digest=sha256:b2d3c8a629c3b05d3f7507a82840cff61b2bfff794a5a7985b9fba3e776111e9

Observation ac57e9c6-b566-4905-a756-dd5818e81d92 · outbound

This paper cites LoRA: Low-Rank Adaptation of Large Language Models.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting LoRA: Low-Rank Adaptation of Large Language Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-05T10:46:05.296884Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:46:05.296884Z digest=sha256:310eea6d4772d42c2cac5d6435868c7a6054dca4fda54ad9d17e6d06c02dad01

Observation aa273c70-1ad5-4176-a12e-25ee2a518a90 · outbound

This paper cites Gloria: A multimodal global-local representation learning framework for label-efficient medical image recognition.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Gloria: A multimodal global-local representation learning framework for label-efficient medical image recognition

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-05T10:46:05.360284Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:46:05.360284Z digest=sha256:3a08541200a5449fa7e60b20c68eb9c4c999706f614b45e140a7bef6c877b71b

Observation c7bf766a-ec65-41da-a482-156f75f1b8da · outbound

This paper cites Enhancing representation in medical vision-language foun- dation models via multi-scale information extraction techniques.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Enhancing representation in medical vision-language foun- dation models via multi-scale information extraction techniques

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.859199Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-05T10:46:05.458401Z digest=sha256:88d3c193379cc8b86fb7e2c66a9371e8aaa6f4496ab93bd6edb23c0a7f7dd1f6

Observation da68f205-0616-4a33-8ce2-23eb68858145 · outbound

This paper cites STU-Net: Scalable and Transferable Medical Image Segmentation Models Empowered by Large-Scale Supervised Pre-training.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting STU-Net: Scalable and Transferable Medical Image Segmentation Models Empowered by Large-Scale Supervised Pre-training

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-05T10:46:05.562057Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:46:05.562057Z digest=sha256:982851d38b8a9373a98a9890d9c813608bce9ff8a3cb5deeaa3034d30b84b1d0

Observation 4dbc9100-63b5-4b79-b258-fb0395cb7c62 · outbound

This paper cites nnu-net: a self-configuring method for deep learning-based biomedical image segmentation.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting nnu-net: a self-configuring method for deep learning-based biomedical image segmentation

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-05T10:46:05.633725Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:46:05.633725Z digest=sha256:504765b9e2b8f12684423afc184ec6243c154b8c504421c485b93313000325e0

Observation 0d20a01c-58e5-4de6-8e64-28d7416906b8 · outbound

This paper cites Fool me twice: delayed diagnoses in radiology with emphasis on perpetuated errors.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Fool me twice: delayed diagnoses in radiology with emphasis on perpetuated errors

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.834498Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-05T10:46:05.740237Z digest=sha256:e6b26de3ad585f00f538a05357b20093cc77fbf4ff5df2718ab2ea099c664235

Observation 0ef9eb6e-3c40-4f58-8234-d3ac1161d8f4 · outbound

This paper cites Generating synthetic data for medical imaging.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Generating synthetic data for medical imaging

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.819212Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-05T10:46:05.837515Z digest=sha256:7dcdfe36eaf770ef4604d8663f6faf080392c482609a9c8d24ce4eb7b87d7f2a

Observation 35d0a13e-01ef-4672-bf63-cc177925cee6 · outbound

This paper cites Cxr-llava: a multimodal large language model for interpreting chest x-ray images.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Cxr-llava: a multimodal large language model for interpreting chest x-ray images

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.803136Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-05T10:46:05.952126Z digest=sha256:f94a70fa63610ee2aebc25d594ed06fff44562eab0db0e8230e3d14c4989d8f2

Observation 71cb91cd-edae-46b0-9c62-195ae4964bf7 · outbound

This paper cites Llava-med: Training a large language-and-vision assistant for biomedicine in one day.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Llava-med: Training a large language-and-vision assistant for biomedicine in one day

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-05T10:46:06.026153Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:46:06.026153Z digest=sha256:fb5a81a475c24fa7a4d5a83f043687ea25b18edbe3d8bddf0ec5495500625aea

Observation 6b547427-1974-4c16-a177-2ff59895bdc3 · outbound

This paper cites Artificial general intelligence for medical imaging analysis.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Artificial general intelligence for medical imaging analysis

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.775179Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-05T10:46:06.144188Z digest=sha256:4c8c2617e20f166538faacc1b5b7555b1294121b49300a819ec832f592e89453

Observation 9920dd3d-25af-4526-9efc-147946bfd878 · outbound

This paper cites Ct-glip: 3d grounded language-image pretraining with ct scans and radiology reports for full-body scenarios.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Ct-glip: 3d grounded language-image pretraining with ct scans and radiology reports for full-body scenarios

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-05T10:46:06.214147Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:46:06.214147Z digest=sha256:e306d8737883e4e593fa13fc14e94c203e6c03c1e90db0855f56f205db549494

Observation 045e8575-14d9-458f-90dc-c8791db2b0af · outbound

This paper cites Pmc-clip: Contrastive language-image pre-training using biomedical documents.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Pmc-clip: Contrastive language-image pre-training using biomedical documents

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-05T10:46:06.347751Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:46:06.347751Z digest=sha256:40a4bb5f74af65d56ea5520c467e7d77d3cd7a891c72dbd92b30901d619d1c0c

Observation 56045d27-ec37-4ba5-a6bd-9f2a0988974a · outbound

This paper cites Can Medical Vision-Language Pre-training Succeed with Purely Synthetic Data?.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Can Medical Vision-Language Pre-training Succeed with Purely Synthetic Data?

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-05T10:46:06.490159Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:46:06.490159Z digest=sha256:f153f9b3d4f87489fcac70f03c035ae772b12766f4b3ccf451c580d272e760bd

Observation 7cfb17f5-68a2-4784-ae44-4289145e1d0d · outbound

This paper cites Improved baselines with visual instruction tuning.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Improved baselines with visual instruction tuning

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-05T10:46:06.654963Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:46:06.654963Z digest=sha256:3301484545a990bd7582c51aa8c32db78dc29b7cc07e6bfb0a9ab073ce2ff8f6

Observation 08c8a7a7-1b04-4673-b65c-8d991d660f66 · outbound

This paper cites Representation Learning with Contrastive Predictive Coding.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Representation Learning with Contrastive Predictive Coding

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-05T10:46:06.818244Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:46:06.818244Z digest=sha256:b2bd6ecc43232bab4048c97c0eb5057fd11883cab8953cfe73ed7ceea517efef

Observation 051d5200-2466-433f-948e-fa46ba9c4c30 · outbound

This paper cites Unsupervised medical image translation with adversarial diffusion models.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Unsupervised medical image translation with adversarial diffusion models

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.736423Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-05T10:46:06.953711Z digest=sha256:038db9d97de233189153432f68754833af45190bbd72318286985bc3e6e4433a

Observation 5a6d5801-2d1a-43c9-bdba-d88c5914c4df · outbound

This paper cites On variational bounds of mutual information.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting On variational bounds of mutual information

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.720929Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-05T10:46:07.058611Z digest=sha256:ea43e0e4f49df9848222ec03a4a66ec6def4a5798ede000ce7767e53f21e8fac

Observation 04d8ab2e-3c09-4969-8f54-f0c86ead5096 · outbound

This paper cites Learning transferable visual models from natural language supervision.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Learning transferable visual models from natural language supervision

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-05T10:46:07.172235Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:46:07.172235Z digest=sha256:bda66740638f8aa12a4338bdcd152cae53065da2cb316c2a1c98f558b00e59fe

Observation 904762dd-f001-4ef0-96f7-33945a9c2c06 · outbound

This paper cites Study of thoracic ct in covid-19: the stoic project.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Study of thoracic ct in covid-19: the stoic project

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.695888Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-05T10:46:07.275215Z digest=sha256:b19ccec3b46894a3c60607692ea3d1ddfd256d477f75f74175e2f79fde38367d

Observation 7526b28c-0615-4fd0-8f25-a789a2a3a0d4 · outbound

This paper cites Deep learning in medical image analysis.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Deep learning in medical image analysis

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.680527Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-05T10:46:07.401048Z digest=sha256:f29585de683e1b2dc63325ea82780c435b623b93d96dda55499994c5e1beee66

Observation 8cea4486-36cd-46ca-a3e0-800c3835fb65 · outbound

This paper cites Large-scale and fine-grained vision- language pre-training for enhanced ct image understanding.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Large-scale and fine-grained vision- language pre-training for enhanced ct image understanding

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.666394Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-05T10:46:07.524178Z digest=sha256:294774813567f46650cbf111eacde7cf6b3ed13118ac52c929d2635265bff6b4

Observation d6fa8c35-2d72-432f-8f29-319783f54832 · outbound

This paper cites Bioclip: A vision foundation model for the tree of life.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Bioclip: A vision foundation model for the tree of life

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.651773Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-05T10:46:07.697454Z digest=sha256:b240efb16532a138509253e48ee3e11df3769b5b936530aa164cfa169a7fb720

Observation 196f8bc6-9cd3-4d50-ba35-9d041cf1c193 · outbound

This paper cites XrayGPT: Chest Radiographs Summarization using Medical Vision-Language Models.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting XrayGPT: Chest Radiographs Summarization using Medical Vision-Language Models

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-05T10:46:07.826433Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:46:07.826433Z digest=sha256:cbacc7c398baf0ff6a01836f44436e5cdbcefa017beac9e438c94d476ef0c028

Observation fe0aa725-db38-4b9a-9523-c2a648ecd696 · outbound

This paper cites Communication errors in radiology–pitfalls and how to avoid them.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Communication errors in radiology–pitfalls and how to avoid them

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.637015Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-05T10:46:07.970509Z digest=sha256:aea9d9eb65e7c91c06d9a9e6a37814f5cd2a383a84f1834c478964ad2424bf7b

Observation 75a176bc-5a08-4d29-9b3e-e14cd42ca5df · outbound

This paper cites Multi- granularity cross-modal alignment for generalized medical visual representation learning.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Multi- granularity cross-modal alignment for generalized medical visual representation learning

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.621985Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-05T10:46:08.098509Z digest=sha256:6bb600766040237c4c31579488e1dc485b06fb03df4a87c9807871ba05602f85

Observation 01ab2a9b-31b2-4afe-9136-fa8d0a2ff7f7 · outbound

This paper cites Totalseg- mentator: robust segmentation of 104 anatomic structures in ct images.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Totalseg- mentator: robust segmentation of 104 anatomic structures in ct images

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.608638Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-05T10:46:08.279430Z digest=sha256:2f3c555f5eb2ab06500a588571a6dc9e41ebf856758c36184947aa3a1da4d448

Observation 06f050c3-69ee-4cae-a1f7-a1026382da7c · outbound

This paper cites Medklip: Medical knowledge enhanced language-image pre-training for x-ray diagnosis.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Medklip: Medical knowledge enhanced language-image pre-training for x-ray diagnosis

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.591474Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-05T10:46:08.403517Z digest=sha256:74a380b06dce8457d88549d4da897ad263575da4cefa33096214e77ca6c09d71

Observation 99fd17d9-2eb1-4594-bf2f-41f070a5863a · outbound

This paper cites Unimiss: Universal medical self-supervised learning via breaking dimensionality barrier.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Unimiss: Universal medical self-supervised learning via breaking dimensionality barrier

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.577184Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-05T10:46:08.565766Z digest=sha256:3169ec3f415f845d09906f1b00a5ec4f84faff148da1aff2f36dffe8ee8316b6

Observation eb497786-556b-4405-a3d4-fa736c636040 · outbound

This paper cites Demystifying CLIP Data.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Demystifying CLIP Data

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-05T10:46:08.732381Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:46:08.732381Z digest=sha256:0f8ef914713d0fe6c2348e70d5a230d0e5f0b4ad54e189035a32748b7584f876

Observation 57eb387b-0cc9-4b4b-947e-0e076d79af67 · outbound

This paper cites Glipv2: unifying local- ization and vl understanding.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Glipv2: unifying local- ization and vl understanding

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.561088Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-05T10:46:08.830792Z digest=sha256:e6d8f6cee69f2ae89fe15fd0561b9fa73c8879b110ec00f2a40920d66077fea3

Observation 0871f1a1-4930-4509-b976-2ae4035b12ae · outbound

This paper cites Biomedgpt: A unified and generalist biomedical generative pre-trained transformer for vision, language, and multimodal tasks.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Biomedgpt: A unified and generalist biomedical generative pre-trained transformer for vision, language, and multimodal tasks

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.546935Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-05T10:46:08.923378Z digest=sha256:87b2d349f2732c3b6c40aba5724071a89ddb06e70d2bd85107c72ba60b31833a

Observation fd585d6f-5915-4473-b7d0-29ad8ac20d8c · outbound

This paper cites RadGenome-Chest CT: A Grounded Vision-Language Dataset for Chest CT Analysis.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting RadGenome-Chest CT: A Grounded Vision-Language Dataset for Chest CT Analysis

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-05T10:46:09.008709Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:46:09.008709Z digest=sha256:090886c871a0d5b281dbde5f0c1b9a6997aa9c390ebb647923789d7e019b2591

Observation 5278f4c0-7029-4709-814e-c6f077d70839 · outbound

This paper cites Development of a large-scale medical visual question-answering dataset.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Development of a large-scale medical visual question-answering dataset

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.532120Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-05T10:46:09.069233Z digest=sha256:a87924887b4e4acaea59ebe1d2362d79f47dc39b72bb3f225fc12e8ae7396c71

Observation 52da4f73-68ac-4a46-ae3b-f2ce4f16ed22 · outbound

This paper cites Each of these claims is supported by theoretical analysis, ablation studies, and experimental results.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Each of these claims is supported by theoretical analysis, ablation studies, and experimental results

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.447833Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-05T10:46:09.130493Z digest=sha256:f3297450c6d4f2d756040fad316ec1d24bdff4bd2c1df24b0ce654647a1ede9f

Observation 2e310e4c-9ace-41a2-81bd-58b7deb97277 · outbound

This paper cites Limitations.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Limitations

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.216943Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-05T10:46:09.166745Z digest=sha256:be458ca4ab87f57dab836c2c82749c2b97129eb7ce6e059502b1f033f3aaf285

Observation 655d84c8-671a-4e46-8840-a2ce1a9282b3 · outbound

This paper cites Guidelines: • The answer NA means that the paper does not include theoretical results.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Guidelines: • The answer NA means that the paper does not include theoretical results

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:12.864349Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-05T10:46:09.222371Z digest=sha256:d27035a5b7c48822cd66d3e50bf0357e4342c6b9209ce6b0c9b1c1836c408bd3

Observation 4babdb4f-17c1-4e46-af10-1ab8db572560 · outbound

This paper cites Guidelines: • The answer NA means that the paper does not include experiments.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Guidelines: • The answer NA means that the paper does not include experiments

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:12.564751Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-05T10:46:09.304182Z digest=sha256:24b2139f89539fcf096aef07ab59ddb520f4d8e8bd1e4f564d26844be720229d

Observation 42aa9b52-0cf8-4efa-b0ce-3849436a662a · outbound

This paper cites Guidelines: • The answer NA means that paper does not include experiments requiring code.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Guidelines: • The answer NA means that paper does not include experiments requiring code

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:12.259108Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-05T10:46:09.363957Z digest=sha256:06d918997fe4c489b46947c5da83077d71857316c4a025186853a48f96ad7f5f

Observation 81e41ff1-e8e3-43ed-9355-b8b9868615ca · outbound

This paper cites Guidelines: • The answer NA means that the paper does not include experiments.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Guidelines: • The answer NA means that the paper does not include experiments

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:12.017340Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-05T10:46:09.437212Z digest=sha256:135a0417bbee82c17cc940e6b7574479c02bc4996e60b0087299a9da89af2bdb

Observation e1fad113-1c30-4f52-88ea-f48e99eed218 · outbound

This paper cites Guidelines: • The answer NA means that the paper does not include experiments.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Guidelines: • The answer NA means that the paper does not include experiments

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:11.928213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-05T10:46:09.500264Z digest=sha256:5211a1cd1b94b585cbd54bdde3a69a60d084f3fe436c0cc0993becd83911ec5d

Observation 60d98338-22c4-40ca-957a-87b29c194562 · outbound

This paper cites Guidelines: • The answer NA means that the paper does not include experiments.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Guidelines: • The answer NA means that the paper does not include experiments

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:11.810015Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-05T10:46:09.569808Z digest=sha256:9c7b66907596c40130fbd57b4a01e6aa1f6665f5f00f5d974113799d60c9de6d

Observation fe82fdc4-fa2e-4d71-bcbe-0c308e46cdd5 · outbound

This paper cites All data used in this study are from publicly available, de- identified medical datasets, and no personally identifiable information (PII) was accessed or used.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting All data used in this study are from publicly available, de- identified medical datasets, and no personally identifiable information (PII) was accessed or used

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:11.712125Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-05T10:46:09.648785Z digest=sha256:d000f616563afd269d86da87571d9f8e7184fc757dadb442987f193eb4cc653c

Observation 4cbe450f-4293-472e-80de-8a150b92989b · outbound

This paper cites Guidelines: 17 • The answer NA means that there is no societal impact of the work performed.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Guidelines: 17 • The answer NA means that there is no societal impact of the work performed

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:11.553592Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-05T10:46:09.694965Z digest=sha256:ed974a7d8da8959c36020ede69bb9e88e31c4908d6059f4e0ca78274377a2a3f

Observation 91091c49-eb63-4fbe-b291-6edfceb106fd · outbound

This paper cites an unresolved cited work.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Unresolved cited work

Reference 55

Resolution
unresolved
raw_fallback, observed 2026-08-05T10:46:11.419872Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-05T10:46:09.756872Z digest=sha256:84b512b2d95125b3b4f92d9b05f37db4d984878503956061cb1d0ec03e283544

Observation 68c2e891-feb6-46ed-b4b6-2720a5c2d289 · outbound

This paper cites Guidelines: • The answer NA means that the paper does not use existing assets.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Guidelines: • The answer NA means that the paper does not use existing assets

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:11.269223Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-05T10:46:09.813237Z digest=sha256:773c02f26392d84876c8e274d8a5cb9188e49e90dcfdd4067c8ee31d5c88b511

Observation 065dba93-e8cc-466b-ba0b-2abbd75001ec · outbound

This paper cites These assets will be released with accompanying documentation upon paper acceptance.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting These assets will be released with accompanying documentation upon paper acceptance

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:11.105435Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-05T10:46:09.853132Z digest=sha256:4c5b55ca78bceaab2dd15741a44cd10f9eb398b6fddf2b91b950e5c1217d8604

Observation 5cf22d1e-5be8-4e8f-9c0d-983ca3baaf39 · outbound

This paper cites All data used are from publicly available, de-identified medical datasets with appropriate licenses and do not involve any direct interaction with individuals.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting All data used are from publicly available, de-identified medical datasets with appropriate licenses and do not involve any direct interaction with individuals

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:10.989629Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-05T10:46:09.939998Z digest=sha256:6d313e3cba49aa120cdba297159d9b2e0832c19b885b12a1d1225ba82ade8b3c

Observation 304aea77-f964-4354-8c2a-855d2a6a01d3 · outbound

This paper cites Therefore, IRB approval was not required.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Therefore, IRB approval was not required

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:10.820337Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-05T10:46:09.978316Z digest=sha256:39f004d549537139b6a0b12ffc283b41491f0432b7064648d80eab84b1245697

Observation 2905636d-5717-4d58-9d1d-4f6aeb97aac9 · outbound

This paper cites Answer: [Yes] Justification: Large language models such as GPT-4o and Qwen2.5, were used to rewrite radiology reports for improving semantic clarity during pretraining.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Answer: [Yes] Justification: Large language models such as GPT-4o and Qwen2.5, were used to rewrite radiology reports for improving semantic clarity during pretraining

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:10.677975Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-05T10:46:09.996574Z digest=sha256:64850a728e120dad9a7f4c8c17b549dfd599b1bc7c43a7061638e07e203711f6

Pith citing papers

Observation 0cd03453-c97c-4b33-9b7c-1d3daeea6f74 · inbound

Self-Supervised Dynamical System Representations for Physiological Time-Series cites this paper.

Self-Supervised Dynamical System Representations for Physiological Time-Series MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-03T19:34:04.544567Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T19:34:04.544567Z digest=sha256:6b4e08a2dc6c5057a9b218890d27e08bfb76d265b9ace670f00b92be14a17cd9

Observation e8ecbf8a-69eb-48e9-9d43-16ea5263acd2 · inbound

ASAP: Advancing Medical Volumetric Representation Learning with Anatomy-aware Semantically-adaptive Pre-training cites this paper.

ASAP: Advancing Medical Volumetric Representation Learning with Anatomy-aware Semantically-adaptive Pre-training MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting

Reference 81

Resolution
verified exact
arxiv_id, observed 2026-06-28T19:22:34.659378Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-06-28T19:16:42.139096Z digest=sha256:8017c8e582ae5bde50789e780ea431e0000dff249d91556e9bcc380b5e3fb2e0