Pith. sign in

Paper Citation Record · LEDGER

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting

As of 22 August 2026, this Paper Citation Record lists 60 of 60 outbound references and 2 inbound Pith citation observations for arXiv:2509.03800.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2509.03800 v1

Coverage vector

measured 60 of 60 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T10:46:09.996574Z

measured 62 of 62 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-03T19:34:04.544567Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-28T19:22:34.657932Z

Reference resolution

60 of 60 outbound references displayed

  • verified exact1
  • verified fuzzy39
  • unresolved20
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a6fa80d1-f5d6-4f4a-8c11-f7590af040bf · outbound

This paper cites Merlin: A vision language foundation model for 3d computed tomography.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Merlin: A vision language foundation model for 3d computed tomography

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-05T10:46:04.243549Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:46:04.243549Z digest=sha256:2455174a5bdc51d462c8ecfcb2d72bc4330b08f46b91a06ee411bf5327e5ea60

Observation b33a7ab9-9f9f-45d3-8445-97b33ff2b16b · outbound

This paper cites A vision–language foundation model for the generation of realistic chest x-ray images.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting A vision–language foundation model for the generation of realistic chest x-ray images

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.958572Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T10:46:04.330865Z digest=sha256:27e1c0d0e0231b7172b5c3afa39bc4b53433a490e77b66b3713bc533a215d14a

Observation cb37fff5-0668-4885-9707-e0ec24f3015e · outbound

This paper cites Making the most of text semantics to improve biomedical vision–language processing.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Making the most of text semantics to improve biomedical vision–language processing

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-05T10:46:04.426003Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:46:04.426003Z digest=sha256:174915b43604b3a8652da8a1bc60274086db1c56016d7cf00460c7c530e8ab1a

Observation d404663f-9670-4008-8297-2b334af143b6 · outbound

This paper cites Understanding and confronting our mis- takes: the epidemiology of error in radiology and strategies for error reduction.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Understanding and confronting our mis- takes: the epidemiology of error in radiology and strategies for error reduction

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.932829Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T10:46:04.516487Z digest=sha256:38fdd804f83ff88eb617cad40fd844471c2c830856356386761c1beb38536260

Observation c143d303-e11e-4e10-bd16-89cf43f4e210 · outbound

This paper cites Joint modeling of chest radiographs and radiology reports for pulmonary edema assessment.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Joint modeling of chest radiographs and radiology reports for pulmonary edema assessment

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.917747Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T10:46:04.585735Z digest=sha256:e33831f9d478718bcfb48c5c77aa8981cc91b5c912b43c5446c4c28680d7ed35

Observation 41ed5110-7327-4e3d-bafa-1a1622ef2ee5 · outbound

This paper cites Contrastive Localized Language-Image Pre-Training.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Contrastive Localized Language-Image Pre-Training

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-05T10:46:04.661105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:46:04.661105Z digest=sha256:783347f5fe10ba11b8e01950633d879938eb5c5df9c77540291b37f6e365d77f

Observation 3f111173-6cf8-4d07-8f39-b1d2c4ddf619 · outbound

This paper cites A review of medical image data augmentation techniques for deep learning applications.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting A review of medical image data augmentation techniques for deep learning applications

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.900852Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T10:46:04.776379Z digest=sha256:83791bef2e8431d08f1e72f305b33285d98c3099cf26992b3adf376ca9eaa32f

Observation f555adc7-97f4-414c-8a7b-5d9ba7d8b2d8 · outbound

This paper cites Machine-learning-based multiple abnormality prediction with large-scale chest computed tomography volumes.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Machine-learning-based multiple abnormality prediction with large-scale chest computed tomography volumes

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.884634Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T10:46:04.842478Z digest=sha256:55e438f24669df264cd2128a91a189a011c029028f5bf5695f40c369c6ee9d1d

Observation 85df312c-28df-4027-8a32-0cf6486e38a7 · outbound

This paper cites Multi-View and Multi-Scale Alignment for Contrastive Language-Image Pre-training in Mammography.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Multi-View and Multi-Scale Alignment for Contrastive Language-Image Pre-training in Mammography

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-08-05T10:46:10.536952Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T10:46:04.962958Z digest=sha256:888527565b03afa30df9f1403a8bf18dffef15b705889d415379e820d133c5e0

Observation 4feab371-5592-4901-b37e-221704c40765 · outbound

This paper cites The Llama 3 Herd of Models.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting The Llama 3 Herd of Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-05T10:46:05.092590Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:46:05.092590Z digest=sha256:01ef52513d8ae509a9ef4e721c2b7cc43950fa0034ac2972ebbb5786483746b9

Observation 13ac3f6d-a019-4f0b-825c-d36d7414fa43 · outbound

This paper cites Devel- oping generalist foundation models from a multimodal dataset for 3d computed tomography.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Devel- oping generalist foundation models from a multimodal dataset for 3d computed tomography

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-05T10:46:05.191575Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:46:05.191575Z digest=sha256:ceaec1c88b772eb2fa5f62890e062060a27e3fc182e9543aaf87176b138573b7

Observation ac57e9c6-b566-4905-a756-dd5818e81d92 · outbound

This paper cites LoRA: Low-Rank Adaptation of Large Language Models.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting LoRA: Low-Rank Adaptation of Large Language Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-05T10:46:05.296884Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:46:05.296884Z digest=sha256:e410c1e8882e0dda4920cbbcccf069102b194246f32e9c2592a1416677a4019f

Observation aa273c70-1ad5-4176-a12e-25ee2a518a90 · outbound

This paper cites Gloria: A multimodal global-local representation learning framework for label-efficient medical image recognition.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Gloria: A multimodal global-local representation learning framework for label-efficient medical image recognition

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-05T10:46:05.360284Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:46:05.360284Z digest=sha256:8080cacd9ca13925bc11c5a64640fb2c924acbec8d9257f61ff884b8ba85deac

Observation c7bf766a-ec65-41da-a482-156f75f1b8da · outbound

This paper cites Enhancing representation in medical vision-language foun- dation models via multi-scale information extraction techniques.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Enhancing representation in medical vision-language foun- dation models via multi-scale information extraction techniques

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.859199Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T10:46:05.458401Z digest=sha256:c4ba27e3ad138cccf570355a8d2b29709aebe5d14eee4eb99c887c6819265717

Observation da68f205-0616-4a33-8ce2-23eb68858145 · outbound

This paper cites STU-Net: Scalable and Transferable Medical Image Segmentation Models Empowered by Large-Scale Supervised Pre-training.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting STU-Net: Scalable and Transferable Medical Image Segmentation Models Empowered by Large-Scale Supervised Pre-training

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-05T10:46:05.562057Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:46:05.562057Z digest=sha256:f8c99e3e3a03c67544e821200453a3125c50968c3e20a965bae3c28cc5a00ddf

Observation 4dbc9100-63b5-4b79-b258-fb0395cb7c62 · outbound

This paper cites nnu-net: a self-configuring method for deep learning-based biomedical image segmentation.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting nnu-net: a self-configuring method for deep learning-based biomedical image segmentation

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-05T10:46:05.633725Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:46:05.633725Z digest=sha256:291abffc33aac13d8066df582d93da9ad9bb83cb7d1995ef375306fbc1c54a29

Observation 0d20a01c-58e5-4de6-8e64-28d7416906b8 · outbound

This paper cites Fool me twice: delayed diagnoses in radiology with emphasis on perpetuated errors.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Fool me twice: delayed diagnoses in radiology with emphasis on perpetuated errors

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.834498Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T10:46:05.740237Z digest=sha256:fa1171a2775b788edca863467278957d039b37be6cc75a039792a70215ca82d4

Observation 0ef9eb6e-3c40-4f58-8234-d3ac1161d8f4 · outbound

This paper cites Generating synthetic data for medical imaging.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Generating synthetic data for medical imaging

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.819212Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T10:46:05.837515Z digest=sha256:b27e69e4d1678817c5747b3b5ca2c52662efeeb4a377771bfdc35bd30a3eba8e

Observation 35d0a13e-01ef-4672-bf63-cc177925cee6 · outbound

This paper cites Cxr-llava: a multimodal large language model for interpreting chest x-ray images.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Cxr-llava: a multimodal large language model for interpreting chest x-ray images

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.803136Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T10:46:05.952126Z digest=sha256:c99d6609aaf33b73098fe8ed207fbac65d3470fa15aeae426b970c11eee05305

Observation 71cb91cd-edae-46b0-9c62-195ae4964bf7 · outbound

This paper cites Llava-med: Training a large language-and-vision assistant for biomedicine in one day.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Llava-med: Training a large language-and-vision assistant for biomedicine in one day

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-05T10:46:06.026153Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:46:06.026153Z digest=sha256:96b50139d4c6e4607fbb07c6fb33a8ce0a6922e684743086d509d582686013ca

Observation 6b547427-1974-4c16-a177-2ff59895bdc3 · outbound

This paper cites Artificial general intelligence for medical imaging analysis.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Artificial general intelligence for medical imaging analysis

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.775179Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T10:46:06.144188Z digest=sha256:a8f0973f4f0ca3f88bbc4ce8dfdcd8f38a0f2368f2ef447cc3895e24ae6b33de

Observation 9920dd3d-25af-4526-9efc-147946bfd878 · outbound

This paper cites Ct-glip: 3d grounded language-image pretraining with ct scans and radiology reports for full-body scenarios.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Ct-glip: 3d grounded language-image pretraining with ct scans and radiology reports for full-body scenarios

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-05T10:46:06.214147Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:46:06.214147Z digest=sha256:98ad9d564b782c2b3b8d3e32a6e920691a9fbfe73510fbd3f263125a109f3737

Observation 045e8575-14d9-458f-90dc-c8791db2b0af · outbound

This paper cites Pmc-clip: Contrastive language-image pre-training using biomedical documents.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Pmc-clip: Contrastive language-image pre-training using biomedical documents

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-05T10:46:06.347751Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:46:06.347751Z digest=sha256:2f6dd13866154b024ffe61cf3205fc5e2ac58e85b2d2c9b3b50e6b19575ab0f9

Observation 56045d27-ec37-4ba5-a6bd-9f2a0988974a · outbound

This paper cites Can Medical Vision-Language Pre-training Succeed with Purely Synthetic Data?.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Can Medical Vision-Language Pre-training Succeed with Purely Synthetic Data?

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-05T10:46:06.490159Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:46:06.490159Z digest=sha256:d096806b43ad7e99b06f0386f3155d50d1e244b0fbf63a12cfb750ad484c65f2

Observation 7cfb17f5-68a2-4784-ae44-4289145e1d0d · outbound

This paper cites Improved baselines with visual instruction tuning.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Improved baselines with visual instruction tuning

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-05T10:46:06.654963Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:46:06.654963Z digest=sha256:1aef12c9d945f2c3cc917fbc9d64ee4c6d5d8b64569bc76cf40a75c56efc6bad

Observation 08c8a7a7-1b04-4673-b65c-8d991d660f66 · outbound

This paper cites Representation Learning with Contrastive Predictive Coding.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Representation Learning with Contrastive Predictive Coding

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-05T10:46:06.818244Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:46:06.818244Z digest=sha256:5984422f7075d2c28c4b1f733fb4ade8b3c46bf91f60ee30483307338f0c87c5

Observation 051d5200-2466-433f-948e-fa46ba9c4c30 · outbound

This paper cites Unsupervised medical image translation with adversarial diffusion models.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Unsupervised medical image translation with adversarial diffusion models

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.736423Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T10:46:06.953711Z digest=sha256:536bf9d19b0cebb584d3f740211ea56e5e0e4fc605fe4796ad9cb6df71ab3896

Observation 5a6d5801-2d1a-43c9-bdba-d88c5914c4df · outbound

This paper cites On variational bounds of mutual information.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting On variational bounds of mutual information

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.720929Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T10:46:07.058611Z digest=sha256:c388c13baf859dc103c6b43675fc5c6578405cae0204990eba3808bf978d84d7

Observation 04d8ab2e-3c09-4969-8f54-f0c86ead5096 · outbound

This paper cites Learning transferable visual models from natural language supervision.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Learning transferable visual models from natural language supervision

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-05T10:46:07.172235Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:46:07.172235Z digest=sha256:7b7805519b7f95b6af1e94ee551402df3ae35d02f1a159805ddfc9c550d11f62

Observation 904762dd-f001-4ef0-96f7-33945a9c2c06 · outbound

This paper cites Study of thoracic ct in covid-19: the stoic project.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Study of thoracic ct in covid-19: the stoic project

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.695888Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T10:46:07.275215Z digest=sha256:bd70f0d6762e0d8e108e2a569d94eb0eda9de2a0f1a0055ab8b23c605ba5e13a

Observation 7526b28c-0615-4fd0-8f25-a789a2a3a0d4 · outbound

This paper cites Deep learning in medical image analysis.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Deep learning in medical image analysis

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.680527Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T10:46:07.401048Z digest=sha256:4fd37a4d083d425b7d19c0d245d479d8f8ade3adc8b72b29fdc205aa1b000c3a

Observation 8cea4486-36cd-46ca-a3e0-800c3835fb65 · outbound

This paper cites Large-scale and fine-grained vision- language pre-training for enhanced ct image understanding.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Large-scale and fine-grained vision- language pre-training for enhanced ct image understanding

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.666394Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T10:46:07.524178Z digest=sha256:ead8098fed923602228c7ce44ad21911eac02ba5e4bc815c61b8f7fbc0786efa

Observation d6fa8c35-2d72-432f-8f29-319783f54832 · outbound

This paper cites Bioclip: A vision foundation model for the tree of life.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Bioclip: A vision foundation model for the tree of life

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.651773Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T10:46:07.697454Z digest=sha256:23f380c3df725f472d3ec13157130b23d1351d5eb02c16f24250a604dba93a7c

Observation 196f8bc6-9cd3-4d50-ba35-9d041cf1c193 · outbound

This paper cites XrayGPT: Chest Radiographs Summarization using Medical Vision-Language Models.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting XrayGPT: Chest Radiographs Summarization using Medical Vision-Language Models

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-05T10:46:07.826433Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:46:07.826433Z digest=sha256:a5bc48163334fbdc6d4b93f467b8931427e41e6ac82d9f78218842087e60a5ff

Observation fe0aa725-db38-4b9a-9523-c2a648ecd696 · outbound

This paper cites Communication errors in radiology–pitfalls and how to avoid them.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Communication errors in radiology–pitfalls and how to avoid them

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.637015Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T10:46:07.970509Z digest=sha256:6163b302a3a3a1df8fbbaccda531f30ec6664dd123481b1dd34151a2e197c63d

Observation 75a176bc-5a08-4d29-9b3e-e14cd42ca5df · outbound

This paper cites Multi- granularity cross-modal alignment for generalized medical visual representation learning.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Multi- granularity cross-modal alignment for generalized medical visual representation learning

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.621985Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T10:46:08.098509Z digest=sha256:fc86c079fd5e4a516c5538f1d78ba123f120d67a24cdd620b22d1b01a533b4ee

Observation 01ab2a9b-31b2-4afe-9136-fa8d0a2ff7f7 · outbound

This paper cites Totalseg- mentator: robust segmentation of 104 anatomic structures in ct images.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Totalseg- mentator: robust segmentation of 104 anatomic structures in ct images

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.608638Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T10:46:08.279430Z digest=sha256:827c12deda9821ab63e77e57663970bb7bab49a82361b277ea086fe555db1e0c

Observation 06f050c3-69ee-4cae-a1f7-a1026382da7c · outbound

This paper cites Medklip: Medical knowledge enhanced language-image pre-training for x-ray diagnosis.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Medklip: Medical knowledge enhanced language-image pre-training for x-ray diagnosis

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.591474Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T10:46:08.403517Z digest=sha256:4320136516e250cc17413a6fcef132a843fe3242987c7ed05333f7986b8019f2

Observation 99fd17d9-2eb1-4594-bf2f-41f070a5863a · outbound

This paper cites Unimiss: Universal medical self-supervised learning via breaking dimensionality barrier.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Unimiss: Universal medical self-supervised learning via breaking dimensionality barrier

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.577184Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T10:46:08.565766Z digest=sha256:2842d8e3b90a7accc9e97b802cbc467a5fd4b277860168faea03bc02a740259a

Observation eb497786-556b-4405-a3d4-fa736c636040 · outbound

This paper cites Demystifying CLIP Data.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Demystifying CLIP Data

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-05T10:46:08.732381Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:46:08.732381Z digest=sha256:c7d4ef576bb88be79205f9190c890fafffbdda6abb1b7bdf0c70cb81a932e06a

Observation 57eb387b-0cc9-4b4b-947e-0e076d79af67 · outbound

This paper cites Glipv2: unifying local- ization and vl understanding.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Glipv2: unifying local- ization and vl understanding

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.561088Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T10:46:08.830792Z digest=sha256:0cef6f0e87afe14dc2d4d273e6e4b32c33674b54830cd63dc53d441e60bc9d8c

Observation 0871f1a1-4930-4509-b976-2ae4035b12ae · outbound

This paper cites Biomedgpt: A unified and generalist biomedical generative pre-trained transformer for vision, language, and multimodal tasks.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Biomedgpt: A unified and generalist biomedical generative pre-trained transformer for vision, language, and multimodal tasks

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.546935Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T10:46:08.923378Z digest=sha256:454f96e00c800018f251d530977d4243a149250881d08b89706ada627d6161ef

Observation fd585d6f-5915-4473-b7d0-29ad8ac20d8c · outbound

This paper cites RadGenome-Chest CT: A Grounded Vision-Language Dataset for Chest CT Analysis.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting RadGenome-Chest CT: A Grounded Vision-Language Dataset for Chest CT Analysis

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-05T10:46:09.008709Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:46:09.008709Z digest=sha256:90b46cb37d332b22d999ec57d1bdf34478f5f03eef841319250260a237c7a4b4

Observation 5278f4c0-7029-4709-814e-c6f077d70839 · outbound

This paper cites Development of a large-scale medical visual question-answering dataset.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Development of a large-scale medical visual question-answering dataset

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.532120Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T10:46:09.069233Z digest=sha256:c82809bcbc6177633831ec0784790af11f4530f8c04ab76555eee79ffef830c5

Observation 52da4f73-68ac-4a46-ae3b-f2ce4f16ed22 · outbound

This paper cites Each of these claims is supported by theoretical analysis, ablation studies, and experimental results.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Each of these claims is supported by theoretical analysis, ablation studies, and experimental results

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.447833Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T10:46:09.130493Z digest=sha256:3d9dd3918ad35af491d8c959a782f8c9753f67029f08fa0db340923dd7894444

Observation 2e310e4c-9ace-41a2-81bd-58b7deb97277 · outbound

This paper cites Limitations.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Limitations

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:13.216943Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T10:46:09.166745Z digest=sha256:ae35d75760e215aab4b3926cf7697dfecf3c639aee4e61c4684a379c11f16d03

Observation 655d84c8-671a-4e46-8840-a2ce1a9282b3 · outbound

This paper cites Guidelines: • The answer NA means that the paper does not include theoretical results.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Guidelines: • The answer NA means that the paper does not include theoretical results

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:12.864349Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T10:46:09.222371Z digest=sha256:74861cfc48f50fb817635fef5be911bc3a2fec2dfbbb42a06076bcacb86874fa

Observation 4babdb4f-17c1-4e46-af10-1ab8db572560 · outbound

This paper cites Guidelines: • The answer NA means that the paper does not include experiments.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Guidelines: • The answer NA means that the paper does not include experiments

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:12.564751Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T10:46:09.304182Z digest=sha256:c333d0380a3e52e259754fd038121c4ea23faaa1698a0563e1116015cf748e3e

Observation 42aa9b52-0cf8-4efa-b0ce-3849436a662a · outbound

This paper cites Guidelines: • The answer NA means that paper does not include experiments requiring code.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Guidelines: • The answer NA means that paper does not include experiments requiring code

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:12.259108Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T10:46:09.363957Z digest=sha256:0174d33b5e5a524fa10d32819404cc1af61ed843663d3046055aa9d1a2b350bc

Observation 81e41ff1-e8e3-43ed-9355-b8b9868615ca · outbound

This paper cites Guidelines: • The answer NA means that the paper does not include experiments.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Guidelines: • The answer NA means that the paper does not include experiments

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:12.017340Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T10:46:09.437212Z digest=sha256:f7409af882fe8f67ece3d8b17562bb812219385142c319b5441d25bd98176f06

Observation e1fad113-1c30-4f52-88ea-f48e99eed218 · outbound

This paper cites Guidelines: • The answer NA means that the paper does not include experiments.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Guidelines: • The answer NA means that the paper does not include experiments

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:11.928213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T10:46:09.500264Z digest=sha256:aa3831dc9249894b4036aa6df1b51189a499034ae12ef7e53923c4b586eef95a

Observation 60d98338-22c4-40ca-957a-87b29c194562 · outbound

This paper cites Guidelines: • The answer NA means that the paper does not include experiments.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Guidelines: • The answer NA means that the paper does not include experiments

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:11.810015Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T10:46:09.569808Z digest=sha256:f6a787db245df042724a31c8ca92971475375d39c2c8838d1d237563f6491fac

Observation fe82fdc4-fa2e-4d71-bcbe-0c308e46cdd5 · outbound

This paper cites All data used in this study are from publicly available, de- identified medical datasets, and no personally identifiable information (PII) was accessed or used.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting All data used in this study are from publicly available, de- identified medical datasets, and no personally identifiable information (PII) was accessed or used

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:11.712125Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T10:46:09.648785Z digest=sha256:45312c5de6843ebbe356401f66f25d0af29fef6c127726ddb1328e4f3c79522e

Observation 4cbe450f-4293-472e-80de-8a150b92989b · outbound

This paper cites Guidelines: 17 • The answer NA means that there is no societal impact of the work performed.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Guidelines: 17 • The answer NA means that there is no societal impact of the work performed

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:11.553592Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T10:46:09.694965Z digest=sha256:d1c9dca2453cde8506cbdc51161be5577a445ababccbf83976ac41026aa37c77

Observation 91091c49-eb63-4fbe-b291-6edfceb106fd · outbound

This paper cites an unresolved cited work.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Unresolved cited work

Reference 55

Resolution
unresolved
raw_fallback, observed 2026-08-05T10:46:11.419872Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T10:46:09.756872Z digest=sha256:0a7c5ca0c3524a9edbe8ee0b66947021d2025bb938a8fff26100bf67680e5b8c

Observation 68c2e891-feb6-46ed-b4b6-2720a5c2d289 · outbound

This paper cites Guidelines: • The answer NA means that the paper does not use existing assets.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Guidelines: • The answer NA means that the paper does not use existing assets

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:11.269223Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T10:46:09.813237Z digest=sha256:248fad8de3ef7207c1b88977832caf635c0ef9316efc081d51d3c2a617c3fc48

Observation 065dba93-e8cc-466b-ba0b-2abbd75001ec · outbound

This paper cites These assets will be released with accompanying documentation upon paper acceptance.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting These assets will be released with accompanying documentation upon paper acceptance

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:11.105435Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T10:46:09.853132Z digest=sha256:2f2a3ec3fa01bfcebbb035abe5ff588574dab6d5c97ff19811817f9d9fbd2e2c

Observation 5cf22d1e-5be8-4e8f-9c0d-983ca3baaf39 · outbound

This paper cites All data used are from publicly available, de-identified medical datasets with appropriate licenses and do not involve any direct interaction with individuals.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting All data used are from publicly available, de-identified medical datasets with appropriate licenses and do not involve any direct interaction with individuals

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:10.989629Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T10:46:09.939998Z digest=sha256:8346c9e094ad47c130880796cb7b8870f9a608c04cee94ebdc9bc1b3c3a62484

Observation 304aea77-f964-4354-8c2a-855d2a6a01d3 · outbound

This paper cites Therefore, IRB approval was not required.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Therefore, IRB approval was not required

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:10.820337Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T10:46:09.978316Z digest=sha256:db43172029a212cba80a688bc9e132e9c9c31f2d5dc0b4e93d6cda532a740397

Observation 2905636d-5717-4d58-9d1d-4f6aeb97aac9 · outbound

This paper cites Answer: [Yes] Justification: Large language models such as GPT-4o and Qwen2.5, were used to rewrite radiology reports for improving semantic clarity during pretraining.

MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting Answer: [Yes] Justification: Large language models such as GPT-4o and Qwen2.5, were used to rewrite radiology reports for improving semantic clarity during pretraining

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:46:10.677975Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T10:46:09.996574Z digest=sha256:1b5f1c365ee51270a432d71afa5db1da6da875a7cac61074da8271831f745ed8

Pith citing papers

Observation 0cd03453-c97c-4b33-9b7c-1d3daeea6f74 · inbound

Self-Supervised Dynamical System Representations for Physiological Time-Series cites this paper.

Self-Supervised Dynamical System Representations for Physiological Time-Series MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-03T19:34:04.544567Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T19:34:04.544567Z digest=sha256:49121d81784f328c5930cff847bfd26a441f808f1ed14f1e50ef19c2fec200ae

Observation e8ecbf8a-69eb-48e9-9d43-16ea5263acd2 · inbound

ASAP: Advancing Medical Volumetric Representation Learning with Anatomy-aware Semantically-adaptive Pre-training cites this paper.

ASAP: Advancing Medical Volumetric Representation Learning with Anatomy-aware Semantically-adaptive Pre-training MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting

Reference 81

Resolution
verified exact
arxiv_id, observed 2026-06-28T19:22:34.659378Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-28T19:16:42.139096Z digest=sha256:b2eb9aa49b5370582dae81662458e48111c10b6acd493d4e22de0794b56a6ac6