REVIEW 13 cited by
Vision Foundation Models for Computed Tomography
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Foundation models (FMs) have shown transformative potential in radiology by performing diverse, complex tasks across imaging modalities. Here, we developed CT-FM, a large-scale 3D image-based pre-trained model designed explicitly for various radiological tasks. CT-FM was pre-trained using 148,000 computed tomography (CT) scans from the Imaging Data Commons through label-agnostic contrastive learning. We evaluated CT-FM across four categories of tasks, namely, whole-body and tumor segmentation, head CT triage, medical image retrieval, and semantic understanding, showing superior performance against state-of-the-art models. Beyond quantitative success, CT-FM demonstrated the ability to cluster regions anatomically and identify similar anatomical and structural concepts across scans. Furthermore, it remained robust across test-retest settings and indicated reasonable salient regions attached to its embeddings. This study demonstrates the value of large-scale medical imaging foundation models and by open-sourcing the model weights, code, and data, aims to support more adaptable, reliable, and interpretable AI solutions in radiology.
Forward citations
Cited by 13 Pith papers
-
Learning from Compressed CT: Feature Attention Style Transfer and Structured Factorized Projections for Resource-Efficient Medical Image Analysis
CT-Lite combines Feature Attention Style Transfer (FAST) and Structured Factorized Projections (SFP) with contrastive learning to reach AUROC within 5-7% of uncompressed baselines on compressed CT volumes across three...
-
Big, Bright, or Invisible: A Frozen-Feature Benchmark of 3D CT Foundation Models
Across ten frozen 3D CT encoders, a finding's detectability is governed by its contrast and spatial extent, not the model choice, and small low-contrast lesions remain undetectable even with linear probing.
-
Semantically Calibrated Evidence Composition for CT Vision-Language Learning
SCOPE composes organ-level CT evidence under whole-volume context and calibrates it with diagnostic-summary supervision, reaching 85.0/72.2 macro AUC on CT-RATE and RadChestCT.
-
Learning How Much, Not Just What: Cross-Patient Burden Order for CT Vision-Language Pretraining
Cross-patient report-based pair mining plus burden-direction alignment improves CT vision-language pretraining, reaching 85.6 AUROC on CT-RATE zero-shot diagnosis.
-
Empirical investigation of 3D CT Foundation Models and Unsupervised Adaptation for Head and Neck Cancer Recurrence Prediction
Existing 3D CT foundation models generalize poorly to external head and neck cancer cohorts; CT-CLIP was the most robust, and gating fusion of imaging with clinical data gave the best prediction.
-
M3Ret: Unleashing Zero-shot Multimodal Medical Image Retrieval via Self-Supervision
One self-supervised encoder trained on unpaired X-ray, ultrasound, endoscopy, and CT data gives competitive zero-shot retrieval and seems to generalize to unseen MRI tasks.
-
Disorder-induced stress-flow misalignment in soft glassy materials revealed using multi-directional shear
Soft glassy materials show a transient stress response orthogonal to a newly applied shear direction, which a mesoscopic elasto-plastic model attributes to local yield-stress disorder.
-
CADS: A Comprehensive Anatomical Dataset and Segmentation for Whole-Body Anatomy in Computed Tomography
A new public dataset of 22,022 CT volumes labeled for 167 structures, and a nnU-Net model trained on it, outperform TotalSegmentator on most shared structures and expand coverage.
-
4KAgent: Agentic Any Image to 4K Super-Resolution
An agentic pipeline that plans and executes image restoration from a toolbox of pretrained models to upscale arbitrary images to 4K, reporting state-of-the-art results on many benchmarks.
-
Same Branches, Different Trees: A Bifurcation Connectedness Metric for Coronary Artery Segmentation and FFR-CT Decision Agreement
Bifurcation Connectedness Score (BCS) measures junction-level vessel connectivity that Dice misses, tracks geometric FFR-CT decision agreement in severe disease, and shows branch recovery and tree connectedness are se...
-
Anatomy Contextualized Adaption of CT Foundation Models
A lightweight inter-anatomy transformer on frozen CT foundation embeddings plus dual anatomy/scan contrastive losses beats global and fine-grained baselines on Merlin and CT-RATE zero-shot finding classification.
-
Comparing the Performance of Foundation Model Derived Embeddings with Traditional Approaches for Distant Metastasis Prediction in Head and Neck Cancer
CT Foundation embeddings predict 2-year distant metastasis in head and neck cancer with AUC 0.791, matching radiomics-plus-deep-learning (0.794) while requiring no tumor contours.
-
Rethinking Artificial Intelligence in Medical Imaging: Assumptions, Reality, and Reframing
Limited bedside impact of medical imaging AI stems from structural misalignment with clinical decision-making across six dimensions, not mainly from weak algorithms, regulation, or explainability.
Discussion (0). Continue with ORCID to comment.