REVIEW 9 cited by
AMOS: A Large-Scale Abdominal Multi-Organ Benchmark for Versatile Medical Image Segmentation
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Despite the considerable progress in automatic abdominal multi-organ segmentation from CT/MRI scans in recent years, a comprehensive evaluation of the models' capabilities is hampered by the lack of a large-scale benchmark from diverse clinical scenarios. Constraint by the high cost of collecting and labeling 3D medical data, most of the deep learning models to date are driven by datasets with a limited number of organs of interest or samples, which still limits the power of modern deep models and makes it difficult to provide a fully comprehensive and fair estimate of various methods. To mitigate the limitations, we present AMOS, a large-scale, diverse, clinical dataset for abdominal organ segmentation. AMOS provides 500 CT and 100 MRI scans collected from multi-center, multi-vendor, multi-modality, multi-phase, multi-disease patients, each with voxel-level annotations of 15 abdominal organs, providing challenging examples and test-bed for studying robust segmentation algorithms under diverse targets and scenarios. We further benchmark several state-of-the-art medical segmentation models to evaluate the status of the existing methods on this new challenging dataset. We have made our datasets, benchmark servers, and baselines publicly available, and hope to inspire future research. Information can be found at https://amos22.grand-challenge.org.
Forward citations
Cited by 9 Pith papers
-
Learning Segmentation from Radiology Reports
R-Super converts tumor count, size, and location information from radiology reports into voxel-wise losses that improve CT tumor segmentation beyond training with masks alone.
-
HyperSORT: Self-Organising Robust Training with hyper-networks
A hyper-network that predicts segmentation UNet weights from per-sample learned latent vectors yields a structured map of annotation styles and a way to flag erroneous labels.
-
DrVD-Bench: Do Vision-Language Models Reason Like Human Doctors in Medical Image Diagnosis?
A new five-level medical imaging benchmark, DrVD-Bench, shows that vision-language models lose accuracy sharply as reasoning complexity grows and often diagnose without grounding in lesion evidence.
-
Curia-MAE: Multi-Modal Multi-Anatomy MAE Pre-Training for 3D Medical Image Segmentation
Curia-MAE, a multi-modal multi-anatomy masked autoencoder, improves frozen-encoder 3D medical segmentation accuracy over a strong MAE baseline on eight benchmarks.
-
SegMoTE: Token-Level Mixture of Experts for Medical Image Segmentation
SegMoTE shows that adding token-level mixture-of-experts routing to a frozen SAM decoder can match or beat medical-segmentation models trained on far more data, using 0.15M curated masks and 17M trainable parameters.
-
Large-scale Multi-sequence Pretraining for Generalizable MRI Analysis in Versatile Clinical Applications
A four-objective self-supervised pretraining recipe on a 336k-volume multi-sequence MRI corpus yields first-rank transfer on 39 of 44 downstream MRI tasks.
-
Is Visual in-Context Learning for Compositional Medical Tasks within Reach?
Training on synthetic compositional task sequences with sequence-level masking lets a transformer-based in-context learner follow multi-step medical imaging instructions on held-out images, but well below codebook upp...
-
Good Enough? An Investigation on the Impact of Label Quality in Large-Scale Medical Datasets
Label quality matters little when pre-training medical segmentation models, but still matters for in-domain deployment; only large quality gaps affect transfer results.
-
The Large Cancer Assistant (LCA): A Model-Agnostic Orchestration Framework for Scalable Clinical Decision Support in Oncology
A formal orchestration framework for oncology AI pipelines is proposed, demonstrating that routing logic and output schema remain invariant under model substitution via a proof-of-concept with synthetic stubs.
Discussion (0). Continue with ORCID to comment.