Pith. sign in

REVIEW 1 cited by

"Name that manufacturer". Relating image acquisition bias with task complexity when training deep learning models: experiments on head CT

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2008.08525 v1 pith:DW56BKLR submitted 2020-08-19 eess.IV cs.CVcs.LG

classification eess.IVcs.CVcs.LG
keywords biasmodelslearningclassificationclinicaldatadatasetsegmentation
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

As interest in applying machine learning techniques for medical images continues to grow at a rapid pace, models are starting to be developed and deployed for clinical applications. In the clinical AI model development lifecycle (described by Lu et al. [1]), a crucial phase for machine learning scientists and clinicians is the proper design and collection of the data cohort. The ability to recognize various forms of biases and distribution shifts in the dataset is critical at this step. While it remains difficult to account for all potential sources of bias, techniques can be developed to identify specific types of bias in order to mitigate their impact. In this work we analyze how the distribution of scanner manufacturers in a dataset can contribute to the overall bias of deep learning models. We evaluate convolutional neural networks (CNN) for both classification and segmentation tasks, specifically two state-of-the-art models: ResNet [2] for classification and U-Net [3] for segmentation. We demonstrate that CNNs can learn to distinguish the imaging scanner manufacturer and that this bias can substantially impact model performance for both classification and segmentation tasks. By creating an original synthesis dataset of brain data mimicking the presence of more or less subtle lesions we also show that this bias is related to the difficulty of the task. Recognition of such bias is critical to develop robust, generalizable models that will be crucial for clinical applications in real-world data distributions.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Understanding Dataset Bias in Medical Imaging: A Case Study on Chest X-rays

    cs.CV 2025-07 conditional novelty 6.0 of 10

    Classifiers can identify the origin of chest X-rays across NIH, CheXpert, MIMIC-CXR, and PadChest with F1 scores up to about 99%, and the bias appears driven mainly by pixel intensity and texture.

Pith tools