Pith. sign in

REVIEW 1 cited by

Cross-validation failure: small sample sizes lead to large error bars

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 1706.07581 v1 pith:ALTB3PJU submitted 2017-06-23 q-bio.QM stat.MEstat.ML

classification q-bio.QMstat.MEstat.ML
keywords errorbarscross-validationlargemanysampleacrossbiomarkers
verification ladder T0 review T1 audit T2 compute T3 formal

Signed reviews

No signed human review yet.

0 comments
abstract

Predictive models ground many state-of-the-art developments in statistical brain image analysis: decoding, MVPA, searchlight, or extraction of biomarkers. The principled approach to establish their validity and usefulness is cross-validation, testing prediction on unseen data. Here, I would like to raise awareness on error bars of cross-validation, which are often underestimated. Simple experiments show that sample sizes of many neuroimaging studies inherently lead to large error bars, eg $\pm$10% for 100 samples. The standard error across folds strongly underestimates them. These large error bars compromise the reliability of conclusions drawn with predictive models, such as biomarkers or methods developments where, unlike with cognitive neuroimaging MVPA approaches, more samples cannot be acquired by repeating the experiment across many subjects. Solutions to increase sample size must be investigated, tackling possible increases in heterogeneity of the data.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Exhaustive Model Selection in $b \to s \ell \ell$ Decays: Pitting Cross-Validation against AIC$_c$

    hep-ph 2019-08 conditional novelty 6.0 of 10

    An exhaustive model selection over 511 Wilson-coefficient combinations for b to s l l decays finds that all surviving scenarios contain O9 (Delta C9 near -1.1 to -1.4), the only surviving one-operator scenario.

Pith tools