Pith. sign in

REVIEW 4 cited by

Safe Testing

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 1906.07801 v5 pith:72H6GACA submitted 2019-06-18 math.ST cs.ITcs.LGmath.ITstat.MEstat.TH

Safe Testing

classification math.ST cs.ITcs.LGmath.ITstat.MEstat.TH
keywords e-valuessafetestingcontinuationoptionalseveraltheoryacceptable
verification ladder T0 review T1 audit T2 compute T3 formal T4 reserved
0 comments
read the original abstract

We develop the theory of hypothesis testing based on the e-value, a notion of evidence that, unlike the p-value, allows for effortlessly combining results from several studies in the common scenario where the decision to perform a new study may depend on previous outcomes. Tests based on e-values are safe, i.e. they preserve Type-I error guarantees, under such optional continuation. We define growth-rate optimality (GRO) as an analogue of power in an optional continuation context, and we show how to construct GRO e-variables for general testing problems with composite null and alternative, emphasizing models with nuisance parameters. GRO e-values take the form of Bayes factors with special priors. We illustrate the theory using several classic examples including a one-sample safe t-test and the 2 x 2 contingency table. Sharing Fisherian, Neymanian and Jeffreys-Bayesian interpretations, e-values may provide a methodology acceptable to adherents of all three schools.

discussion (0)

Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.

Forward citations

Cited by 4 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

  1. Optimal Posterior E-values with Non-Convex Parameter Sets with Applications to Voting Systems

    math.ST 2026-06 unverdicted novelty 7.0

    A framework for optimal posterior e-values with non-convex composite hypotheses, demonstrated via statistical tests for multiple voting systems including the first treatment of Schulze.

  2. Asymptotically Log-Optimal Bayes-Assisted Confidence Sequences for Bounded Means

    stat.ML 2026-05 unverdicted novelty 7.0

    A Bayesian predictive model adaptively constructs asymptotically log-optimal confidence sequences for bounded means using test martingales.

  3. Asymptotically Log-Optimal Bayes-Assisted Confidence Sequences for Bounded Means

    stat.ML 2026-05 unverdicted novelty 7.0

    A Bayesian predictive model adaptively selects martingale factors to construct asymptotically log-optimal confidence sequences for bounded means while preserving anytime validity under misspecification.

  4. Efficient Sequential Evaluation of Large Language Models

    stat.ML 2026-07 conditional novelty 5.0

    A confidence-sequence framework for sequentially estimating an LLM's average benchmark accuracy under adaptive question selection, with growth-oriented sampling rules that in practice often lose to uniform sampling.