Pith. sign in

REVIEW 4 cited by

Machine learning assisted Bayesian model comparison: learnt harmonic mean estimator

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2111.12720 v3 pith:TSOGHXH5 submitted 2021-11-24 stat.ME astro-ph.IMstat.CO

classification stat.MEastro-ph.IMstat.CO
keywords estimatorharmonicmeanlearntoriginalvariancebayesianlarge
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
abstract

We resurrect the infamous harmonic mean estimator for computing the marginal likelihood (Bayesian evidence) and solve its problematic large variance. The marginal likelihood is a key component of Bayesian model selection to evaluate model posterior probabilities; however, its computation is challenging. The original harmonic mean estimator, first proposed by Newton and Raftery in 1994, involves computing the harmonic mean of the likelihood given samples from the posterior. It was immediately realised that the original estimator can fail catastrophically since its variance can become very large (possibly not finite). A number of variants of the harmonic mean estimator have been proposed to address this issue although none have proven fully satisfactory. We present the \emph{learnt harmonic mean estimator}, a variant of the original estimator that solves its large variance problem. This is achieved by interpreting the harmonic mean estimator as importance sampling and introducing a new target distribution. The new target distribution is learned to approximate the optimal but inaccessible target, while minimising the variance of the resulting estimator. Since the estimator requires samples of the posterior only, it is agnostic to the sampling strategy used. We validate the estimator on a variety of numerical experiments, including a number of pathological examples where the original harmonic mean estimator fails catastrophically. We also consider a cosmological application, where our approach leads to $\sim$ 3 to 6 times more samples than current state-of-the-art techniques in 1/3 of the time. In all cases our learnt harmonic mean estimator is shown to be highly accurate. The estimator is computationally scalable and can be applied to problems of dimension $O(10^3)$ and beyond. Code implementing the learnt harmonic mean estimator is made publicly available

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 4 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Self-consistent population synthesis of AGN from observational constraints in the X-rays

    astro-ph.CO 2025-06 conditional novelty 7.0 of 10

    A self-consistent population synthesis using ray-tracing and simulation-based inference simultaneously reproduces the cosmic X-ray background and AGN absorption statistics, yielding an intrinsic Compton-thick fraction...

  2. General Relativistic Entropic Acceleration at the perturbation level: a CLASS implementation and first Boltzmann-code constraints

    gr-qc 2026-07 conditional novelty 6.0 of 10

    First full Boltzmann-code implementation of entropic dark energy, with MCMC constraints from CMB+BAO+SN, yields α≈1 and a fit statistically indistinguishable from ΛCDM.

  3. Savage-Dickey density ratio estimation with normalizing flows for Bayesian model comparison

    astro-ph.CO 2025-06 conditional novelty 5.0 of 10

    A normalizing flow estimates the normalized marginal posterior in the Savage-Dickey density ratio, enabling Bayes factors for nested models with many extra parameters.

  4. Estimating Marginal Likelihoods in Likelihood-Free Inference via Neural Density Estimation

    stat.CO 2025-07 conditional novelty 4.0 of 10

    SNLE outputs can be used to estimate marginal likelihoods via importance sampling and a telescoping product of per-round likelihood ratios.

Pith tools