REVIEW 4 cited by
posteriordb: Testing, Benchmarking and Developing Bayesian Inference Algorithms
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
The generality and robustness of inference algorithms is critical to the success of widely used probabilistic programming languages such as Stan, PyMC, Pyro, and Turing.jl. When designing a new general-purpose inference algorithm, whether it involves Monte Carlo sampling or variational approximation, the fundamental problem arises in evaluating its accuracy and efficiency across a range of representative target models. To solve this problem, we propose posteriordb, a database of models and data sets defining target densities along with reference Monte Carlo draws. We further provide a guide to the best practices in using posteriordb for model evaluation and comparison. To provide a wide range of realistic target densities, posteriordb currently comprises 120 representative models and has been instrumental in developing several general inference algorithms.
Forward citations
Cited by 4 Pith papers
-
Is Retrieval All You Need? Assessment and Emergence of Novelty in Protein Structure Generation
Full-chain structural novelty is not evidence of fold invention, because generated backbones mostly contain known domains and a zero-training retrieval baseline reproduces the same novelty profile.
-
Diffeomorphic Markov Chain Monte Carlo: fast mixing for heavy-tailed distributions
DCS pulls heavy-tailed targets back to a Euclidean ball and proves uniform ergodicity for any polynomial tail, with O(d^3) Hit-and-Run mixing for Student-t targets.
-
MCBench: A Benchmark Suite for Monte Carlo Sampling Algorithms
MCBench is a modular Julia benchmark suite that scores Monte Carlo samplers by comparing their samples against IID reference samples using metrics like sliced Wasserstein distance and maximum mean discrepancy.
-
Disentangling impact of capacity, objective, batchsize, estimators, and step-size on flow VI
A careful ablation shows normalizing-flow variational inference with large capacity and large batchsize matches turnkey HMC, so complex objectives and estimators are unnecessary.
Discussion (0). Continue with ORCID to comment.