Pith. sign in

REVIEW 2 cited by

Wasserstein Gradient Flow over Variational Parameter Space for Variational Inference

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2310.16705 v5 pith:WN6YQAJJ submitted 2023-10-25 cs.LG stat.ML

classification cs.LGstat.ML
keywords optimizationvariationalgradientdescentnatural-gradientwassersteinblack-boxinference
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Variational inference (VI) can be cast as an optimization problem in which the variational parameters are tuned to closely align a variational distribution with the true posterior. The optimization task can be approached through vanilla gradient descent in black-box VI or natural-gradient descent in natural-gradient VI. In this work, we reframe VI as the optimization of an objective that concerns probability distributions defined over a \textit{variational parameter space}. Subsequently, we propose Wasserstein gradient descent for tackling this optimization problem. Notably, the optimization techniques, namely black-box VI and natural-gradient VI, can be reinterpreted as specific instances of the proposed Wasserstein gradient descent. To enhance the efficiency of optimization, we develop practical methods for numerically solving the discrete gradient flows. We validate the effectiveness of the proposed methods through empirical experiments on a synthetic dataset, supplemented by theoretical analyses.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Accelerated Multiple Wasserstein Gradient Flows for Multi-objective Distributional Optimization

    cs.LG 2026-01 conditional novelty 6.0 of 10

    A-MWGraD accelerates multi-objective Wasserstein gradient descent, achieving O(1/t^2) and exponential merit-function convergence rates in continuous time for convex and strongly convex objectives.

  2. Multiple Wasserstein Gradient Descent Algorithm for Multi-Objective Distributional Optimization

    cs.LG 2025-05 conditional novelty 5.0 of 10

    MWGraD aggregates multiple Wasserstein gradients with dynamically updated weights to find Pareto-stationary distributions, with convergence guarantees and improved multi-task accuracy.

Pith tools