REVIEW 3 cited by
Learning with minibatch Wasserstein : asymptotic and gradient properties
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
Optimal transport distances are powerful tools to compare probability distributions and have found many applications in machine learning. Yet their algorithmic complexity prevents their direct use on large scale datasets. To overcome this challenge, practitioners compute these distances on minibatches {\em i.e.} they average the outcome of several smaller optimal transport problems. We propose in this paper an analysis of this practice, which effects are not well understood so far. We notably argue that it is equivalent to an implicit regularization of the original problem, with appealing properties such as unbiased estimators, gradients and a concentration bound around the expectation, but also with defects such as loss of distance property. Along with this theoretical analysis, we also conduct empirical experiments on gradient flows, GANs or color transfer that highlight the practical interest of this strategy.
Forward citations
Cited by 3 Pith papers
-
Wasserstein normalized autoencoder for anomaly detection
A Wasserstein-distance-trained normalized autoencoder detects semivisible jets in simulated LHC events with AUCs around 0.69–0.77, outperforming standard and normalized autoencoders on a ttbar background.
-
Model alignment using inter-modal bridges
A semi-supervised conditional flow matching method, guided by a new inter-modal bridge cost, aligns latent spaces of distinct pre-trained models using only a small set of paired samples.
-
Modeling Stochastic Conditional Dynamics from Sparse Observations via Kernel-Stabilized Flow Matching
CVFM learns temporal evolution of conditional probability densities from unpaired state-condition observations by coupling state and conditioning flows with a mismatch kernel.
Discussion (0). Continue with ORCID to comment.