REVIEW 2 cited by
Wasserstein PAC-Bayes Learning: Exploiting Optimisation Guarantees to Explain Generalisation
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Wasserstein PAC-Bayes Learning: Exploiting Optimisation Guarantees to Explain Generalisation
read the original abstract
PAC-Bayes learning is an established framework to both assess the generalisation ability of learning algorithms, and design new learning algorithm by exploiting generalisation bounds as training objectives. Most of the exisiting bounds involve a \emph{Kullback-Leibler} (KL) divergence, which fails to capture the geometric properties of the loss function which are often useful in optimisation. We address this by extending the emerging \emph{Wasserstein PAC-Bayes} theory. We develop new PAC-Bayes bounds with Wasserstein distances replacing the usual KL, and demonstrate that sound optimisation guarantees translate to good generalisation abilities. In particular we provide generalisation bounds for the \emph{Bures-Wasserstein SGD} by exploiting its optimisation properties.
Forward citations
Cited by 2 Pith papers
-
Benign Overfitting Does Not Occur in Diffusion Models
Benign overfitting and double descent do not occur in diffusion models: population and empirical score-matching losses cannot both be small without exponentially many samples.
-
PAC-Bayesian Reinforcement Learning Trains Generalizable Policies
A mixing-time-aware PAC-Bayes bound is turned into PB-SAC, an algorithm that computes tightening certified performance lower bounds during SAC training on MuJoCo tasks.
discussion (0)
Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.