Pith. sign in

REVIEW 1 cited by

Auto-Encoding Total Correlation Explanation

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 1802.05822 v1 pith:TWUA257M submitted 2018-02-16 cs.LG stat.ML

classification cs.LGstat.ML
keywords learningtotalboundcorexcorrelationgenerationinformationinformation-theoretic
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Advances in unsupervised learning enable reconstruction and generation of samples from complex distributions, but this success is marred by the inscrutability of the representations learned. We propose an information-theoretic approach to characterizing disentanglement and dependence in representation learning using multivariate mutual information, also called total correlation. The principle of total Cor-relation Ex-planation (CorEx) has motivated successful unsupervised learning applications across a variety of domains, but under some restrictive assumptions. Here we relax those restrictions by introducing a flexible variational lower bound to CorEx. Surprisingly, we find that this lower bound is equivalent to the one in variational autoencoders (VAE) under certain conditions. This information-theoretic view of VAE deepens our understanding of hierarchical VAE and motivates a new algorithm, AnchorVAE, that makes latent codes more interpretable through information maximization and enables generation of richer and more realistic samples.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Geometric Disentanglement for Generative Latent Shape Models

    cs.CV 2019-08 conditional novelty 6.0 of 10

    An unsupervised VAE for 3D point clouds is structured so that latent variables separately control intrinsic shape and extrinsic pose, using Laplace-Beltrami spectra and hierarchical disentanglement penalties.

Pith tools