Pith. sign in

REVIEW 1 cited by

Semi-Amortized Variational Autoencoders

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 1802.02550 v7 pith:GBGPWRLA submitted 2018-02-07 stat.ML cs.CLcs.LG

classification stat.MLcs.CLcs.LG
keywords variationalinferenceapproachgenerativeautoencoderslocalmodelsnetwork
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Amortized variational inference (AVI) replaces instance-specific local inference with a global inference network. While AVI has enabled efficient training of deep generative models such as variational autoencoders (VAE), recent empirical work suggests that inference networks can produce suboptimal variational parameters. We propose a hybrid approach, to use AVI to initialize the variational parameters and run stochastic variational inference (SVI) to refine them. Crucially, the local SVI procedure is itself differentiable, so the inference network and generative model can be trained end-to-end with gradient-based optimization. This semi-amortized approach enables the use of rich generative models without experiencing the posterior-collapse phenomenon common in training VAEs for problems like text generation. Experiments show this approach outperforms strong autoregressive and variational baselines on standard text and image datasets.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Case Studies of Generative Machine Learning Models for Dynamical Systems

    eess.SY 2025-08 conditional novelty 5.0 of 10

    Physics-informed VAEs with Hamiltonian-based losses generate trajectories that match training distributions and satisfy optimal-control equations from as few as 200 to 500 samples.

Pith tools