REVIEW 2 cited by
Stochastic Variational Optimization
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
Variational Optimization forms a differentiable upper bound on an objective. We show that approaches such as Natural Evolution Strategies and Gaussian Perturbation, are special cases of Variational Optimization in which the expectations are approximated by Gaussian sampling. These approaches are of particular interest because they are parallelizable. We calculate the approximate bias and variance of the corresponding gradient estimators and demonstrate that using antithetic sampling or a baseline is crucial to mitigate their problems. We contrast these methods with an alternative parallelizable method, namely Directional Derivatives. We conclude that, for differentiable objectives, using Directional Derivatives is preferable to using Variational Optimization to perform parallel Stochastic Gradient Descent.
Forward citations
Cited by 2 Pith papers
-
Neural Image Compression and Explanation
NICE trains a stochastic binary mask that marks decision-relevant pixels and turns the rest into a low-resolution background, giving both an explanation and about 1.6x PNG compression with a small accuracy drop.
-
Neural Plasticity Networks
A single parameter k in a binary-gate network training method interpolates between dropout, standard training, and sparse or expanded architectures.
Discussion (0). Continue with ORCID to comment.