pith. sign in

arxiv: 1607.07329 · v1 · pith:EPNLAVKAnew · submitted 2016-07-25 · 🧮 math.OC · stat.ML

Accelerating Stochastic Composition Optimization

classification 🧮 math.OC stat.ML
keywords stochasticasc-pgcompositionmethodgradientoptimizationproblemproximal
0
0 comments X
read the original abstract

Consider the stochastic composition optimization problem where the objective is a composition of two expected-value functions. We propose a new stochastic first-order method, namely the accelerated stochastic compositional proximal gradient (ASC-PG) method, which updates based on queries to the sampling oracle using two different timescales. The ASC-PG is the first proximal gradient method for the stochastic composition problem that can deal with nonsmooth regularization penalty. We show that the ASC-PG exhibits faster convergence than the best known algorithms, and that it achieves the optimal sample-error complexity in several important special cases. We further demonstrate the application of ASC-PG to reinforcement learning and conduct numerical experiments.

This paper has not been read by Pith yet.

discussion (0)

Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.