Pith. sign in

REVIEW 2 cited by

Stochastic Newton and Cubic Newton Methods with Simple Local Linear-Quadratic Rates

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 1912.01597 v1 pith:OHVGYAL7 submitted 2019-12-03 cs.LG math.OCstat.ML

classification cs.LGmath.OCstat.ML
keywords methodsstochasticnewtonconvergencelocalmethodexistingiteration
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

We present two new remarkably simple stochastic second-order methods for minimizing the average of a very large number of sufficiently smooth and strongly convex functions. The first is a stochastic variant of Newton's method (SN), and the second is a stochastic variant of cubically regularized Newton's method (SCN). We establish local linear-quadratic convergence results. Unlike existing stochastic variants of second order methods, which require the evaluation of a large number of gradients and/or Hessians in each iteration to guarantee convergence, our methods do not have this shortcoming. For instance, the simplest variants of our methods in each iteration need to compute the gradient and Hessian of a {\em single} randomly selected function only. In contrast to most existing stochastic Newton and quasi-Newton methods, our approach guarantees local convergence faster than with first-order oracle and adapts to the problem's curvature. Interestingly, our method is not unbiased, so our theory provides new intuition for designing new stochastic methods.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Physics-informed neural networks for high-dimensional solutions and snaking bifurcations in nonlinear lattices

    math.NA 2025-07 conditional novelty 5.0 of 10

    A PINN framework with Levenberg-Marquardt optimization approximates solutions, snaking bifurcation diagrams, and linear stability of the discrete Allen-Cahn equation in up to five spatial dimensions.

  2. A Unified Gradient-based Framework for Task-agnostic Continual Learning-Unlearning

    cs.LG 2025-05 conditional novelty 5.0 of 10

    UG-CLU derives a four-part gradient update for continual learning and unlearning and shows it outperforms task-level CLU methods on new fine-grained benchmarks.

Pith tools