Nonasymptotic analysis of stochastic gradient descent with the richardson- romberg extrapolation

Nonasymptotic analysis of stochastic gradient descent with the richardson-romberg extrapolation , author= · 2024 · arXiv 2410.05106

3 Pith papers cite this work. Polarity classification is still indexing.

3 Pith papers citing it

read on arXiv browse 3 citing papers

representative citing papers

Shuffling the Data, Stretching the Step-size: Sharper Bias in constant step-size SGD

math.OC · 2026-04-11 · unverdicted · novelty 7.0

Combining random reshuffling and Richardson-Romberg extrapolation yields cubic bias refinement and better MSE for constant-step SGD on structured non-monotone variational inequalities.

SGD at the Edge of Stability: Stochastic Stabilization with Large Learning Rates

stat.ML · 2026-06-29 · unverdicted · novelty 6.0

SGD on multiclass cross-entropy loss alternates between curvature-driven oscillations and stable regimes but self-stabilizes to enable best-iterate convergence with large learning rates for linear and two-layer models.

Gaussian Approximation and Multiplier Bootstrap for Stochastic Gradient Descent

stat.ML · 2025-02-10 · unverdicted · novelty 6.0

Proves the first fully non-asymptotic bound on the accuracy of multiplier bootstrap for constructing confidence sets from Polyak-Ruppert SGD iterates, achieving convex-distance rates up to 1/sqrt(n) under regularity conditions.

citing papers explorer

Showing 3 of 3 citing papers after filters.

Shuffling the Data, Stretching the Step-size: Sharper Bias in constant step-size SGD math.OC · 2026-04-11 · unverdicted · none · ref 124
Combining random reshuffling and Richardson-Romberg extrapolation yields cubic bias refinement and better MSE for constant-step SGD on structured non-monotone variational inequalities.
SGD at the Edge of Stability: Stochastic Stabilization with Large Learning Rates stat.ML · 2026-06-29 · unverdicted · none · ref 215
SGD on multiclass cross-entropy loss alternates between curvature-driven oscillations and stable regimes but self-stabilizes to enable best-iterate convergence with large learning rates for linear and two-layer models.
Gaussian Approximation and Multiplier Bootstrap for Stochastic Gradient Descent stat.ML · 2025-02-10 · unverdicted · none · ref 41
Proves the first fully non-asymptotic bound on the accuracy of multiplier bootstrap for constructing confidence sets from Polyak-Ruppert SGD iterates, achieving convex-distance rates up to 1/sqrt(n) under regularity conditions.

Nonasymptotic analysis of stochastic gradient descent with the richardson- romberg extrapolation

fields

years

verdicts

representative citing papers

citing papers explorer