REVIEW 3 cited by
Online Bootstrap Inference with Nonconvex Stochastic Gradient Descent Estimator
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
In this paper, we investigate the theoretical properties of stochastic gradient descent (SGD) for statistical inference in the context of nonconvex optimization problems, which have been relatively unexplored compared to convex settings. Our study is the first to establish provable inferential procedures using the SGD estimator for general nonconvex objective functions, which may contain multiple local minima. We propose two novel online inferential procedures that combine SGD and the multiplier bootstrap technique. The first procedure employs a consistent covariance matrix estimator, and we establish its error convergence rate. The second procedure approximates the limit distribution using bootstrap SGD estimators, yielding asymptotically valid bootstrap confidence intervals. We validate the effectiveness of both approaches through numerical experiments. Furthermore, our analysis yields an intermediate result: the in-expectation error convergence rate for the original SGD estimator in nonconvex settings, which is comparable to existing results for convex problems. We believe this novel finding holds independent interest and enriches the literature on optimization and statistical inference.
Forward citations
Cited by 3 Pith papers
-
Statistical Inference for Stochastic Gradient Descent: Beyond Finite Variance
Presents a self-normalized subsampling procedure for asymptotically valid confidence regions from SGD iterates under both finite and infinite variance assumptions.
-
Beyond Sin-Squared Error: Linear-Time Entrywise Uncertainty Quantification for Streaming PCA
Entrywise concentration bounds, a central limit theorem, and a median-of-means variance estimator give linear-time coordinate-wise uncertainty quantification for streaming PCA with Oja's algorithm.
-
Statistical inference for Linear Stochastic Approximation with Markovian Noise
Polyak-Ruppert averaged linear stochastic approximation with Markovian noise achieves Berry-Esseen rate O(n^{-1/4}) in Kolmogorov distance, and a multiplier subsample bootstrap achieves coverage error O(n^{-1/10}).
Discussion (0). Continue with ORCID to comment.