Pith. sign in

Paper Citation Record · LEDGER

Convergence Analysis of the Last Iterate in Distributed Stochastic Gradient Descent with Momentum

As of 19 August 2026, this Paper Citation Record lists 38 of 38 outbound references and 0 inbound Pith citation observations for arXiv:2505.10889.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.10889 v1

Coverage vector

measured 38 of 38 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T21:16:05.092207Z

measured 38 of 38 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

38 of 38 outbound references displayed

  • verified exact0
  • verified fuzzy30
  • unresolved8
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 1b5f86fc-1484-450f-b64d-78e01a15e9ff · outbound

This paper cites A stochastic approximation method,.

Convergence Analysis of the Last Iterate in Distributed Stochastic Gradient Descent with Momentum A stochastic approximation method,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-15T21:16:04.870210Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:16:04.870210Z digest=sha256:321870567d90eea49284f595d1c7adccc7c76e3406196c05213e0e22aaadb8d4

Observation 7585abd5-81d3-41f0-8ca3-fcf92f29d43f · outbound

This paper cites Some methods of speeding up the convergence of iteration methods,.

Convergence Analysis of the Last Iterate in Distributed Stochastic Gradient Descent with Momentum Some methods of speeding up the convergence of iteration methods,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:16:05.816393Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T21:16:04.875811Z digest=sha256:0e7d08f06ad4a6cb130e9926e4c7e1f0d2b5d91a12b8bddc71b580a134bd91a5

Observation b807a7cb-5537-4f4d-a557-2a99fad654c7 · outbound

This paper cites Adaptive deep feature learning network with Nesterov momentum and its application to rotating machinery fault diagnosis,.

Convergence Analysis of the Last Iterate in Distributed Stochastic Gradient Descent with Momentum Adaptive deep feature learning network with Nesterov momentum and its application to rotating machinery fault diagnosis,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:16:05.799392Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T21:16:04.881935Z digest=sha256:a8350805ea9824f2127087b8aad8855e3c556b523b00173c13ca270696b3e4dd

Observation 9e5d4f43-4c8a-4c92-a6fc-8b05117e671b · outbound

This paper cites Combining ordered subsets and momentum for accelerated X-ray CT image reconstruction,.

Convergence Analysis of the Last Iterate in Distributed Stochastic Gradient Descent with Momentum Combining ordered subsets and momentum for accelerated X-ray CT image reconstruction,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:16:05.783110Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T21:16:04.887847Z digest=sha256:415bd7f91588833dd304ff4ffc148a53c62042c6d01fc449e8b81dd0c58a5872

Observation d25caf16-974f-4e6e-9b89-42683cde325b · outbound

This paper cites Speech recognition with deep recurrent neural networks,.

Convergence Analysis of the Last Iterate in Distributed Stochastic Gradient Descent with Momentum Speech recognition with deep recurrent neural networks,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:16:05.764791Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T21:16:04.900026Z digest=sha256:ee368f5825c985fa12270a7315aa2759fbc2e41592c0a1467b6fb5f2b3d34635

Observation 2311f3d8-5143-4f18-ac09-c0031f7e3d24 · outbound

This paper cites SGD and Hogwild! convergence without the bounded gradients assumption,.

Convergence Analysis of the Last Iterate in Distributed Stochastic Gradient Descent with Momentum SGD and Hogwild! convergence without the bounded gradients assumption,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:16:05.745177Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T21:16:04.909718Z digest=sha256:73ffcf827155f5a54495bc748a089e5843cedf8664090e648511e76619656066

Observation 61dccd6c-5dc3-4985-a3e2-482351608cf1 · outbound

This paper cites Reducing the dimensionality of data with neural networks,.

Convergence Analysis of the Last Iterate in Distributed Stochastic Gradient Descent with Momentum Reducing the dimensionality of data with neural networks,

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-15T21:16:04.917004Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:16:04.917004Z digest=sha256:3287c96529ea7c872c4d077e643a0b6949d9edc591d8b1c4df002fac1655b077

Observation 28293377-7842-44e8-bbee-1a415bf18e64 · outbound

This paper cites Distributed training strategies for the structured perceptron,.

Convergence Analysis of the Last Iterate in Distributed Stochastic Gradient Descent with Momentum Distributed training strategies for the structured perceptron,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:16:05.712772Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T21:16:04.923055Z digest=sha256:d48d7ce8f3af5e6d3440bc455466ad215bbb4564468da016cc6ed87572e2d90f

Observation 09b2f917-3779-4625-9eea-956a35f780f4 · outbound

This paper cites Deep learning with elastic averaging sgd,.

Convergence Analysis of the Last Iterate in Distributed Stochastic Gradient Descent with Momentum Deep learning with elastic averaging sgd,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:16:05.694937Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T21:16:04.928793Z digest=sha256:ed4b310530025debee5b2ef9620f2c95574c1304d990bc3dd99a2fbf9debbcb0

Observation be118f36-a1fd-4cd2-adfc-7a1d16366ecf · outbound

This paper cites Network topology and communication-computation tradeoffs in decentralized optimization,.

Convergence Analysis of the Last Iterate in Distributed Stochastic Gradient Descent with Momentum Network topology and communication-computation tradeoffs in decentralized optimization,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:16:05.677857Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T21:16:04.934245Z digest=sha256:4ae866cb03c083d1b1d8ae8f2d7037884cbe0e8d59e02eb23ebaf76f7d5862d0

Observation 8ba79a71-f9ed-4ec2-b28a-5e96becd7255 · outbound

This paper cites Local SGD Converges Fast and Communicates Little.

Convergence Analysis of the Last Iterate in Distributed Stochastic Gradient Descent with Momentum Local SGD Converges Fast and Communicates Little

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-15T21:16:04.939109Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:16:04.939109Z digest=sha256:d37a35f09770eb95292617f3f21b6a7aeb834ac062034773fef29a4bdfdfdbea

Observation 3c49b6d1-4c9e-4dce-994f-4bbc64568414 · outbound

This paper cites Parallel restarted sgd with faster convergence and less communication: Demystifying why model averaging works for deep learning,.

Convergence Analysis of the Last Iterate in Distributed Stochastic Gradient Descent with Momentum Parallel restarted sgd with faster convergence and less communication: Demystifying why model averaging works for deep learning,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:16:05.658384Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T21:16:04.945190Z digest=sha256:cc801b4a1f5a511cd7415bcb16f803c3fe4c07df185606c1696104e47b3e9305

Observation e04901f1-2fae-44e9-86a5-c19c7cef8610 · outbound

This paper cites A linear speedup analysis of distributed deep learning with sparse and quantized communication,.

Convergence Analysis of the Last Iterate in Distributed Stochastic Gradient Descent with Momentum A linear speedup analysis of distributed deep learning with sparse and quantized communication,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:16:05.638473Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T21:16:04.950550Z digest=sha256:a97b4c913aac1e7f8b3002e11a58445042a291cde253099909bd5bb955fab534

Observation d7cc35b9-7ac1-4333-b872-59d8a830cf17 · outbound

This paper cites Can decentralized algorithms outperform centralized algorithms? a case study for decentralized parallel stochastic gradient descent,.

Convergence Analysis of the Last Iterate in Distributed Stochastic Gradient Descent with Momentum Can decentralized algorithms outperform centralized algorithms? a case study for decentralized parallel stochastic gradient descent,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-15T21:16:04.959122Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:16:04.959122Z digest=sha256:b3beede8eeaa3e694ecc614338c3c34bd2be2e6b6893dab776895186f217a910

Observation bd5fc640-7e5a-4828-8cea-377fc323f43a · outbound

This paper cites Collaborative deep learning in fixed topology networks,.

Convergence Analysis of the Last Iterate in Distributed Stochastic Gradient Descent with Momentum Collaborative deep learning in fixed topology networks,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:16:05.609126Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T21:16:04.965081Z digest=sha256:0cf2be148ac4d6950ed8196c3d279443d58f763eba8617ce8c58c1302b45cad8

Observation 54d6075a-44b8-4aa7-89e9-dae226ecf6aa · outbound

This paper cites On nonconvex decentralized gradient descent,.

Convergence Analysis of the Last Iterate in Distributed Stochastic Gradient Descent with Momentum On nonconvex decentralized gradient descent,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-15T21:16:04.970086Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:16:04.970086Z digest=sha256:4537a937029d2c737b8924478db30e52fdc94d20de8bd00081e74bbc04fba06b

Observation 21478628-8e38-4c4f-bc5a-787d95886490 · outbound

This paper cites Cooperative sgd: A unified framework for the design and analysis of local-update sgd algorithms,.

Convergence Analysis of the Last Iterate in Distributed Stochastic Gradient Descent with Momentum Cooperative sgd: A unified framework for the design and analysis of local-update sgd algorithms,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:16:05.578632Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T21:16:04.974941Z digest=sha256:8421ad57768cd6e98a800b6d93d46ec00bd08ffb7a0aea31600b3faf81e78431

Observation 5968c1dd-000c-4260-85ef-567626f646cd · outbound

This paper cites Imagenet classification with deep convolutional neural networks,.

Convergence Analysis of the Last Iterate in Distributed Stochastic Gradient Descent with Momentum Imagenet classification with deep convolutional neural networks,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:16:05.560596Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T21:16:04.981179Z digest=sha256:5580591262b479811804785f7d2f8c12a14797105dd1f45a9c2ae97c50e24c29

Observation 063a5a98-a5ae-47e7-bb83-de7c48a3e4b8 · outbound

This paper cites A Unified Analysis of Stochastic Momentum Methods for Deep Learning.

Convergence Analysis of the Last Iterate in Distributed Stochastic Gradient Descent with Momentum A Unified Analysis of Stochastic Momentum Methods for Deep Learning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-15T21:16:04.987777Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:16:04.987777Z digest=sha256:c1e65ee0d78d54212161d37d551d1306cd9a1890d0a491ea741eaf8ee91cfcd3

Observation 22d8a116-b5a8-472d-827d-3cd5064a03e6 · outbound

This paper cites On the importance of initialization and momentum in deep learning,.

Convergence Analysis of the Last Iterate in Distributed Stochastic Gradient Descent with Momentum On the importance of initialization and momentum in deep learning,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:16:05.542612Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T21:16:04.995072Z digest=sha256:9d13a1f16dae5490c0db11e6e8cacdb9a1f160eed205a238d46ac21809a82716

Observation d3c4ab34-6398-4381-a1d0-c753a7476d50 · outbound

This paper cites On the linear speedup analysis of communication efficient momentum sgd for distributed non- convex optimization,.

Convergence Analysis of the Last Iterate in Distributed Stochastic Gradient Descent with Momentum On the linear speedup analysis of communication efficient momentum sgd for distributed non- convex optimization,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:16:05.524880Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T21:16:05.001300Z digest=sha256:e5e397fdbf6d89c7a79de84aaffc0cbd51b4fd484818dd3f285dc3e6ea106405

Observation 615cb8d7-1509-4876-838e-b19ce8728ba0 · outbound

This paper cites On consensus- optimality trade-offs in collaborative deep learning,.

Convergence Analysis of the Last Iterate in Distributed Stochastic Gradient Descent with Momentum On consensus- optimality trade-offs in collaborative deep learning,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:16:05.507346Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T21:16:05.006388Z digest=sha256:0511b8f6984c3581aad81715832c35c59ecfdd515cef5ad94c904bc077bb8c5b

Observation 3be9fce3-33cf-4350-af62-f99cc5b9cbdc · outbound

This paper cites Deep gradient compression: Reducing the communication bandwidth for distributed training,.

Convergence Analysis of the Last Iterate in Distributed Stochastic Gradient Descent with Momentum Deep gradient compression: Reducing the communication bandwidth for distributed training,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:16:05.484749Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T21:16:05.011429Z digest=sha256:7003c72791a9624c419a0a7f3849635e3b0939a647aec387e0a9ccb96c13323b

Observation 351bc966-cd6c-4988-b96b-29baf7d61ca0 · outbound

This paper cites Can decentralized algorithms outperform centralized algorithms? a case study for decentralized parallel stochastic gradient descent,.

Convergence Analysis of the Last Iterate in Distributed Stochastic Gradient Descent with Momentum Can decentralized algorithms outperform centralized algorithms? a case study for decentralized parallel stochastic gradient descent,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:16:05.463015Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T21:16:05.016562Z digest=sha256:c8de423d80c96be16b32da6fde092d078db7be3d2b6d595a9cae2092b35d8834

Observation 2af36f5a-24eb-46ea-b0fa-ca8a190c034b · outbound

This paper cites Decentlam: Decentralized momentum sgd for large-batch deep training,.

Convergence Analysis of the Last Iterate in Distributed Stochastic Gradient Descent with Momentum Decentlam: Decentralized momentum sgd for large-batch deep training,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:16:05.445325Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T21:16:05.022135Z digest=sha256:3d47e8e206c17b2be8e87164cfc355c5c85faa7647e72f918f48073bf01fefb1

Observation 800f32fa-3c7f-41c4-a694-f859fd295700 · outbound

This paper cites 2020SQuARM: Communication-efficient momentum SGD for decentralized optimization,.

Convergence Analysis of the Last Iterate in Distributed Stochastic Gradient Descent with Momentum 2020SQuARM: Communication-efficient momentum SGD for decentralized optimization,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:16:05.422343Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T21:16:05.026972Z digest=sha256:9969f2e4aa40cc5b8278de56bc05ab06db22d49f384096c5afee26c755f5cf33

Observation 03643d56-567b-42d2-bde3-7f2e2674825f · outbound

This paper cites Periodic Stochastic Gradient Descent with Momentum for Decentralized Training.

Convergence Analysis of the Last Iterate in Distributed Stochastic Gradient Descent with Momentum Periodic Stochastic Gradient Descent with Momentum for Decentralized Training

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-15T21:16:05.031704Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:16:05.031704Z digest=sha256:4a457724b5e93c33b2348b5e2a54c01a5425375578c220d7f3065722bd577c89

Observation e6b9600d-3953-418e-a5ae-5b625c772bdc · outbound

This paper cites Decentralized deep learning using momentum-accelerated consensus,.

Convergence Analysis of the Last Iterate in Distributed Stochastic Gradient Descent with Momentum Decentralized deep learning using momentum-accelerated consensus,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:16:05.403893Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T21:16:05.037280Z digest=sha256:6981a4546a49391fa5019d0d9e21fa1e4478c358d60ee09939835ca0c32735c7

Observation 2039e531-6c49-4595-b67b-0fedeebd52be · outbound

This paper cites Parallel restarted sgd with faster convergence and less communication: Demystifying why model averaging works for deep learning,.

Convergence Analysis of the Last Iterate in Distributed Stochastic Gradient Descent with Momentum Parallel restarted sgd with faster convergence and less communication: Demystifying why model averaging works for deep learning,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:16:05.384740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T21:16:05.042769Z digest=sha256:bc54fa5e393bc895ede8ffc929ec5ab862585c3ccbd0930203609c29294ae5e1

Observation 58a2e080-9d1c-455e-84a1-8ad434aad158 · outbound

This paper cites On the convergence of mSGD and AdaGrad for stochastic optimization,.

Convergence Analysis of the Last Iterate in Distributed Stochastic Gradient Descent with Momentum On the convergence of mSGD and AdaGrad for stochastic optimization,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:16:05.365253Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T21:16:05.047649Z digest=sha256:8b1dd473930d36bfd95d93faf5541d421f305c4f60eae65886cb0313a6db3acb

Observation 4cbe6bf7-67b3-45b1-bdbf-c4747bbdeefc · outbound

This paper cites Don't Decay the Learning Rate, Increase the Batch Size.

Convergence Analysis of the Last Iterate in Distributed Stochastic Gradient Descent with Momentum Don't Decay the Learning Rate, Increase the Batch Size

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-15T21:16:05.053234Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:16:05.053234Z digest=sha256:7f46ff827cc2e9ec6af34a82ee3f2dd246f344c922b8512672371c497f03cf1e

Observation 696e5d03-2482-4598-9f3f-27ea1dba391d · outbound

This paper cites Bayesian learning via stochastic gradient langevin dynamics,.

Convergence Analysis of the Last Iterate in Distributed Stochastic Gradient Descent with Momentum Bayesian learning via stochastic gradient langevin dynamics,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:16:05.346931Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T21:16:05.059295Z digest=sha256:4535ac6fe43bcccfd107e7da67b9fafa960c3bf606007c838786317fb744de31

Observation cfc784d5-d69d-4e7c-9cc0-f008fbdc5273 · outbound

This paper cites Convergence of proximal-gradient stochastic variational inference under non-decreasing step-size sequence,.

Convergence Analysis of the Last Iterate in Distributed Stochastic Gradient Descent with Momentum Convergence of proximal-gradient stochastic variational inference under non-decreasing step-size sequence,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:16:05.328386Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T21:16:05.065059Z digest=sha256:de59fee11282591d19b9453e7f3e3014793b69bfaa02812169d404c4386d46b3

Observation d656d89c-cd16-4138-b152-92cb33613746 · outbound

This paper cites Understanding the role of momentum in stochastic gradient methods,.

Convergence Analysis of the Last Iterate in Distributed Stochastic Gradient Descent with Momentum Understanding the role of momentum in stochastic gradient methods,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:16:05.308628Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T21:16:05.070661Z digest=sha256:d61a00f4b19c7fe625be522864a6d84a762f64258d6b88a1464637ad87d1727a

Observation a5415d8d-8810-476f-b99f-1868005ba493 · outbound

This paper cites Deep residual learning for image recognition,.

Convergence Analysis of the Last Iterate in Distributed Stochastic Gradient Descent with Momentum Deep residual learning for image recognition,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:16:05.288603Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T21:16:05.075459Z digest=sha256:ea46d3f5f7c2744fa3f941aa79f88fb7e6309d766b99234f087fdef8a2a99e78

Observation 0214823a-791f-4ca3-93aa-1578f2ad0091 · outbound

This paper cites Nesterov,Introductory Lectures on Convex Optimization: A Basic Course.

Convergence Analysis of the Last Iterate in Distributed Stochastic Gradient Descent with Momentum Nesterov,Introductory Lectures on Convex Optimization: A Basic Course

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:16:05.266884Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T21:16:05.081368Z digest=sha256:b510c4ea82552b8f74c33f038cb32090a84776724b7827dfe1de17ee206852bb

Observation e8d7a3ba-bd4b-439c-8769-b5e6561f3fe8 · outbound

This paper cites Revisit last- iterate convergence of msgd under milder requirement on step size.

Convergence Analysis of the Last Iterate in Distributed Stochastic Gradient Descent with Momentum Revisit last- iterate convergence of msgd under milder requirement on step size

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:16:05.245733Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T21:16:05.086621Z digest=sha256:3259e3aa94648500a8eb9d93290abd87fc0de6810137473cd1954652a898f988

Observation 4d126e17-d5ef-42e8-9b5e-e2e73f0ad451 · outbound

This paper cites On almost sure convergence for sums of stochastic sequence,.

Convergence Analysis of the Last Iterate in Distributed Stochastic Gradient Descent with Momentum On almost sure convergence for sums of stochastic sequence,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:16:05.225176Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T21:16:05.092207Z digest=sha256:c2955a2f3080704649fe02306f424fae625981088745a950a5ed3959a4314f0e

Pith citing papers

No inbound Pith citation observations are available.