Pith. sign in

Paper Citation Record · LEDGER

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning

As of 14 August 2026, this Paper Citation Record lists 55 of 55 outbound references and 2 inbound Pith citation observations for arXiv:2505.22085.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.22085 v1

Coverage vector

measured 55 of 55 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:20:49.832951Z

measured 57 of 57 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-03T10:09:36.912553Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T07:29:38.373231Z

Reference resolution

55 of 55 outbound references displayed

  • verified exact4
  • verified fuzzy25
  • unresolved26
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 0ede2bde-50d0-4e6b-b914-b26768d3a5f4 · outbound

This paper cites Adam with model exponential moving average is effective for nonconvex optimization.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning Adam with model exponential moving average is effective for nonconvex optimization

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T13:20:44.027164Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:20:44.027164Z digest=sha256:36515a634cf34ff50b36bed8f772ab4c483789d00f33bde1559320d09edd68ac

Observation ce1f34b2-c3bc-4452-a082-0e5e2cfea8f5 · outbound

This paper cites General framework for online-to-nonconvex conversion: Schedule-free SGD is also effective for nonconvex optimization.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning General framework for online-to-nonconvex conversion: Schedule-free SGD is also effective for nonconvex optimization

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T13:20:44.103140Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:20:44.103140Z digest=sha256:136cb5d35383a46761b2aa811d4c257d27ca451bc52169b46a8d1c04cc940e39

Observation dbc51ba4-6467-4b70-b652-51e18dec4760 · outbound

This paper cites an unresolved cited work.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:20:59.424343Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T13:20:44.251006Z digest=sha256:55e642894aab377afd95a02b38aa6f93cd873849bb53d22fda515d65f6777a1c

Observation f655e2b0-de84-4ab0-aca7-6b406d49b244 · outbound

This paper cites Learning Theory from First Principles.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning Learning Theory from First Principles

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:20:59.101490Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T13:20:44.335953Z digest=sha256:0d864e276d1632e011894fb0f1af5cc9ddae1c7ba3b6f9f6b0101bd2526b3355

Observation 003918ea-9e35-4d1e-ac82-3ef4f2fff3e5 · outbound

This paper cites Convergence and dynamical behavior of the Adam algorithm for nonconvex stochastic optimization.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning Convergence and dynamical behavior of the Adam algorithm for nonconvex stochastic optimization

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:20:58.683762Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T13:20:44.420377Z digest=sha256:cbd92fc011d4afea011bb7d3772201d5f9696dd8e7c30fdc3eefe46f3c959545

Observation 66f837b2-360b-4c7a-9b98-1a88ab5c6eb6 · outbound

This paper cites Solving the Kolmogorov PDE by means of deep learning.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning Solving the Kolmogorov PDE by means of deep learning

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:20:58.425929Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T13:20:44.520870Z digest=sha256:7a9f4c917dd4a009dbfab364ff723bba2c880ff50b519d30e30e2e798d2ddb7d

Observation 603fecab-d84e-4768-8b14-0f59c951ab3d · outbound

This paper cites An overview on deep learning-based approximation methods for partial differential equations.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning An overview on deep learning-based approximation methods for partial differential equations

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:20:58.111151Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T13:20:44.587912Z digest=sha256:b73340aaab5d9f9fe2fce8a122c0a8debb6cf2d2fe2ddd6f756bf90317d94c7b

Observation a9c69ac5-8cb3-460f-9215-91f249d6b530 · outbound

This paper cites Solving high-dimensional optimal stopping problems using deep learning.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning Solving high-dimensional optimal stopping problems using deep learning

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:20:57.837111Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T13:20:44.675429Z digest=sha256:2fd2f5495869ea176ded8011f9d4967ccd7dd254b07dc9e206b1dd03fda1c2dc

Observation fb6639ea-405b-44e7-a370-97f4dd87b3ba · outbound

This paper cites an unresolved cited work.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:20:57.565534Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T13:20:44.774332Z digest=sha256:1080839ad91607cd55bea9f82612e56765f89a0b0aa0cd694b0106c79c49888d

Observation 10695ec8-d505-4f1c-ab3c-475bc9ce6d2b · outbound

This paper cites G., Suau Cuadros, X., and Webb, R.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning G., Suau Cuadros, X., and Webb, R

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:20:57.266491Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T13:20:44.911287Z digest=sha256:58084fcded5e956ddfe6028ecbce2dca55d3a4a526addb47b55123e4de95a059

Observation 0cf9b7d8-f3cb-4c5b-b55e-232ad0c71007 · outbound

This paper cites Scientific machine learning through physics-informed neural networks: where we are and what’s next.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning Scientific machine learning through physics-informed neural networks: where we are and what’s next

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:20:56.991187Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T13:20:45.018646Z digest=sha256:99163173b2d40d0d224b435beaea78a77ac45e45beed366db100948e339a0fca

Observation a8dddf37-23af-43f4-b8ac-005545f927e8 · outbound

This paper cites The Road Less Scheduled.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning The Road Less Scheduled

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T13:20:45.123106Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:20:45.123106Z digest=sha256:77b1af10678ebb2158add4507be99fd5b596a8d1a8a9a6a96d3be5d8936b0fc5

Observation 51a48e79-e6d9-435a-9995-498bf0e27a9b · outbound

This paper cites A Simple Convergence Proof of Adam and Adagrad.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning A Simple Convergence Proof of Adam and Adagrad

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:20:56.670464Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T13:20:45.226863Z digest=sha256:88cf2355ad334284ce89d2a851b9a9963372bcd9a7e988716392b5601602ecf9

Observation c3cb9ab0-2fc9-45ed-ad59-5f53b7904919 · outbound

This paper cites General multilevel adaptations for stochastic approximation algorithms II: CLTs.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning General multilevel adaptations for stochastic approximation algorithms II: CLTs

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:20:56.382981Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T13:20:45.364610Z digest=sha256:9da3e0f897d79b13bc91082de8c38ce9c59b99685c04c896ddc56d52f40fa086

Observation 7561269f-221b-42ab-b3e7-386f24565866 · outbound

This paper cites Convergence rates for the Adam optimizer.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning Convergence rates for the Adam optimizer

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T13:20:45.494765Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:20:45.494765Z digest=sha256:66cde64c0c8382e013f4b8be81b2a05a2282cc05e0283d9309ffd86bbeadf52a

Observation 48240a19-5a1f-4a6d-9ae4-a7e6e0a34b9d · outbound

This paper cites On the existence of minimizers in shallow residual relu neural network optimization landscapes.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning On the existence of minimizers in shallow residual relu neural network optimization landscapes

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:20:56.058509Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T13:20:45.533294Z digest=sha256:e56c4658709f2b4d2d465ba2382309ff1ced1e75db67d4aa0b0e95af0692eecc

Observation 87d7e869-f19e-427e-8c0a-7b832eba50bc · outbound

This paper cites Averaged Adam accelerates stochastic optimization in the training of deep neural network approximations for partial differential equation and optimal control problems.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning Averaged Adam accelerates stochastic optimization in the training of deep neural network approximations for partial differential equation and optimal control problems

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T13:20:45.699370Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:20:45.699370Z digest=sha256:98ec97dd59553f2a3805cb4949c4e598e1423f53a336f9f4bbdac9b01694f2a1

Observation 080081b6-b3f0-4147-9f37-997a168118e0 · outbound

This paper cites Central limit theorems for stochastic gradient descent with averaging for stable manifolds.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning Central limit theorems for stochastic gradient descent with averaging for stable manifolds

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:20:55.827850Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T13:20:45.799000Z digest=sha256:c5c8bd1a273346789119117f047b11c3d645e350949a818e647856c27a9a7ad1

Observation e29d8abd-3823-40d6-97db-4159f880da90 · outbound

This paper cites On the existence of optimal shallow feedforward networks with ReLU activation.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning On the existence of optimal shallow feedforward networks with ReLU activation

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:20:55.531121Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T13:20:45.895036Z digest=sha256:90ba066f6b699ca419a407bd4b5ad5f229f9935772d8fd71e8f11bb946dc3973

Observation a998fdc4-adee-4388-aae0-cab3df4962ea · outbound

This paper cites General multilevel adaptations for stochastic approximation algorithms of Robbins-Monro and Polyak-Ruppert type.Numer.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning General multilevel adaptations for stochastic approximation algorithms of Robbins-Monro and Polyak-Ruppert type.Numer

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:20:55.228206Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T13:20:46.038818Z digest=sha256:2d3c13012a09037d8815309e28510f9d27f7fa554b11664acad854788594cc22

Observation 7765285a-e5f8-44e7-8862-38d86e513211 · outbound

This paper cites Uniform convergence guarantees for the deep ritz method for nonlinear problems.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning Uniform convergence guarantees for the deep ritz method for nonlinear problems

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:20:54.928881Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T13:20:46.144661Z digest=sha256:e1b0d1d7b41aad20975fdbb4af4ca8f2f0c54f000f2783eb484c511247cb0b4f

Observation 2e25f9d3-d276-4245-8d69-89a207cbf049 · outbound

This paper cites Deep learning-based numerical methods for high- dimensional parabolic partial differential equations and backward stochastic differential equations.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning Deep learning-based numerical methods for high- dimensional parabolic partial differential equations and backward stochastic differential equations

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:20:54.610718Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T13:20:46.247445Z digest=sha256:5f9e1e1f815db3e3b478251063c2c27c2b2f9fdaf6caacd492656871d9b343aa

Observation dd51ade3-a143-42ee-869d-c433c46f5dce · outbound

This paper cites Algorithms for solving high dimensional PDEs: from nonlinear Monte Carlo to machine learning.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning Algorithms for solving high dimensional PDEs: from nonlinear Monte Carlo to machine learning

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:20:54.327060Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T13:20:46.383740Z digest=sha256:797e202b319a50e6ceb62a6b6bc39573ad9866b5aac22cbfd79350037afcec09

Observation e4bdcae6-5852-49e7-afb1-ee27a3958610 · outbound

This paper cites The deep Ritz method: a deep learning-based numerical algorithm for solving variational problems.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning The deep Ritz method: a deep learning-based numerical algorithm for solving variational problems

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:20:54.005662Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T13:20:46.481643Z digest=sha256:8b674e415f0c2c0890181cfd79a3780439d9209ae963a99f928659c93a76103b

Observation 91ddcca1-97a5-4f04-9fca-01d5749814f3 · outbound

This paper cites Optimal non-asymptotic bound of the Ruppert-Polyak averaging without strong convexity.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning Optimal non-asymptotic bound of the Ruppert-Polyak averaging without strong convexity

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-08-07T13:20:51.292006Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T13:20:46.576139Z digest=sha256:2bb4c04b3392326c124684a7c7393b42ce0063977c163a6017fdd3d36deaf489

Observation 3df7c9c6-1abd-410a-bfde-b0a24629b2e4 · outbound

This paper cites Blow up phenomena for gradient descent optimization methods in the training of artificial neural networks.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning Blow up phenomena for gradient descent optimization methods in the training of artificial neural networks

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T13:20:46.654207Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:20:46.654207Z digest=sha256:3310bac4b882e28a5cf31e453823afdeb16993f77aa4cc34dd95473222c2eb10

Observation 40721e7e-24dd-4a0c-856e-320d5fdbda52 · outbound

This paper cites Neural networks-based algorithms for stochastic control and PDEs in finance.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning Neural networks-based algorithms for stochastic control and PDEs in finance

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T13:20:46.707687Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:20:46.707687Z digest=sha256:836e94020cdab83248ce5381efeaa0763334d6d296818ce1c3894b604ea20234

Observation 911bee71-a7b2-49a9-87b7-2c8b01730968 · outbound

This paper cites Stochastic weight averaging revisited.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning Stochastic weight averaging revisited

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:20:53.734446Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T13:20:46.832622Z digest=sha256:e025af187ec61a1b0a9bfd3ccf0043922f2a7e61457edac4c0e2321117062780

Observation 74dad194-4c96-4a2f-844e-990a800d4b01 · outbound

This paper cites an unresolved cited work.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:20:53.453359Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T13:20:46.957269Z digest=sha256:2289e787957420362fed2dfc8d0ead41c08356203fc3b71d10073596b609e727

Observation dd6775d0-2b8e-4823-ac1a-76cdd95391a3 · outbound

This paper cites Recent developments in machine learning methods for stochastic control and games.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning Recent developments in machine learning methods for stochastic control and games

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:20:53.089257Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T13:20:47.034083Z digest=sha256:6a9ed8d2bc3e82bb3177f62f0a7766d579a499c6637b9ce45aaa788ed7bbf91f

Observation f24dec9c-4fdd-40dd-a4a8-135a1fe8edcd · outbound

This paper cites Averaging Weights Leads to Wider Optima and Better Generalization.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning Averaging Weights Leads to Wider Optima and Better Generalization

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T13:20:47.187971Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:20:47.187971Z digest=sha256:ed1095478d4ea1f3831b16713f291d418e7b9cabd2334bb8acf9039a033ba841

Observation 8916bea2-64ea-48d1-b5c0-54fa55f3d09e · outbound

This paper cites Mathematical Introduction to Deep Learning: Methods, Implementations, and Theory.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning Mathematical Introduction to Deep Learning: Methods, Implementations, and Theory

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T13:20:47.254649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:20:47.254649Z digest=sha256:90b601878cafd219cd6adb94e415fbe3817d1a8e3b78f78208b784ca17dfee8d

Observation e75e8d49-a8c5-49ac-91b3-de4e0cc451f8 · outbound

This paper cites On the existence of global minima and convergence analyses for gradient descent methods in the training of deep neural networks.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning On the existence of global minima and convergence analyses for gradient descent methods in the training of deep neural networks

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:20:52.838309Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T13:20:47.384573Z digest=sha256:7dff689eddb4e310d88d0b962b51b895fa63af0f75dba971da48d33b6f6600b6

Observation 71cdcc58-6054-416f-b9c1-e954c0495dd4 · outbound

This paper cites Adam: A Method for Stochastic Optimization.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning Adam: A Method for Stochastic Optimization

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T13:20:47.424978Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:20:47.424978Z digest=sha256:162ed2ec3f5be94f0c9c7216fcd7c3d199986120017fbd632a8581d53f983d04

Observation d50492e0-b898-4191-a8c6-b43068b2341c · outbound

This paper cites SAD Neural Net- works: Divergent Gradient Flows and Asymptotic Optimality via o-minimal Structures.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning SAD Neural Net- works: Divergent Gradient Flows and Asymptotic Optimality via o-minimal Structures

Reference 35

Resolution
verified exact
raw_fallback, observed 2026-08-07T13:20:50.894625Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T13:20:47.515572Z digest=sha256:61343cd84c126cf753726d609dc606c2a9f8785710a66e30e775beb55e93f558

Observation 4b85d118-57fa-461d-80bd-5ab60eaed08c · outbound

This paper cites Convergence of Adam Under Relaxed Assumptions.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning Convergence of Adam Under Relaxed Assumptions

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T13:20:47.571118Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:20:47.571118Z digest=sha256:354de98986be4a5f11fefc5b4f873beba92cbd51a94545d0ff885982c41724f2

Observation 8102fa5d-3cb1-4def-8a85-973b59d35ec3 · outbound

This paper cites Understanding SGD with Exponential Moving Average: A Case Study in Linear Regression.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning Understanding SGD with Exponential Moving Average: A Case Study in Linear Regression

Reference 37

Resolution
verified exact
local_arxiv, observed 2026-08-07T13:20:50.561668Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T13:20:47.643359Z digest=sha256:dbec08893b7d3a1fb653e807214cec0e358ca040962b5c660c41c5370bfc995c

Observation e11fa63e-2b52-4fda-8849-89de09810771 · outbound

This paper cites Summary of ChatGPT-Related Research and Perspective Towards the Future of Large Language Models.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning Summary of ChatGPT-Related Research and Perspective Towards the Future of Large Language Models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T13:20:47.730894Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:20:47.730894Z digest=sha256:f90c93bd96ecd63246913d8f01cb79d38b0d52e139f7615696865265a3c7ad2c

Observation dd8c1174-40ba-4975-a786-50ceab7bdbc9 · outbound

This paper cites Decoupled Weight Decay Regularization.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning Decoupled Weight Decay Regularization

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T13:20:47.787142Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:20:47.787142Z digest=sha256:b98284cb20bf1141c82823699ae0f4e591659d6668190ff8ad17fc8b3fc9b335

Observation ff0b6f8b-69f6-458d-8324-e607f2d38ff7 · outbound

This paper cites Gradient Descent Maximizes the Margin of Homogeneous Neural Networks.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning Gradient Descent Maximizes the Margin of Homogeneous Neural Networks

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T13:20:47.870928Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:20:47.870928Z digest=sha256:2d9b94129c93a57fc4e8bfa872a3291134595647c9dacf60303baf1d8bd18103

Observation f5a2318d-ef61-414c-80ee-bb1a7df5a500 · outbound

This paper cites Stochastic Gradient Descent as Approximate Bayesian Inference.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning Stochastic Gradient Descent as Approximate Bayesian Inference

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T13:20:47.962359Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:20:47.962359Z digest=sha256:d6e9da0ab34685704cb026f1cfbfb628947b2312ce1c9d9e740b04226a58b499

Observation 5c447aaa-cef6-4293-a0f0-4205e3852af2 · outbound

This paper cites Exponential moving average of weights in deep learning: Dynamics and benefits.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning Exponential moving average of weights in deep learning: Dynamics and benefits

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:20:52.605227Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T13:20:48.034105Z digest=sha256:f5dbe98f234a9352de00732c39d3cc93af620d99f6320fae8fa41cdc23fa24b5

Observation 394582bf-e2c6-4e71-8955-e6c29c87a6dd · outbound

This paper cites Topological properties of the set of functions generated by neural networks of fixed size.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning Topological properties of the set of functions generated by neural networks of fixed size

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:20:52.390186Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T13:20:48.183493Z digest=sha256:d45fd963451b646642fe6a1ff32581e5bb22617453f4ba94ccae3226a1d6da9f

Observation 0d60ffcf-2a5e-498c-96fa-3b4ddff012c6 · outbound

This paper cites Continuous-time stochastic control and optimization with financial applications, vol.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning Continuous-time stochastic control and optimization with financial applications, vol

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:20:52.189882Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T13:20:48.301435Z digest=sha256:55634677890767be4b61af7d00e50924f2c1be9ce955849bf662212627c8b98d

Observation 20f15452-6034-473a-ba11-f349d65d8dc8 · outbound

This paper cites an unresolved cited work.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning Unresolved cited work

Reference 45

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:20:52.012419Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T13:20:48.411360Z digest=sha256:39ed391cc57ae83d6789aa7790814f7d4b094fd0dbbe4730e3c5624274827a15

Observation dd5bac77-c224-42ca-91f4-df1fe0e02d0c · outbound

This paper cites T., and Juditsky, A.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning T., and Juditsky, A

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:20:51.839610Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T13:20:48.626059Z digest=sha256:4174a53ca07e4882682e800b86c2ed6ac7e23f121c258cb37b4d7a6f3e5a258e

Observation 821a4347-7511-40bb-aac1-0a49b15ef796 · outbound

This paper cites Zero-Shot Text-to-Image Generation.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning Zero-Shot Text-to-Image Generation

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T13:20:48.749797Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:20:48.749797Z digest=sha256:ea708a86b96a00f735ddef339cf1709893571d3b9862506135076df89bce3565

Observation 8de2a81c-251e-457e-ad7f-3f692c0d2f2d · outbound

This paper cites On the Convergence of Adam and Beyond.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning On the Convergence of Adam and Beyond

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T13:20:48.882281Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:20:48.882281Z digest=sha256:4953b88c086ddf41660290d6c31bb023c621e9f1629c84a7d4eaeb6b8a160233

Observation 5ceea0a7-d7e6-42e1-83a8-f50dc56cd332 · outbound

This paper cites High-Resolution Image Synthesis with Latent Diffusion Models.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning High-Resolution Image Synthesis with Latent Diffusion Models

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T13:20:49.032440Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:20:49.032440Z digest=sha256:4671043cf2d4e5cbdf153d8b26005e7daca52cce2ebafe3c219ace5bb6723c74

Observation 08ca1edf-a5b2-41d6-998e-8c2c25e34313 · outbound

This paper cites An overview of gradient descent optimization algorithms.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning An overview of gradient descent optimization algorithms

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T13:20:49.150415Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:20:49.150415Z digest=sha256:3c7369974438d3dbccbfd1b3a601b31bd3a54525b598d685ba2828f5357b3ec9

Observation b7389969-55b8-4b52-9f5c-8df99e018d28 · outbound

This paper cites Efficient estimations from a slowly convergent Robbins-Monro process.Cor- nell University Operations Research and Industrial Engineering, hdl.handle.net/1813/8664 (1988), 1–34.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning Efficient estimations from a slowly convergent Robbins-Monro process.Cor- nell University Operations Research and Industrial Engineering, hdl.handle.net/1813/8664 (1988), 1–34

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:20:51.601182Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T13:20:49.266841Z digest=sha256:0ac596dcfdf6577bb8870c84a7a36ddd4fc2430f8fdd6c3ab6fefe63240be095

Observation 9ae9a0e3-f7c2-4cd3-a529-38c93cbaaf54 · outbound

This paper cites Photorealistic Text-to-Image Diffusion Models with Deep Language Understanding.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning Photorealistic Text-to-Image Diffusion Models with Deep Language Understanding

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T13:20:49.372617Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:20:49.372617Z digest=sha256:88b29277510f8a15872a3d246f55b75fb9a1dad7806e71f4ebbc03ae53cf69bb

Observation 59a99ae4-2b76-4bb0-ad7f-ecea03b0c639 · outbound

This paper cites Training trajectories, mini-batch losses and the curious role of the learning rate.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning Training trajectories, mini-batch losses and the curious role of the learning rate

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T13:20:49.539468Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:20:49.539468Z digest=sha256:a2a7148c584fcc92417f05329a0a5d437f6c3e7cc2cbf455c09550c89755a4d4

Observation 503a1d71-526c-4b3a-8797-e03a34ddf718 · outbound

This paper cites On Margin Maximization in Linear and ReLU Networks.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning On Margin Maximization in Linear and ReLU Networks

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-07T13:20:49.699637Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:20:49.699637Z digest=sha256:35c7cea223943ddb5ce09b8f650c2389d4ed8fa3edea55b98a4e63c42d207cdf

Observation de806598-2911-43d2-bd22-b44927348aae · outbound

This paper cites Deep learning with Elastic Averaging SGD.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning Deep learning with Elastic Averaging SGD

Reference 55

Resolution
verified exact
local_arxiv, observed 2026-08-07T13:20:50.167971Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T13:20:49.832951Z digest=sha256:6763d7c7cb0b8bbb47b24c4a89d12fbe1685f8d94769e75dad4bf626dfbb4f6f

Pith citing papers

Observation c4250ee0-78b1-45b3-8508-61ed43f75d3c · inbound

On the Provable Suboptimality of Momentum SGD in Nonstationary Stochastic Optimization cites this paper.

On the Provable Suboptimality of Momentum SGD in Nonstationary Stochastic Optimization PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-03T10:09:36.912553Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:09:36.912553Z digest=sha256:f7c9bcc0588cb66b5358a065027560a7cb19405835047f719c419787f38de4bf

Observation b0941b4b-8e5c-49ff-ba18-6a59519f5548 · inbound

Central limit theorem for the averaged Adam optimizer cites this paper.

Central limit theorem for the averaged Adam optimizer PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-07-04T07:29:38.374677Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-26T13:29:08.750421Z digest=sha256:684f9e59cb1791bd858b596ddf1b7e3405b16621423606bb6c2ceee1466229ef