Pith. sign in

Paper Citation Record · LEDGER

Sharpness-Aware Minimization with Adaptive Regularization for Training Deep Neural Networks

As of 19 August 2026, this Paper Citation Record lists 25 of 25 outbound references and 1 inbound Pith citation observation for arXiv:2412.16854.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.16854 v1

Coverage vector

measured 25 of 25 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T10:20:32.005290Z

measured 26 of 26 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:32:06.484361Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-07T12:32:10.377062Z

Reference resolution

25 of 25 outbound references displayed

  • verified exact1
  • verified fuzzy20
  • unresolved4
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 6f898516-4b53-4da1-b482-e93e40f8baf0 · outbound

This paper cites Adaptive regularization with cubics on manifolds.

Sharpness-Aware Minimization with Adaptive Regularization for Training Deep Neural Networks Adaptive regularization with cubics on manifolds

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:20:32.375324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T10:20:31.912207Z digest=sha256:763b9125fb748a06a6a0b4c2ec14e7168c23395d6ce3d72664caea52f75ee331

Observation 26b54dc2-7a68-4aaa-b6ec-3dcac239176e · outbound

This paper cites To- wards understanding sharpness-aware minimization.

Sharpness-Aware Minimization with Adaptive Regularization for Training Deep Neural Networks To- wards understanding sharpness-aware minimization

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:20:32.361467Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T10:20:31.916493Z digest=sha256:fe97c058ece3748681fcacdd0fcb404a28da8beed69e92adcc2429b001710919

Observation 2bc010f3-c479-4d51-ab4e-f6fa1a2144bf · outbound

This paper cites Stochastic analysis of an adaptive cubic regularization method under inexact gradient evaluations and dynamic hessian accu- racy.

Sharpness-Aware Minimization with Adaptive Regularization for Training Deep Neural Networks Stochastic analysis of an adaptive cubic regularization method under inexact gradient evaluations and dynamic hessian accu- racy

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:20:32.346168Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T10:20:31.920149Z digest=sha256:1a3fbd110f668f757cd797be970ec5410aec418e9e2f2896f4ca9ab951f25482

Observation 3f732cfd-1a1f-4cdb-949e-70d6f48f13ce · outbound

This paper cites Opti- mization methods for large-scale machine learning.

Sharpness-Aware Minimization with Adaptive Regularization for Training Deep Neural Networks Opti- mization methods for large-scale machine learning

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:20:32.331627Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T10:20:31.923814Z digest=sha256:4ae176e64957f864285c398965f0bc174f879aebeb91e788f07792fc29f1ec15

Observation 40cff607-9e95-4234-b1af-1197415105dc · outbound

This paper cites an unresolved cited work.

Sharpness-Aware Minimization with Adaptive Regularization for Training Deep Neural Networks Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-11T10:20:32.317081Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T10:20:31.927607Z digest=sha256:fa62f82b937c30de4fa52be8178a9b334b838a0c91cf157f9f38c696e9a0fbdb

Observation bd274ac9-a97d-4a53-8fb6-bac3306abcc7 · outbound

This paper cites Sharpness-aware minimization for efficiently improving generalization.

Sharpness-Aware Minimization with Adaptive Regularization for Training Deep Neural Networks Sharpness-aware minimization for efficiently improving generalization

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:20:32.304095Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T10:20:31.931378Z digest=sha256:e5fa4045d5b2145906ef2374463fff15f16068415545ed0bea489746a4ebbd2d

Observation 24f3da99-4886-4248-8fa3-bdbfc32ce352 · outbound

This paper cites Understanding the role of momentum in stochas- tic gradient methods.

Sharpness-Aware Minimization with Adaptive Regularization for Training Deep Neural Networks Understanding the role of momentum in stochas- tic gradient methods

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:20:32.292401Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T10:20:31.935326Z digest=sha256:98ee8a09a98648a0fa747650c084efdabf29aaa9797b232433ff8dd2663113f2

Observation c3c0b375-e7ae-4ae8-afd9-60b20206de4c · outbound

This paper cites Deep residual learning for image recognition.

Sharpness-Aware Minimization with Adaptive Regularization for Training Deep Neural Networks Deep residual learning for image recognition

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-11T10:20:31.939573Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T10:20:31.939573Z digest=sha256:17d03c663ca20009f5f3aceae334d8d337e3d5cf48cd4e63887fa70109ea5870

Observation eb84e23b-6a02-4f4e-9ee7-ef63ae4b32a0 · outbound

This paper cites Fantastic generalization measures and where to find them.

Sharpness-Aware Minimization with Adaptive Regularization for Training Deep Neural Networks Fantastic generalization measures and where to find them

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:20:32.273567Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T10:20:31.943782Z digest=sha256:fa4a3f4163c261a34154d4b226e15c2f68f4cca9274557d9cea7c79266080e8b

Observation c92bf434-3cb6-44da-b30e-50b19ecada3f · outbound

This paper cites On large-batch training for deep learning: Generalization gap and sharp minima.

Sharpness-Aware Minimization with Adaptive Regularization for Training Deep Neural Networks On large-batch training for deep learning: Generalization gap and sharp minima

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:20:32.262894Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T10:20:31.947465Z digest=sha256:fdbf2f60a38038bdadf92558641ba4fa50851a28c77c4e55824c0b6d8f0b47d8

Observation 1bb894df-d559-48ab-8f60-9606d57ff7bc · outbound

This paper cites Fundamental Convergence Analysis of Sharpness-Aware Minimization.

Sharpness-Aware Minimization with Adaptive Regularization for Training Deep Neural Networks Fundamental Convergence Analysis of Sharpness-Aware Minimization

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-11T10:20:31.951325Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T10:20:31.951325Z digest=sha256:fa3c6566df85b3baa857ee57a7c68841f54d62c4483015462744744f054693ef

Observation be592894-de9b-41b8-a350-75f12ec42aec · outbound

This paper cites Sub-sampled cubic regularization for non-convex optimization.

Sharpness-Aware Minimization with Adaptive Regularization for Training Deep Neural Networks Sub-sampled cubic regularization for non-convex optimization

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:20:32.252523Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T10:20:31.955828Z digest=sha256:00b7e1e41053dec8d85b6329dc8104a5bab753d27fec38b3780ba5499db27096

Observation d0beaf03-1ded-4ec0-a98c-7f7677f4b1a4 · outbound

This paper cites Asam: Adaptive sharpness-aware min- imization for scale-invariant learning of deep neural networks.

Sharpness-Aware Minimization with Adaptive Regularization for Training Deep Neural Networks Asam: Adaptive sharpness-aware min- imization for scale-invariant learning of deep neural networks

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:20:32.240526Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T10:20:31.960046Z digest=sha256:edc4038dda84bfbda67a0d0d2c70d838b8cbdc51a39a1e6507519bc425f0a4e8

Observation 7519dd12-06c5-416f-8185-480e501ff804 · outbound

This paper cites Giannakis.

Sharpness-Aware Minimization with Adaptive Regularization for Training Deep Neural Networks Giannakis

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:20:32.228256Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T10:20:31.963945Z digest=sha256:63b444b49b038b0a46cfcd305439aa3a5bec90f2a5bda7690d753f82f68a1a33

Observation 95230a08-8dfa-4d6e-9be7-4a8983f28d6c · outbound

This paper cites Friendly sharpness-aware minimization.

Sharpness-Aware Minimization with Adaptive Regularization for Training Deep Neural Networks Friendly sharpness-aware minimization

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:20:32.216957Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T10:20:31.967542Z digest=sha256:6719649ac528dbeac83dfc21d24b09235358cd1ce42eacb7c4cd45fe7e720b98

Observation 7b9bca27-3cff-4c7c-9630-eee894d1d20c · outbound

This paper cites A second look at exponential and cosine step sizes: Simplicity, adaptivity, and performance.

Sharpness-Aware Minimization with Adaptive Regularization for Training Deep Neural Networks A second look at exponential and cosine step sizes: Simplicity, adaptivity, and performance

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:20:32.193661Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T10:20:31.971305Z digest=sha256:bbe30f2874ebaecc004568f446cadc5b86b80ffed07082bca0c6dc100e12fbf4

Observation 513477bc-5877-4d30-9b1c-4967507fa690 · outbound

This paper cites Towards efficient and scalable sharpness- aware minimization.

Sharpness-Aware Minimization with Adaptive Regularization for Training Deep Neural Networks Towards efficient and scalable sharpness- aware minimization

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:20:32.181028Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T10:20:31.975105Z digest=sha256:410cd3fd8d000e6c36927e0dc60b4e6ed054bd0e55e779392397bf217f517d71

Observation 0e5b978f-f53a-4e90-8b56-58320f64454d · outbound

This paper cites Make sharpness- aware minimization stronger: A sparsified perturbation approach.

Sharpness-Aware Minimization with Adaptive Regularization for Training Deep Neural Networks Make sharpness- aware minimization stronger: A sparsified perturbation approach

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:20:32.167538Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T10:20:31.978966Z digest=sha256:6a401392c99f82d0f967cfdfce32357a91729b3d50dc6a1df6ac96f836372172

Observation bd22974f-cb28-4c98-9dfc-ab1a2a37e4d9 · outbound

This paper cites Adasam: Boosting sharpness-aware minimization with adaptive learning rate and momentum for training deep neural networks.

Sharpness-Aware Minimization with Adaptive Regularization for Training Deep Neural Networks Adasam: Boosting sharpness-aware minimization with adaptive learning rate and momentum for training deep neural networks

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:20:32.148697Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T10:20:31.983216Z digest=sha256:c17a8cd060ad1a30307c73ef285966a257fbd73582fe2ef07972d0c1e3353f3e

Observation a9f5790d-aa5b-443c-8ee5-15d6ae6a1474 · outbound

This paper cites Adaptive random walk gradient descent for decentralized optimiza- tion.

Sharpness-Aware Minimization with Adaptive Regularization for Training Deep Neural Networks Adaptive random walk gradient descent for decentralized optimiza- tion

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:20:32.135049Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T10:20:31.986859Z digest=sha256:0c28d5baa0741cf257f9514b6344542be0dc9cb2a1fa147b7252420f5a03242e

Observation 072c6cd5-7ced-40a1-9b2e-e18b90b2efdb · outbound

This paper cites Training Deep Neural Networks with Adaptive Momentum Inspired by the Quadratic Optimization.

Sharpness-Aware Minimization with Adaptive Regularization for Training Deep Neural Networks Training Deep Neural Networks with Adaptive Momentum Inspired by the Quadratic Optimization

Reference 21

Resolution
verified exact
local_arxiv, observed 2026-08-11T10:20:32.061155Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T10:20:31.990626Z digest=sha256:3e838acfb2df854fd8c13e863e12f8cc3e2aaba7d10c3abba466616bd2b211bd

Observation 678d2c40-e0ec-468c-9a8f-77bc6ee769e7 · outbound

This paper cites Novel convergence results of adaptive stochastic gradi- ent descents.

Sharpness-Aware Minimization with Adaptive Regularization for Training Deep Neural Networks Novel convergence results of adaptive stochastic gradi- ent descents

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:20:32.122097Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T10:20:31.994611Z digest=sha256:cdd37d7119113ab0efc06e5f36a75b4c5f58345a38a0b39552b570a97a6582ab

Observation 522755ad-3f02-4708-8351-2fe59989ca29 · outbound

This paper cites General proximal incremental aggregated gradient algo- rithms: Better and novel results under general scheme.

Sharpness-Aware Minimization with Adaptive Regularization for Training Deep Neural Networks General proximal incremental aggregated gradient algo- rithms: Better and novel results under general scheme

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:20:32.104600Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T10:20:31.997863Z digest=sha256:66a54e986fd5eba5e0d3bcb64dfdde9834f679fb3cb4090226fbe602911561bc

Observation a911a13d-aecc-4f63-8110-599d9f403150 · outbound

This paper cites Normalized stochastic heavy ball with adaptive momentum.

Sharpness-Aware Minimization with Adaptive Regularization for Training Deep Neural Networks Normalized stochastic heavy ball with adaptive momentum

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:20:32.089647Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T10:20:32.001407Z digest=sha256:826e33575134c1cb4a474523a8113429d2d0641f211e8219462fc3dac17fdfe2

Observation e1147683-b61a-434b-a8de-64b63ff08014 · outbound

This paper cites Wide Residual Networks.

Sharpness-Aware Minimization with Adaptive Regularization for Training Deep Neural Networks Wide Residual Networks

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-11T10:20:32.005290Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T10:20:32.005290Z digest=sha256:3da2b5f60884c447a20dcf8f1f6e96458382ac9294b4ebb9151c66492dbc82c9

Pith citing papers

Observation ab012672-d34c-472e-98f9-ad6716cbbf86 · inbound

LightSAM: Parameter-Agnostic Sharpness-Aware Minimization cites this paper.

LightSAM: Parameter-Agnostic Sharpness-Aware Minimization Sharpness-Aware Minimization with Adaptive Regularization for Training Deep Neural Networks

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-08-07T12:32:10.443643Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T12:32:06.484361Z digest=sha256:cb66ff90cf00bce7f3ac288b730d1f44260701f17422279d2497ddcef5bf32d6