Pith. sign in

Paper Citation Record · LEDGER

Training trajectories, mini-batch losses and the curious role of the learning rate

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:2301.02312.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2301.02312 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 5 of 5 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-09T21:56:50.782654Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-20T12:23:16.782786Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 40e4be6d-56a9-4089-8150-8e11153d32d5 · inbound

The Surprising Agreement Between Convex Optimization Theory and Learning-Rate Scheduling for Large Model Training cites this paper.

The Surprising Agreement Between Convex Optimization Theory and Learning-Rate Scheduling for Large Model Training Training trajectories, mini-batch losses and the curious role of the learning rate

Reference 1970

Resolution
unresolved
no resolver link, observed 2026-08-09T21:56:50.782654Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T21:56:50.782654Z digest=sha256:c8813e7f4140bdc6a301c7d3f65bd56713dbff6074663fc9a22a98475fe8cc3c

Observation 59a99ae4-2b76-4bb0-ad7f-ecea03b0c639 · inbound

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning cites this paper.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning Training trajectories, mini-batch losses and the curious role of the learning rate

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T13:20:49.539468Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:20:49.539468Z digest=sha256:53f926cd23eecb216112a4b3bc56c0e587713ead6721a0354672e5b2975c0738

Observation 9f1d570c-bd0e-43dd-801f-db8643107920 · inbound

WSM: Decay-Free Learning Rate Schedule via Checkpoint Merging for LLM Pre-training cites this paper.

WSM: Decay-Free Learning Rate Schedule via Checkpoint Merging for LLM Pre-training Training trajectories, mini-batch losses and the curious role of the learning rate

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-06T14:49:40.322544Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:49:40.322544Z digest=sha256:3fa9566b0ef34ea821b9312dd5a598610184a788f2ab5c49fdcf11da01589a38

Observation 8a87a096-54f8-4fa8-9708-00f1012e76a9 · inbound

EMA Without the Lag: Bias-Corrected Iterate Averaging Schemes cites this paper.

EMA Without the Lag: Bias-Corrected Iterate Averaging Schemes Training trajectories, mini-batch losses and the curious role of the learning rate

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-06T10:23:51.923341Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:23:51.923341Z digest=sha256:3c8ace6eff1cb08499bebee3acda9e6603524cec3612452b851758d91869ac10

Observation 8bf73e49-b023-47a1-b2f6-0d8210f4ed28 · inbound

ScheduleFree+: Scaling Learning-Rate-Free & Schedule-Free Learning to Large Language Models cites this paper.

ScheduleFree+: Scaling Learning-Rate-Free & Schedule-Free Learning to Large Language Models Training trajectories, mini-batch losses and the curious role of the learning rate

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-20T12:23:16.784243Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-20T12:22:30.263086Z digest=sha256:1dd10b40af9af057eed5d2a7de2286a263438196df94225e3bb28d6edf8abd00