Pith. sign in

Paper Citation Record · LEDGER

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent

As of 21 August 2026, this Paper Citation Record lists 56 of 56 outbound references and 0 inbound Pith citation observations for arXiv:2505.21651.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.21651 v1

Coverage vector

measured 56 of 56 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:30:29.286244Z

measured 56 of 56 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

56 of 56 outbound references displayed

  • verified exact2
  • verified fuzzy42
  • unresolved12
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 4ddf0f5f-4ee3-4ef8-8117-8dfa3e70f1be · outbound

This paper cites How Free is Parameter-Free Stochastic Optimization?.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent How Free is Parameter-Free Stochastic Optimization?

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T13:30:23.064115Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:30:23.064115Z digest=sha256:414e38ccf02c512dd80d04b1ab9c0d8c382df8cac6f078191a5b76b630fb8938

Observation aee387a0-0d2a-4cc7-9874-868a878da780 · outbound

This paper cites Calculus , volume 1.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Calculus , volume 1

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:38.322528Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T13:30:23.187262Z digest=sha256:d3817b80b09fdd63b34f3d79915e69f64c15a6ca7a47bef5de11c2ef6dab2bc3

Observation fd78edb9-5eaa-446c-be1a-9cf9327857e3 · outbound

This paper cites Gradient descent converges linearly for logistic regression on separable data.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Gradient descent converges linearly for logistic regression on separable data

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:38.105432Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T13:30:23.284766Z digest=sha256:64741f5873486da608a96fc3d35e991112404117fbcc85cbbd2bfc2b6fd3aafb

Observation 204a1cdd-8cea-4e8b-a5b1-2ac44fe571ee · outbound

This paper cites Julia: A fresh approach to numerical computing.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Julia: A fresh approach to numerical computing

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:37.889146Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T13:30:23.335801Z digest=sha256:0f2900fc92778d9f4fef2585a5390c05e4e196e2b095ff4ae2007a25a986079d

Observation dc942eac-f20d-4881-9c73-6ccf8118e61b · outbound

This paper cites Making SGD parameter-free.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Making SGD parameter-free

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:37.708704Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T13:30:23.434935Z digest=sha256:df230438e02e92d491413e43c54dae24c457a1780a4406737cff73e0731ef1c8

Observation 649b4055-4bd7-4127-91ae-22caed5799c1 · outbound

This paper cites Understanding and detecting convergence for stochastic gradient descent with momentum.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Understanding and detecting convergence for stochastic gradient descent with momentum

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:37.455861Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T13:30:23.514854Z digest=sha256:cfb6e491d284b8488068d105f653bda46df82481d234359cb7f3ef498e912fe7

Observation bfe847ce-4044-43d8-bb02-0ea0237f626a · outbound

This paper cites Convergence diagnostics for stochastic gradient descent with constant step size.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Convergence diagnostics for stochastic gradient descent with constant step size

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-08-07T13:30:29.874881Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T13:30:23.643611Z digest=sha256:3513cfd9eb268df57cc75cabdd815b9331d14f8673b0df67066e179082311e56

Observation 42083acd-ec66-4610-9a39-1d291963493f · outbound

This paper cites Automatically constructing a corpus of sentential paraphrases.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Automatically constructing a corpus of sentential paraphrases

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:37.288214Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T13:30:23.724831Z digest=sha256:561b04189084ffb9cc696da6956ee604eae60b833c665ec82710cd76c5308e27

Observation 9a2ecdec-5fd7-49f3-80ca-fc787084035f · outbound

This paper cites Robust, accurate stochastic optimization for variational inference.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Robust, accurate stochastic optimization for variational inference

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:36.998998Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T13:30:23.809334Z digest=sha256:5318caed020ea9247c43fd62fa6ad432306e2f07aaff95830d8a197c875ca3bd

Observation 4171f955-c560-453a-a9de-dfcea1481669 · outbound

This paper cites Adaptive subgradient methods for online learning and stochastic optimization.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Adaptive subgradient methods for online learning and stochastic optimization

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T13:30:23.876555Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:30:23.876555Z digest=sha256:4f86d5b09c461969d78efb9fcd9d8deea73f5d020a74761a17923ecbb4eac1cc

Observation 574adbda-bd38-4bb3-94fa-ad42f88a16f6 · outbound

This paper cites Learning-rate-free learning by D - A daptation.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Learning-rate-free learning by D - A daptation

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:36.763690Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T13:30:23.965152Z digest=sha256:e68fc83fb25f3e2ab0dd9336a00806bb346ea4178eae169410e1f3069ab2ddc7

Observation af7dafda-77e2-414a-8acb-756552a0c142 · outbound

This paper cites Markov Chains.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Markov Chains

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:36.610874Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T13:30:24.009501Z digest=sha256:df44d6d170e861c4b71fa6f4471a628c232828f9219489175a1a2b3cde019280

Observation aa54e8af-cc3b-455b-9b0e-3665668e6992 · outbound

This paper cites Probability: Theory and Examples.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Probability: Theory and Examples

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:36.444990Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T13:30:24.084769Z digest=sha256:12457d390a8724c04be56a8085bac31a7fd965a1622b62a9ac098919bf8a6ead

Observation b622ca89-67b0-408d-b08e-385de1e777df · outbound

This paper cites The road less scheduled.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent The road less scheduled

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:36.211598Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T13:30:24.161205Z digest=sha256:aac6ac8d9f068a886f2a484638052e53433a94e2f6e680ecb1c03fc79f561d2d

Observation abd233f5-8329-4f89-8b88-1d11de747c4e · outbound

This paper cites Bayesian Data Analysis.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Bayesian Data Analysis

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:36.005226Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T13:30:24.217374Z digest=sha256:785df5448ff59b38ec0dcd645be51f7210be48584a52de5f6a3f07b85824db49

Observation 7d99c6f7-0b53-42b8-a1aa-b86f26760ca8 · outbound

This paper cites Handbook of Convergence Theorems for (Stochastic) Gradient Methods.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Handbook of Convergence Theorems for (Stochastic) Gradient Methods

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T13:30:24.271122Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:30:24.271122Z digest=sha256:e4f05f3ef4b2137ad93f228282d2d070a2e3aff8903c90c97ba8ea47f3e4cedf

Observation f3906197-8710-4483-afbb-550483e6c94e · outbound

This paper cites Inference from iterative simulation using multiple sequences.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Inference from iterative simulation using multiple sequences

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:35.833611Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T13:30:24.344053Z digest=sha256:8625486fb3800ac5f38b740cdfb229254ea4592e8c0904b33190b857d6941e35

Observation f269af48-9b48-43d1-983c-85bc5d2e6d37 · outbound

This paper cites Don't be so monotone: R elaxing stochastic line search in over-parameterized models.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Don't be so monotone: R elaxing stochastic line search in over-parameterized models

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:35.705094Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T13:30:24.514784Z digest=sha256:58840adc28323396bcec88e6c3151ecdc86ce07ac49fcb4d59e45919c8b1acc4

Observation 0380042f-3f8f-460a-a8b9-8bc45ae1a90f · outbound

This paper cites Variance-reduced methods for machine learning.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Variance-reduced methods for machine learning

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:35.526086Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T13:30:24.644887Z digest=sha256:441a1820dd1dca8b5f6a47aa2dc41a9f82c54d8ba0813369a1348877892a14ea

Observation 36e6b8fe-9663-4636-b810-db23e70870ae · outbound

This paper cites Srivastava, and K.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Srivastava, and K

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:35.292698Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T13:30:24.734965Z digest=sha256:da2b539401f5bd72c2023262431290681d525a33c2938e8a9b6dc300b5ba95ec

Observation fb0b79d0-6103-4375-8ba0-994978c72d94 · outbound

This paper cites Deep residual learning for image recognition.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Deep residual learning for image recognition

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T13:30:24.804785Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:30:24.804785Z digest=sha256:f95a065b9a8f1b7c10896dad5303ac90038ae47de82632a7e3fa798f8989670b

Observation e243249d-3c26-45c5-80ad-d1392bbc1618 · outbound

This paper cites DoG is SGD 's best friend: A parameter-free dynamic step size schedule.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent DoG is SGD 's best friend: A parameter-free dynamic step size schedule

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:35.106516Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T13:30:24.917999Z digest=sha256:5039997f4431e8b281ebda45490b5e982d7e48a9c5fd7294772fe7ee491d7fa0

Observation a7c3fab8-dca2-42b3-b1d3-4e9e0e683339 · outbound

This paper cites Parallelizing stochastic gradient descent for least squares regression: mini-batching, averaging, and model misspecification.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Parallelizing stochastic gradient descent for least squares regression: mini-batching, averaging, and model misspecification

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:34.900072Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T13:30:25.014606Z digest=sha256:b6b284864d4b657629930857a8f616160e44af91bca18bd2b00f88ac137830fe

Observation 8c143f95-dbd4-4722-969e-b57ec402f1a1 · outbound

This paper cites Adam: A Method for Stochastic Optimization.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Adam: A Method for Stochastic Optimization

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T13:30:25.236602Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:30:25.236602Z digest=sha256:bf238247c572c8d96b67e76f7592774da2c155f6d9fbefaf0db04612d143ea9b

Observation e9f5da1e-7290-4058-a06b-8bc3a8160200 · outbound

This paper cites Accelerated parameter-free stochastic optimization.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Accelerated parameter-free stochastic optimization

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:34.699088Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T13:30:25.364739Z digest=sha256:6b31030777dc2f284ddede59e9a49ffe4be2d0c40603f8eaf4c125f3d453c97e

Observation f1dbf6ab-fc57-427a-9682-0f28089a6926 · outbound

This paper cites Tuning-Free Stochastic Optimization.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Tuning-Free Stochastic Optimization

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T13:30:25.644802Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:30:25.644802Z digest=sha256:db494c6b1b8ac84a2da34bb573d21d9c189748f10bfb397db8197e7efe1b3021

Observation 332916ce-44bc-4a4c-a515-c14d1a4eabfb · outbound

This paper cites Linear convergence of black-box variational inference: S hould we stick the landing? In International Conference on Artificial Intelligence and Statistics , pages 235--243.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Linear convergence of black-box variational inference: S hould we stick the landing? In International Conference on Artificial Intelligence and Statistics , pages 235--243

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:34.500231Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T13:30:25.815236Z digest=sha256:c7e180e0d206ef992c9efd9489cfd272fadd00451d9a503e5e6d6a289741f2a6

Observation cc97eaae-22fe-433a-a526-03e65b9c99e0 · outbound

This paper cites DoWG unleashed: A n efficient universal parameter-free gradient descent method.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent DoWG unleashed: A n efficient universal parameter-free gradient descent method

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:34.345067Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T13:30:25.935152Z digest=sha256:8363b380fc34a4782788c1485037fa871ace1a8c6ae66af5fcb6226758e198e9

Observation 3d872e16-63cf-40c7-8062-5de2ee988c80 · outbound

This paper cites Learning multiple layers of features from tiny images.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Learning multiple layers of features from tiny images

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T13:30:26.009608Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:30:26.009608Z digest=sha256:745e371f3ab7739a58b90f184c9a93fc864a4f4db6ec4eeb8406ce99f5c73502

Observation 3a8ed22e-5ebb-4893-875d-715a7d6c99ab · outbound

This paper cites RoBERTa: A Robustly Optimized BERT Pretraining Approach.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent RoBERTa: A Robustly Optimized BERT Pretraining Approach

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T13:30:26.125015Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:30:26.125015Z digest=sha256:85389d2a2300162a271999f5881ab4612afb8fd44751f1358c39ed46aa5bac56

Observation 8666a607-6920-4c21-8721-f33055dabb10 · outbound

This paper cites Stochastic polyak step-size for SGD : A n adaptive learning rate for fast convergence.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Stochastic polyak step-size for SGD : A n adaptive learning rate for fast convergence

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:34.051792Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T13:30:26.255069Z digest=sha256:dca0fc8e6c19a5c5983fa360ddf8f5fe59142b9e60829f2a0f3e669226603d3f

Observation 6bb0fb86-b258-49aa-aa2b-d30b47fdaf7e · outbound

This paper cites Using statistics to automate stochastic optimization.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Using statistics to automate stochastic optimization

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:33.896297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T13:30:26.386965Z digest=sha256:a7f9f4c53b7967420c1ab2de8a6988cf1c37a569c1e5009e92845719208a696d

Observation ea3a5ea4-6857-4233-9231-459089494ea2 · outbound

This paper cites Prodigy: An Expeditiously Adaptive Parameter-Free Learner.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Prodigy: An Expeditiously Adaptive Parameter-Free Learner

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T13:30:26.479605Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:30:26.479605Z digest=sha256:b6780469dbe10aa926769f3afd4a4be2bffef62e116e94f687c0322a33322ea6

Observation 62a50b7a-aa01-4eb5-a9bc-605cdc44fc4f · outbound

This paper cites Adaptive Gradient Descent without Descent.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Adaptive Gradient Descent without Descent

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T13:30:26.665057Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:30:26.665057Z digest=sha256:9307265342a7834cd2649758b6033c86c87b3661c82abf79f1cd49e04d4ada20

Observation 23103646-e9af-4aaf-8a31-f8b19b14b154 · outbound

This paper cites Beating SGD saturation with tail-averaging and minibatching.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Beating SGD saturation with tail-averaging and minibatching

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:33.733177Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T13:30:26.964874Z digest=sha256:5c16b01124d611845a429c7be12057e79cd837f864006a6f48e534b223806e7f

Observation 662b3f75-7b27-4b96-8588-01f6aa9ba61f · outbound

This paper cites Let's make block coordinate descent converge faster: F aster greedy rules, message-passing, active-set complexity, and superlinear convergence.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Let's make block coordinate descent converge faster: F aster greedy rules, message-passing, active-set complexity, and superlinear convergence

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:33.571767Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T13:30:27.065084Z digest=sha256:f3e9327fe0784fe59aeba755498c229f67a21a659325d9793e93d03fa19cb251

Observation fcb7d2fb-32de-4f7e-a141-c063f88ba817 · outbound

This paper cites Dynamics of SGD with stochastic P olyak stepsizes: T ruly adaptive variants and convergence to exact solution.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Dynamics of SGD with stochastic P olyak stepsizes: T ruly adaptive variants and convergence to exact solution

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:33.370986Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T13:30:27.182224Z digest=sha256:8235355ac0c8cd0450d462b4ae966853ffa169d16a24726395ed1267f0fd4298

Observation efce4528-360d-4278-8025-944fffac552d · outbound

This paper cites Training deep networks without learning rates through coin betting.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Training deep networks without learning rates through coin betting

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:33.098648Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T13:30:27.414972Z digest=sha256:8b2cea5c954e7a08924dabf82733fc3cb58eb30b3434dc3419f06e611d64f650

Observation b1f77e31-f3f6-4b17-aa45-471c6eeaebac · outbound

This paper cites On convergence-diagnostic based step sizes for stochastic gradient descent.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent On convergence-diagnostic based step sizes for stochastic gradient descent

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:32.885767Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T13:30:27.584877Z digest=sha256:4d4e50a20496d3cb56395694e0f253ffd3561f5286bef1de3b7dbe79b40da06a

Observation c25f0525-7fc9-440f-9ef4-81f3c216a59f · outbound

This paper cites On the determination of the step size in stochastic quasigradient methods.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent On the determination of the step size in stochastic quasigradient methods

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:32.639217Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T13:30:27.666340Z digest=sha256:fabd431a500b19e70f6cb1c2c8264d7c546ec97915bcdce4b033cde633645234

Observation b6ffb9cd-dac4-46c0-8cec-a7c171abf7f7 · outbound

This paper cites Non-asymptotic confidence bounds for stochastic approximation algorithms with constant step size.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Non-asymptotic confidence bounds for stochastic approximation algorithms with constant step size

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:32.392834Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T13:30:27.774917Z digest=sha256:30570823ae4db4d23a38cc6863f3019deb6c3752bb42ab813370efab9bf4c4b8

Observation 359e5e04-282e-4398-89dc-05087705d8ab · outbound

This paper cites Py T orch: A n imperative style, high-performance deep learning library.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Py T orch: A n imperative style, high-performance deep learning library

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:32.241807Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T13:30:27.905036Z digest=sha256:8ece6336337251ab6b1e798081999502ef3c74823851333a9bf788c0e8d9d673

Observation acf787f1-2a75-4fed-967d-d50e6342a6a0 · outbound

This paper cites https://huggingface.co/microsoft/resnet-18, 2025.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent https://huggingface.co/microsoft/resnet-18, 2025

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:32.018044Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T13:30:28.010649Z digest=sha256:fb1a673e26b92399f7b91b842cc3f7340a774639ce3555c20b25ed9f07224640

Observation 6454f534-0c17-463f-9b38-3a1d916d3ddc · outbound

This paper cites A stochastic approximation method.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent A stochastic approximation method

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:31.783988Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T13:30:28.121239Z digest=sha256:dd5a3fa2d4ce43f4d94402be8b6c0dcab43d790dc0b2672d87edbfc50242d7f9

Observation 4a60d70a-43fc-4b69-9cd4-e010c9f1012e · outbound

This paper cites https://huggingface.co/FacebookAI/roberta-base, 2025.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent https://huggingface.co/FacebookAI/roberta-base, 2025

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:31.514873Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T13:30:28.463148Z digest=sha256:7f1deadee94215a4d75f517c10e8fc6a5fd959eef906c9b9f56296a34f07a36c

Observation 7927b180-2727-410e-b59e-a6b26519e75e · outbound

This paper cites Anytime Tail Averaging.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Anytime Tail Averaging

Reference 46

Resolution
verified exact
local_arxiv, observed 2026-08-07T13:30:29.584753Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T13:30:28.540616Z digest=sha256:d637d2609ddfa3d21688a488e3eb4623e1d5fed9a5f91e114888ca81961f33c3

Observation dbfa0b42-1d36-4900-9839-ef3ce529f51a · outbound

This paper cites Making Gradient Descent Optimal for Strongly Convex Stochastic Optimization.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Making Gradient Descent Optimal for Strongly Convex Stochastic Optimization

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T13:30:28.585346Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:30:28.585346Z digest=sha256:b4f03a9f5846bb6596c1d0ae93c99df325906c2c197e3fc0c34f3adef57f1228

Observation 2f10f48c-f16a-44cb-941d-d5d633ef795d · outbound

This paper cites Sticking the landing: S imple, lower-variance gradient estimators for variational inference.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Sticking the landing: S imple, lower-variance gradient estimators for variational inference

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:31.347644Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T13:30:28.687107Z digest=sha256:3fdabb8ce26696f8e9f555e924a39ce2d3574ff208ea52e04e5d9ee98d9cd144

Observation 14442b01-b4df-40e7-b73f-8a897ade788d · outbound

This paper cites SQuAD : 100,000+ questions for machine comprehension of text.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent SQuAD : 100,000+ questions for machine comprehension of text

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:31.130714Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T13:30:28.737065Z digest=sha256:587841ff7d4d110210207c5b8fd8c3f23012c99ee7ff2c3d9d7f1c4dae73a9f8

Observation 8d42f786-343b-4ec1-8444-9576d3126183 · outbound

This paper cites Virtual library of simulation experiments: T est functions and datasets, 2013.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Virtual library of simulation experiments: T est functions and datasets, 2013

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:30.933615Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T13:30:28.815036Z digest=sha256:6f6710cfd20cdc8a8074430ed49a0fb3c0b7f1d5fd577f00dfcdf070a5347d43

Observation 0acc27dd-ad8a-41c7-8cbe-df18780bf2d4 · outbound

This paper cites Recursive deep models for semantic compositionality over a sentiment treebank.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Recursive deep models for semantic compositionality over a sentiment treebank

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:30.762131Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T13:30:28.856405Z digest=sha256:a9c0b32f7218a25306900c8ce4d63c544dee3b51ae3990f50153c5336298e0f7

Observation b597a55d-446a-46dd-83f6-c4175b93c23b · outbound

This paper cites Stochastic gradient descent for non-smooth optimization: C onvergence results and optimal averaging schemes.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Stochastic gradient descent for non-smooth optimization: C onvergence results and optimal averaging schemes

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:30.583910Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T13:30:28.911294Z digest=sha256:7645e86e62fa04f307e19ac60a27dd652db6c1103eb18780b874af10047f27d0

Observation f81a50cd-0413-41b8-9408-88f6d4dd081f · outbound

This paper cites Painless stochastic gradient: I nterpolation, line-search, and convergence rates.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Painless stochastic gradient: I nterpolation, line-search, and convergence rates

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:30.444768Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T13:30:28.941336Z digest=sha256:9bd3349ff027942b6961c1eab92fd0042d1edc454d8a90a2430ae067ff5bb9e5

Observation 19139d0b-134b-46c5-a154-68bf21d15add · outbound

This paper cites A framework for improving the reliability of black-box variational inference.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent A framework for improving the reliability of black-box variational inference

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:30.280853Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T13:30:29.073747Z digest=sha256:c40424383cbf2bfa08ff868dd41b26b00a164fed9a173efa98ae7fddd661e408

Observation f258329d-7b20-498c-a848-44bc1f716d5b · outbound

This paper cites GLUE : A multi-task benchmark and analysis platform for natural language understanding.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent GLUE : A multi-task benchmark and analysis platform for natural language understanding

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:30.093008Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T13:30:29.175993Z digest=sha256:5564c22b535b5507868300b688ce84d50943441a013ed856ff5ef70e7f7e9fbb

Observation 563ed8bc-d756-4a3a-a11e-8f9e3f48bb71 · outbound

This paper cites Fluctuation-dissipation relations for stochastic gradient descent.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Fluctuation-dissipation relations for stochastic gradient descent

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-07T13:30:29.286244Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:30:29.286244Z digest=sha256:0b225f828d8fab7006e7f985b92254b223a485fd6382ac64b24e0e59bb394c68

Pith citing papers

No inbound Pith citation observations are available.