Pith. sign in

Paper Citation Record · LEDGER

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent

As of 10 August 2026, this Paper Citation Record lists 56 of 56 outbound references and 0 inbound Pith citation observations for arXiv:2505.21651.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.21651 v1

Coverage vector

measured 56 of 56 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:30:29.286244Z

measured 56 of 56 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

56 of 56 outbound references displayed

  • verified exact2
  • verified fuzzy42
  • unresolved12
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 4ddf0f5f-4ee3-4ef8-8117-8dfa3e70f1be · outbound

This paper cites How Free is Parameter-Free Stochastic Optimization?.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent How Free is Parameter-Free Stochastic Optimization?

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T13:30:23.064115Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:30:23.064115Z digest=sha256:44ad38895b306a15119ed46a7cc0b54cf079caffb68b7d9e8166feead91bc928

Observation aee387a0-0d2a-4cc7-9874-868a878da780 · outbound

This paper cites Calculus , volume 1.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Calculus , volume 1

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:38.322528Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:30:23.187262Z digest=sha256:6cf40e02975526d1532fdd49be974861b34ce02f4e769f4115571d1be5327aaa

Observation fd78edb9-5eaa-446c-be1a-9cf9327857e3 · outbound

This paper cites Gradient descent converges linearly for logistic regression on separable data.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Gradient descent converges linearly for logistic regression on separable data

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:38.105432Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:30:23.284766Z digest=sha256:bf0037c52eb37fab6c46adf817f119527576d0de057408cbbff508f54640f6a0

Observation 204a1cdd-8cea-4e8b-a5b1-2ac44fe571ee · outbound

This paper cites Julia: A fresh approach to numerical computing.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Julia: A fresh approach to numerical computing

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:37.889146Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:30:23.335801Z digest=sha256:314a8f30bddafcda629b14513817728b89b109bc6e55614bc5be816f7a404619

Observation dc942eac-f20d-4881-9c73-6ccf8118e61b · outbound

This paper cites Making SGD parameter-free.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Making SGD parameter-free

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:37.708704Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:30:23.434935Z digest=sha256:9ab722b2ba0f28d7a8631f5a257e166f44067693a43195967493b6eb8a3bf476

Observation 649b4055-4bd7-4127-91ae-22caed5799c1 · outbound

This paper cites Understanding and detecting convergence for stochastic gradient descent with momentum.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Understanding and detecting convergence for stochastic gradient descent with momentum

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:37.455861Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:30:23.514854Z digest=sha256:cf085221d5350f8c22ecf19be8f1e7c775817cd93b35a7108cbdca73c384a727

Observation bfe847ce-4044-43d8-bb02-0ea0237f626a · outbound

This paper cites Convergence diagnostics for stochastic gradient descent with constant step size.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Convergence diagnostics for stochastic gradient descent with constant step size

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-08-07T13:30:29.874881Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:30:23.643611Z digest=sha256:023f14216ec424f4c9ad212ab4821fd4b126abbf61e0f4ac991efe3019673c10

Observation 42083acd-ec66-4610-9a39-1d291963493f · outbound

This paper cites Automatically constructing a corpus of sentential paraphrases.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Automatically constructing a corpus of sentential paraphrases

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:37.288214Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:30:23.724831Z digest=sha256:d65cd698aab8953129dc40cd7037518f67b763782e71c4e1b401ad7d0a9dadf4

Observation 9a2ecdec-5fd7-49f3-80ca-fc787084035f · outbound

This paper cites Robust, accurate stochastic optimization for variational inference.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Robust, accurate stochastic optimization for variational inference

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:36.998998Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:30:23.809334Z digest=sha256:f4d9c1e3d10d8af701aa64b3b6f7a199cc315f26df4db0f8719123e43b05b1b3

Observation 4171f955-c560-453a-a9de-dfcea1481669 · outbound

This paper cites Adaptive subgradient methods for online learning and stochastic optimization.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Adaptive subgradient methods for online learning and stochastic optimization

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T13:30:23.876555Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:30:23.876555Z digest=sha256:3ff3b93d9fbeb80f75a8723c69369ca3dcb84ede4fd35e3a67180fc473be6258

Observation 574adbda-bd38-4bb3-94fa-ad42f88a16f6 · outbound

This paper cites Learning-rate-free learning by D - A daptation.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Learning-rate-free learning by D - A daptation

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:36.763690Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:30:23.965152Z digest=sha256:47271a5dbfb8593006cb92e1cf43b8de7d1e2e18fef87800d366f89b77d6257a

Observation af7dafda-77e2-414a-8acb-756552a0c142 · outbound

This paper cites Markov Chains.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Markov Chains

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:36.610874Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:30:24.009501Z digest=sha256:7e3653006e6e2d043388b2c43664df8d7880e31bd7746b59cecc607fc5761792

Observation aa54e8af-cc3b-455b-9b0e-3665668e6992 · outbound

This paper cites Probability: Theory and Examples.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Probability: Theory and Examples

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:36.444990Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:30:24.084769Z digest=sha256:d4eb639d9d7e766b31be197624cb9907c8b0cbfea881c4141541cc3218465f95

Observation b622ca89-67b0-408d-b08e-385de1e777df · outbound

This paper cites The road less scheduled.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent The road less scheduled

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:36.211598Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:30:24.161205Z digest=sha256:4d5e895af008e72386e6d2b8f6fca3cccbdc5686d96ccc55351bb741a62e2d50

Observation abd233f5-8329-4f89-8b88-1d11de747c4e · outbound

This paper cites Bayesian Data Analysis.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Bayesian Data Analysis

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:36.005226Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:30:24.217374Z digest=sha256:9a412526a8874fcabdac1177f935f50bce5de28d254bcee16fdecc291ad188af

Observation 7d99c6f7-0b53-42b8-a1aa-b86f26760ca8 · outbound

This paper cites Handbook of Convergence Theorems for (Stochastic) Gradient Methods.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Handbook of Convergence Theorems for (Stochastic) Gradient Methods

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T13:30:24.271122Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:30:24.271122Z digest=sha256:fe87fa45afe2afabe6d7cd51ad4b937b126329020b59c1a9ba477d352bf4d00c

Observation f3906197-8710-4483-afbb-550483e6c94e · outbound

This paper cites Inference from iterative simulation using multiple sequences.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Inference from iterative simulation using multiple sequences

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:35.833611Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:30:24.344053Z digest=sha256:4f7e522e0f9c9adc2761ca8104f158eb287e7f7593995cf3171c40a956dd13db

Observation f269af48-9b48-43d1-983c-85bc5d2e6d37 · outbound

This paper cites Don't be so monotone: R elaxing stochastic line search in over-parameterized models.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Don't be so monotone: R elaxing stochastic line search in over-parameterized models

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:35.705094Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:30:24.514784Z digest=sha256:c9983591217a294f57c05a332fb6bfe08a6fa9f9917c89c498224706d9765bb4

Observation 0380042f-3f8f-460a-a8b9-8bc45ae1a90f · outbound

This paper cites Variance-reduced methods for machine learning.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Variance-reduced methods for machine learning

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:35.526086Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:30:24.644887Z digest=sha256:cfc65b7231ededd6d6758ba6ad40d1530bc2dd98b1685a579c597631290cb888

Observation 36e6b8fe-9663-4636-b810-db23e70870ae · outbound

This paper cites Srivastava, and K.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Srivastava, and K

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:35.292698Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:30:24.734965Z digest=sha256:c89a1af9d0f81b9f4f6981e8ae626e74b921e5f59b990322a280d4f9a53dd0b1

Observation fb0b79d0-6103-4375-8ba0-994978c72d94 · outbound

This paper cites Deep residual learning for image recognition.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Deep residual learning for image recognition

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T13:30:24.804785Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:30:24.804785Z digest=sha256:3ab0191390515c7ca2fcd91c4c5911001a54e8401c009b4e400d070cd32348b2

Observation e243249d-3c26-45c5-80ad-d1392bbc1618 · outbound

This paper cites DoG is SGD 's best friend: A parameter-free dynamic step size schedule.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent DoG is SGD 's best friend: A parameter-free dynamic step size schedule

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:35.106516Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:30:24.917999Z digest=sha256:00f679ee28f9ad5c0658b9c1bba623dda686db5525f1c6e8c83e6e755053a5c5

Observation a7c3fab8-dca2-42b3-b1d3-4e9e0e683339 · outbound

This paper cites Parallelizing stochastic gradient descent for least squares regression: mini-batching, averaging, and model misspecification.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Parallelizing stochastic gradient descent for least squares regression: mini-batching, averaging, and model misspecification

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:34.900072Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:30:25.014606Z digest=sha256:d2e1356ce4bb2aed6e934938677f5b947d8befe6a4662231c39695f0536a6393

Observation 8c143f95-dbd4-4722-969e-b57ec402f1a1 · outbound

This paper cites Adam: A Method for Stochastic Optimization.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Adam: A Method for Stochastic Optimization

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T13:30:25.236602Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:30:25.236602Z digest=sha256:658392a826c1b0eb336450c6b7617db37ffd34fd4410dbb5700040328234842f

Observation e9f5da1e-7290-4058-a06b-8bc3a8160200 · outbound

This paper cites Accelerated parameter-free stochastic optimization.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Accelerated parameter-free stochastic optimization

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:34.699088Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:30:25.364739Z digest=sha256:c99173eea3f211d9e1060f63ac71f26c89e7fdc8be5ec8552345989d3a0beaa7

Observation f1dbf6ab-fc57-427a-9682-0f28089a6926 · outbound

This paper cites Tuning-Free Stochastic Optimization.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Tuning-Free Stochastic Optimization

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T13:30:25.644802Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:30:25.644802Z digest=sha256:578ac25564ee5c19ec86876801ab42fb552f143f0c42382a6b9838fa302e5882

Observation 332916ce-44bc-4a4c-a515-c14d1a4eabfb · outbound

This paper cites Linear convergence of black-box variational inference: S hould we stick the landing? In International Conference on Artificial Intelligence and Statistics , pages 235--243.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Linear convergence of black-box variational inference: S hould we stick the landing? In International Conference on Artificial Intelligence and Statistics , pages 235--243

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:34.500231Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:30:25.815236Z digest=sha256:044f147ed2ee4b1926742ce12b21cbe178096dbc8f8274b74368ec4abfd3eb88

Observation cc97eaae-22fe-433a-a526-03e65b9c99e0 · outbound

This paper cites DoWG unleashed: A n efficient universal parameter-free gradient descent method.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent DoWG unleashed: A n efficient universal parameter-free gradient descent method

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:34.345067Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:30:25.935152Z digest=sha256:d27be8e85ab7d8704ef565d9a310fa559441a4e69585002fb8cff7c3992fa4ce

Observation 3d872e16-63cf-40c7-8062-5de2ee988c80 · outbound

This paper cites Learning multiple layers of features from tiny images.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Learning multiple layers of features from tiny images

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T13:30:26.009608Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:30:26.009608Z digest=sha256:53748e6b2eb3d7a6900467c2be4028a0c59d7386619f1b98ed31601b36279697

Observation 3a8ed22e-5ebb-4893-875d-715a7d6c99ab · outbound

This paper cites RoBERTa: A Robustly Optimized BERT Pretraining Approach.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent RoBERTa: A Robustly Optimized BERT Pretraining Approach

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T13:30:26.125015Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:30:26.125015Z digest=sha256:0fc56ea1f5b0b74c6a5a98241cb097e44f4cf35e9f0759661935550da6cf21db

Observation 8666a607-6920-4c21-8721-f33055dabb10 · outbound

This paper cites Stochastic polyak step-size for SGD : A n adaptive learning rate for fast convergence.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Stochastic polyak step-size for SGD : A n adaptive learning rate for fast convergence

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:34.051792Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:30:26.255069Z digest=sha256:7c3aebe687d4ffff428e83d3e38db75e002465911b0d60e655ed9ce474266169

Observation 6bb0fb86-b258-49aa-aa2b-d30b47fdaf7e · outbound

This paper cites Using statistics to automate stochastic optimization.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Using statistics to automate stochastic optimization

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:33.896297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:30:26.386965Z digest=sha256:e495deebc0ab7f37c38439babecacf9871dc9614a34462078d72df26293d57fa

Observation ea3a5ea4-6857-4233-9231-459089494ea2 · outbound

This paper cites Prodigy: An Expeditiously Adaptive Parameter-Free Learner.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Prodigy: An Expeditiously Adaptive Parameter-Free Learner

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T13:30:26.479605Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:30:26.479605Z digest=sha256:f1d8d0969f9f052d7b111de7f56c5d23398c2b85f22d35b628634d032b895f5a

Observation 62a50b7a-aa01-4eb5-a9bc-605cdc44fc4f · outbound

This paper cites Adaptive Gradient Descent without Descent.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Adaptive Gradient Descent without Descent

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T13:30:26.665057Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:30:26.665057Z digest=sha256:b53b5e8bbd69b0bdddc0e9e8ad91a43661740253b819b4d9f41528bc5f18ea98

Observation 23103646-e9af-4aaf-8a31-f8b19b14b154 · outbound

This paper cites Beating SGD saturation with tail-averaging and minibatching.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Beating SGD saturation with tail-averaging and minibatching

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:33.733177Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:30:26.964874Z digest=sha256:da8d397080b7909f8b151ae406a63da4c7bb0fcdae693479a2b461270f87b112

Observation 662b3f75-7b27-4b96-8588-01f6aa9ba61f · outbound

This paper cites Let's make block coordinate descent converge faster: F aster greedy rules, message-passing, active-set complexity, and superlinear convergence.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Let's make block coordinate descent converge faster: F aster greedy rules, message-passing, active-set complexity, and superlinear convergence

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:33.571767Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:30:27.065084Z digest=sha256:7ddf07c499fe9bc57a55bbdbb9e15aa8eeee650a45583861d4d19d3fe9ef4aae

Observation fcb7d2fb-32de-4f7e-a141-c063f88ba817 · outbound

This paper cites Dynamics of SGD with stochastic P olyak stepsizes: T ruly adaptive variants and convergence to exact solution.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Dynamics of SGD with stochastic P olyak stepsizes: T ruly adaptive variants and convergence to exact solution

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:33.370986Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:30:27.182224Z digest=sha256:f7ce03292f377ae0d3e1a25a49c51c7930cca1aa6a7dd1c59da72701f54d3f90

Observation efce4528-360d-4278-8025-944fffac552d · outbound

This paper cites Training deep networks without learning rates through coin betting.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Training deep networks without learning rates through coin betting

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:33.098648Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:30:27.414972Z digest=sha256:76a26be07d822412f3ac1e83ef0f5790f4acd28e3783951daabcddcc5a8bc7cb

Observation b1f77e31-f3f6-4b17-aa45-471c6eeaebac · outbound

This paper cites On convergence-diagnostic based step sizes for stochastic gradient descent.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent On convergence-diagnostic based step sizes for stochastic gradient descent

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:32.885767Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:30:27.584877Z digest=sha256:4b535e7c9620f6c8533c2071393d60f363b2b8027a22f902255db8cb061dadf1

Observation c25f0525-7fc9-440f-9ef4-81f3c216a59f · outbound

This paper cites On the determination of the step size in stochastic quasigradient methods.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent On the determination of the step size in stochastic quasigradient methods

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:32.639217Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:30:27.666340Z digest=sha256:47b2a8e395d5f3df153c3ce253ed7eb222339cb749e40d492757f2c303cf22f4

Observation b6ffb9cd-dac4-46c0-8cec-a7c171abf7f7 · outbound

This paper cites Non-asymptotic confidence bounds for stochastic approximation algorithms with constant step size.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Non-asymptotic confidence bounds for stochastic approximation algorithms with constant step size

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:32.392834Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:30:27.774917Z digest=sha256:1e0af915fbd506ef5f0695a2e7b469bc082788b35b8d1dba5c06d44f7de7400d

Observation 359e5e04-282e-4398-89dc-05087705d8ab · outbound

This paper cites Py T orch: A n imperative style, high-performance deep learning library.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Py T orch: A n imperative style, high-performance deep learning library

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:32.241807Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:30:27.905036Z digest=sha256:a679f6089e3a7c7f4ca38dd02dd4c1dc881dcc4e73f7d5561e569dda2b51d663

Observation acf787f1-2a75-4fed-967d-d50e6342a6a0 · outbound

This paper cites https://huggingface.co/microsoft/resnet-18, 2025.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent https://huggingface.co/microsoft/resnet-18, 2025

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:32.018044Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:30:28.010649Z digest=sha256:3183e5c66760ca48b22e3ab41c1f3858677465a44764ac71014e99ba5b788985

Observation 6454f534-0c17-463f-9b38-3a1d916d3ddc · outbound

This paper cites A stochastic approximation method.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent A stochastic approximation method

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:31.783988Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:30:28.121239Z digest=sha256:f84159f734070bd21e50a2ecab9c0e2e27c74b08f313b9fa3deccc9eb528e0ef

Observation 4a60d70a-43fc-4b69-9cd4-e010c9f1012e · outbound

This paper cites https://huggingface.co/FacebookAI/roberta-base, 2025.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent https://huggingface.co/FacebookAI/roberta-base, 2025

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:31.514873Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:30:28.463148Z digest=sha256:75b137e35753491ec7606c4a330e9653a302cbe2b473a65f09ed2c22448c9fb1

Observation 7927b180-2727-410e-b59e-a6b26519e75e · outbound

This paper cites Anytime Tail Averaging.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Anytime Tail Averaging

Reference 46

Resolution
verified exact
local_arxiv, observed 2026-08-07T13:30:29.584753Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:30:28.540616Z digest=sha256:a74a19356efe739d048c6299de9a165a8250435e3cd0f6fcb5cffbc2cf6511d2

Observation dbfa0b42-1d36-4900-9839-ef3ce529f51a · outbound

This paper cites Making Gradient Descent Optimal for Strongly Convex Stochastic Optimization.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Making Gradient Descent Optimal for Strongly Convex Stochastic Optimization

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T13:30:28.585346Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:30:28.585346Z digest=sha256:74cfaefec747acd51d9fcb01ff802feb4d373ff9d979318b42d125f78486d28d

Observation 2f10f48c-f16a-44cb-941d-d5d633ef795d · outbound

This paper cites Sticking the landing: S imple, lower-variance gradient estimators for variational inference.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Sticking the landing: S imple, lower-variance gradient estimators for variational inference

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:31.347644Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:30:28.687107Z digest=sha256:096e5c046fc05c6b7c27ed2a4a9c9e651b747a4edc3468952c8a429445470dda

Observation 14442b01-b4df-40e7-b73f-8a897ade788d · outbound

This paper cites SQuAD : 100,000+ questions for machine comprehension of text.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent SQuAD : 100,000+ questions for machine comprehension of text

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:31.130714Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:30:28.737065Z digest=sha256:6c31d1fb6c454697ce82b3addd312d9cdcfe80193e9d9894ed94f500384081b3

Observation 8d42f786-343b-4ec1-8444-9576d3126183 · outbound

This paper cites Virtual library of simulation experiments: T est functions and datasets, 2013.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Virtual library of simulation experiments: T est functions and datasets, 2013

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:30.933615Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:30:28.815036Z digest=sha256:c6d17a753e288d5afa52ddc5d0e59c9bea3f84be5d4c4aca8961644784a70ca9

Observation 0acc27dd-ad8a-41c7-8cbe-df18780bf2d4 · outbound

This paper cites Recursive deep models for semantic compositionality over a sentiment treebank.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Recursive deep models for semantic compositionality over a sentiment treebank

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:30.762131Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:30:28.856405Z digest=sha256:a701c39fa04ade72b466681625cd7fd17e807d5b36874273207d0c0389b6503c

Observation b597a55d-446a-46dd-83f6-c4175b93c23b · outbound

This paper cites Stochastic gradient descent for non-smooth optimization: C onvergence results and optimal averaging schemes.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Stochastic gradient descent for non-smooth optimization: C onvergence results and optimal averaging schemes

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:30.583910Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:30:28.911294Z digest=sha256:76deb354515133f9ac48ca05968e43ff5524a9b1eaf9b9365c6dba2403262cb8

Observation f81a50cd-0413-41b8-9408-88f6d4dd081f · outbound

This paper cites Painless stochastic gradient: I nterpolation, line-search, and convergence rates.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Painless stochastic gradient: I nterpolation, line-search, and convergence rates

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:30.444768Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:30:28.941336Z digest=sha256:249c3652f1895caad2e1b9994fbda2735d70130335d0e3f4dfb9ae934f410a6f

Observation 19139d0b-134b-46c5-a154-68bf21d15add · outbound

This paper cites A framework for improving the reliability of black-box variational inference.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent A framework for improving the reliability of black-box variational inference

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:30.280853Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:30:29.073747Z digest=sha256:bde3a8e459ab5900df7ff55e46a684b2f153996d4cbb6c21ea91a4da49fe200e

Observation f258329d-7b20-498c-a848-44bc1f716d5b · outbound

This paper cites GLUE : A multi-task benchmark and analysis platform for natural language understanding.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent GLUE : A multi-task benchmark and analysis platform for natural language understanding

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:30.093008Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:30:29.175993Z digest=sha256:71ef1862a1a909b12d88c2cb08f9400d0e5fa0ff3dc7d621f6767529221aaacf

Observation 563ed8bc-d756-4a3a-a11e-8f9e3f48bb71 · outbound

This paper cites Fluctuation-dissipation relations for stochastic gradient descent.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Fluctuation-dissipation relations for stochastic gradient descent

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-07T13:30:29.286244Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:30:29.286244Z digest=sha256:604bacb07c12abfffc3fb4d35629b018d664a53d04e0f98444f319e865078a5b

Pith citing papers

No inbound Pith citation observations are available.