Pith. sign in

Paper Citation Record · LEDGER

Constant Stepsize Q-learning: Distributional Convergence, Bias and Extrapolation

As of 13 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 10 inbound Pith citation observations for arXiv:2401.13884.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2401.13884 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 10 of 10 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 10 of 10 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-13T10:46:25.335756Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T13:05:45.530664Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 7059be44-e7d7-47d7-b37c-ac070274758a · inbound

From Set Convergence to Pointwise Convergence: Finite-Time Guarantees for Average-Reward Q-Learning with Adaptive Stepsizes cites this paper.

From Set Convergence to Pointwise Convergence: Finite-Time Guarantees for Average-Reward Q-Learning with Adaptive Stepsizes Constant Stepsize Q-learning: Distributional Convergence, Bias and Extrapolation

Reference 87

Resolution
verified exact
arxiv_id, observed 2026-05-22T17:35:00.842255Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-22T17:34:49.191496Z digest=sha256:a7d5fbe3bd547ec5d1d288881ffff2366f0717e0b97875a983e44a6ffd23fbd6

Observation ea9749e5-69e4-4847-9ef0-ed364951065e · inbound

Central Limit Theorems for Asynchronous Averaged Q-Learning cites this paper.

Central Limit Theorems for Asynchronous Averaged Q-Learning Constant Stepsize Q-learning: Distributional Convergence, Bias and Extrapolation

Reference 12

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T14:56:30.395591Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-18T14:55:49.034505Z digest=sha256:e9cb1d442612d4e60d65ea1605584977aa696d3c2423457fa93a16e9004edcd7

Observation 97155cf2-5cfc-4627-92be-fd8d4a93244e · inbound

A Minimal-Assumption Analysis of Q-Learning with Time-Varying Policies cites this paper.

A Minimal-Assumption Analysis of Q-Learning with Time-Varying Policies Constant Stepsize Q-learning: Distributional Convergence, Bias and Extrapolation

Reference 38

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T05:50:57.154573Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-18T05:47:47.782246Z digest=sha256:76e4687c5dfb4fd14d36ab242b26f6a2b424068b15be5e4e37c5661705f00b76

Observation 138a38e0-9a1b-42ef-8472-7956e3581bb7 · inbound

Wasserstein-p Central Limit Theorem Rates: From Local Dependence to Markov Chains cites this paper.

Wasserstein-p Central Limit Theorem Rates: From Local Dependence to Markov Chains Constant Stepsize Q-learning: Distributional Convergence, Bias and Extrapolation

Reference 76

Resolution
malformed identifier
arxiv_id, observed 2026-05-16T15:23:02.079774Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-16T15:21:02.872523Z digest=sha256:43029c46cf315f1adf6d4fd49b7759a7b6f7c83d40c0362c6e9e8c3b3ef4e27e

Observation 82258d0c-a47d-4238-99e1-9b2d32e949e7 · inbound

Sharp asymptotic theory for Q-learning with LDTZ learning rate and its generalization cites this paper.

Sharp asymptotic theory for Q-learning with LDTZ learning rate and its generalization Constant Stepsize Q-learning: Distributional Convergence, Bias and Extrapolation

Reference 75

Resolution
unresolved
no resolver link, observed 2026-07-13T10:46:25.335756Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-13T10:46:25.335756Z digest=sha256:adb570de360a5093539c4ba42d7964b0402b661c0b9921e2ff2c1989b06bb8da

Observation 93b7c764-836c-4c16-80d8-d05af1def6a2 · inbound

Shuffling the Data, Stretching the Step-size: Sharper Bias in constant step-size SGD cites this paper.

Shuffling the Data, Stretching the Step-size: Sharper Bias in constant step-size SGD Constant Stepsize Q-learning: Distributional Convergence, Bias and Extrapolation

Reference 141

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T11:01:03.712977Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-10T15:13:59.395799Z digest=sha256:966efcd11891c9d29f881c849793aba5cebc51b147be89e784bfb7abb7738076

Observation 7d748d63-82fb-49f6-8a9b-afccb39e68aa · inbound

Revisiting the Constant Stepsize Stochastic Approximation with Decision-Dependent Markovian Noise cites this paper.

Revisiting the Constant Stepsize Stochastic Approximation with Decision-Dependent Markovian Noise Constant Stepsize Q-learning: Distributional Convergence, Bias and Extrapolation

Reference 32

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T13:45:27.427927Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-10T13:45:04.730863Z digest=sha256:cdcc795df5496d7c9e21aa04139b200af11ff4d4779480242332430c951e4b0c

Observation cca55c52-38fc-4b2f-8279-dea5dc13086f · inbound

Elephant random walk with attributed steps and extractions of random sizes cites this paper.

Elephant random walk with attributed steps and extractions of random sizes Constant Stepsize Q-learning: Distributional Convergence, Bias and Extrapolation

Reference 60

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T06:11:20.120298Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-10T06:09:29.594763Z digest=sha256:8ff76c91536ca1140dfdda1c293072b59c566241985d6d5608f61a89741d0edc

Observation a9ac911b-bb77-45f5-85b8-1191a1b91f4c · inbound

Gaussian Approximation and Multiplier Bootstrap for Federated Linear Stochastic Approximation cites this paper.

Gaussian Approximation and Multiplier Bootstrap for Federated Linear Stochastic Approximation Constant Stepsize Q-learning: Distributional Convergence, Bias and Extrapolation

Reference 30

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T02:12:58.456482Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-20T02:10:02.167114Z digest=sha256:606218771dd64469174685a6f535c42300b673b5fea1ef43e7bf46b25aa71746

Observation 429ac6e2-176d-4cb7-80e2-78b514eb8ed3 · inbound

SGD at the Edge of Stability: Stochastic Stabilization with Large Learning Rates cites this paper.

SGD at the Edge of Stability: Stochastic Stabilization with Large Learning Rates Constant Stepsize Q-learning: Distributional Convergence, Bias and Extrapolation

Reference 221

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T13:05:45.532477Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-07-01T01:05:17.842447Z digest=sha256:861e84c66b9551b5c140ce4409e59347f717764cdca11501d987c946a7b98c32