Pith. sign in

Paper Citation Record · LEDGER

A Lyapunov Theory for Finite-Sample Guarantees of Asynchronous Q-Learning and TD-Learning Variants

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 12 inbound Pith citation observations for arXiv:2102.01567.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2102.01567 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 12 of 12 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 12 of 12 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T16:15:15.653672Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-24T02:18:44.885808Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 886d6936-c725-4e7f-8730-96008331033e · inbound

Toward a Unified Lyapunov-Certified ODE Convergence Analysis of Smooth Q-Learning with p-Norms cites this paper.

Toward a Unified Lyapunov-Certified ODE Convergence Analysis of Smooth Q-Learning with p-Norms A Lyapunov Theory for Finite-Sample Guarantees of Asynchronous Q-Learning and TD-Learning Variants

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-24T02:18:44.890270Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-24T02:16:04.596951Z digest=sha256:7c9cb6194e51bb45073fe351b5d311cc56c8365d659c1f4f211249432ad57de8

Observation 067d16b7-c757-4fb4-ad9c-3064f461fe02 · inbound

Statistical and Algorithmic Foundations of Reinforcement Learning cites this paper.

Statistical and Algorithmic Foundations of Reinforcement Learning A Lyapunov Theory for Finite-Sample Guarantees of Asynchronous Q-Learning and TD-Learning Variants

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.653672Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.653672Z digest=sha256:c04f52e423299b0ff3d9011c41860bc855eefe867679dbcbfd69427fbaaca31e

Observation 504916bd-aa98-4ab5-81db-a0103d7ecc21 · inbound

Central Limit Theorems for Asynchronous Averaged Q-Learning cites this paper.

Central Limit Theorems for Asynchronous Averaged Q-Learning A Lyapunov Theory for Finite-Sample Guarantees of Asynchronous Q-Learning and TD-Learning Variants

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T14:56:30.379155Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-18T14:55:49.034505Z digest=sha256:7dbad89ccec453d08cd2490a5cb99154273ca3ae04fcf240f691ad69f7950e0b

Observation 55c7007f-6f0c-49e4-925b-76cfe072e2b1 · inbound

Sharp asymptotic theory for Q-learning with LDTZ learning rate and its generalization cites this paper.

Sharp asymptotic theory for Q-learning with LDTZ learning rate and its generalization A Lyapunov Theory for Finite-Sample Guarantees of Asynchronous Q-Learning and TD-Learning Variants

Reference 11

Resolution
unresolved
no resolver link, observed 2026-07-13T10:46:25.335756Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-13T10:46:25.335756Z digest=sha256:6b55edbdb6a10c76a3006aaa786f8ca98c238896903ced4993645a12c1dd3aff

Observation a67af2fe-7a3d-441c-9e6e-90de4947eaeb · inbound

Finite-Time Analysis of Q-Value Iteration for General-Sum Stackelberg Games cites this paper.

Finite-Time Analysis of Q-Value Iteration for General-Sum Stackelberg Games A Lyapunov Theory for Finite-Sample Guarantees of Asynchronous Q-Learning and TD-Learning Variants

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-10T22:30:50.085352Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T19:49:31.844061Z digest=sha256:7f943c9d5f6205005d9548a97a9513bb28bb1afd2e11ddc52fc923ca92c3f74e

Observation 02176971-3da5-4684-95e8-c211ba9b87ed · inbound

Lyapunov-Certified Direct Switching Theory for Q-Learning cites this paper.

Lyapunov-Certified Direct Switching Theory for Q-Learning A Lyapunov Theory for Finite-Sample Guarantees of Asynchronous Q-Learning and TD-Learning Variants

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-11T13:06:02.465577Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T02:24:29.722541Z digest=sha256:c0ae979713a5e8bbf8fd71020581586e0304f6a49f3257501b0bb1d3a47875c0

Observation 2ac93f8b-a011-4a4a-8a12-3801873cf073 · inbound

Reinforcement Learning for Exponential Utility: Algorithms and Convergence in Discounted MDPs cites this paper.

Reinforcement Learning for Exponential Utility: Algorithms and Convergence in Discounted MDPs A Lyapunov Theory for Finite-Sample Guarantees of Asynchronous Q-Learning and TD-Learning Variants

Reference 45

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T03:50:57.361222Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-11T02:11:22.846629Z digest=sha256:8fab2b685757261b85a744610f0c134136be35bf4a7a2472365ced3f4fe9478d

Observation 0fb8ab85-932c-437f-a653-053f80fc5d41 · inbound

A Switching System Theory of Q-Learning with Linear Function Approximation cites this paper.

A Switching System Theory of Q-Learning with Linear Function Approximation A Lyapunov Theory for Finite-Sample Guarantees of Asynchronous Q-Learning and TD-Learning Variants

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-13T06:17:22.853925Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-13T06:16:23.809510Z digest=sha256:484a21e1b0d24f211ecc2617261c917080e82eb250a2cd110e6f7cf026665761

Observation 3a59870b-a982-4202-a0ae-7432229b9f58 · inbound

A Switching System Theory of Q-Learning with Linear Function Approximation cites this paper.

A Switching System Theory of Q-Learning with Linear Function Approximation A Lyapunov Theory for Finite-Sample Guarantees of Asynchronous Q-Learning and TD-Learning Variants

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-20T22:34:09.258848Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-20T22:34:05.838427Z digest=sha256:ddcae397064d56c6bbc3c9351d3c673f90d079e54a9f4b7389f84dc33ff712b9

Observation 8805dc15-5c74-4b07-939a-21c19f122d0d · inbound

Sign-Separated Asymmetric Finite-Time Error Analysis of Q-Learning cites this paper.

Sign-Separated Asymmetric Finite-Time Error Analysis of Q-Learning A Lyapunov Theory for Finite-Sample Guarantees of Asynchronous Q-Learning and TD-Learning Variants

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-20T18:08:50.710327Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-20T18:06:25.278993Z digest=sha256:3dcdc76268dcc64ea097a85c84258c96d6b0c073e95ff0225aa7b8311d141a30

Observation f2d90220-2bda-4215-a106-3bf936689f5e · inbound

Concentration of General Stochastic Approximation Under Heavy-Tailed Markovian Noise cites this paper.

Concentration of General Stochastic Approximation Under Heavy-Tailed Markovian Noise A Lyapunov Theory for Finite-Sample Guarantees of Asynchronous Q-Learning and TD-Learning Variants

Reference 206

Resolution
metadata mismatch
arxiv_id, observed 2026-05-21T02:29:25.065553Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-21T02:27:24.989781Z digest=sha256:2e296fe0d11fe9e844885e084926bc5d4ea192e443e060a73b1bbe2fa5ac5f83

Observation 76bf0df5-3a2a-401e-a94f-0243c08c8e9c · inbound

Finite-Time Analysis of Discounted Exponential-Utility Reinforcement Learning cites this paper.

Finite-Time Analysis of Discounted Exponential-Utility Reinforcement Learning A Lyapunov Theory for Finite-Sample Guarantees of Asynchronous Q-Learning and TD-Learning Variants

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-04T18:22:18.999794Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T18:22:18.999794Z digest=sha256:c15d044ffc0f931eca79dec107f17d6dc7175f04730c06cab83b53924e48bf59