Pith. sign in

Paper Citation Record · LEDGER

A Lyapunov Theory for Finite-Sample Guarantees of Asynchronous Q-Learning and TD-Learning Variants

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 12 inbound Pith citation observations for arXiv:2102.01567.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2102.01567 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 12 of 12 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 12 of 12 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T16:15:15.653672Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-24T02:18:44.885808Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 886d6936-c725-4e7f-8730-96008331033e · inbound

Toward a Unified Lyapunov-Certified ODE Convergence Analysis of Smooth Q-Learning with p-Norms cites this paper.

Toward a Unified Lyapunov-Certified ODE Convergence Analysis of Smooth Q-Learning with p-Norms A Lyapunov Theory for Finite-Sample Guarantees of Asynchronous Q-Learning and TD-Learning Variants

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-24T02:18:44.890270Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-24T02:16:04.596951Z digest=sha256:cef20d4a1456d3d4f0642f1678dca01f67b7ea1e00ff54c9928f16164bcb2a39

Observation 067d16b7-c757-4fb4-ad9c-3064f461fe02 · inbound

Statistical and Algorithmic Foundations of Reinforcement Learning cites this paper.

Statistical and Algorithmic Foundations of Reinforcement Learning A Lyapunov Theory for Finite-Sample Guarantees of Asynchronous Q-Learning and TD-Learning Variants

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T16:15:15.653672Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:15:15.653672Z digest=sha256:c04f52e423299b0ff3d9011c41860bc855eefe867679dbcbfd69427fbaaca31e

Observation 504916bd-aa98-4ab5-81db-a0103d7ecc21 · inbound

Central Limit Theorems for Asynchronous Averaged Q-Learning cites this paper.

Central Limit Theorems for Asynchronous Averaged Q-Learning A Lyapunov Theory for Finite-Sample Guarantees of Asynchronous Q-Learning and TD-Learning Variants

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T14:56:30.379155Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-18T14:55:49.034505Z digest=sha256:430def62f8f986bb6854d030ee5ec118dadb37f64b46dfd2e6158268488e6810

Observation 55c7007f-6f0c-49e4-925b-76cfe072e2b1 · inbound

Sharp asymptotic theory for Q-learning with LDTZ learning rate and its generalization cites this paper.

Sharp asymptotic theory for Q-learning with LDTZ learning rate and its generalization A Lyapunov Theory for Finite-Sample Guarantees of Asynchronous Q-Learning and TD-Learning Variants

Reference 11

Resolution
unresolved
no resolver link, observed 2026-07-13T10:46:25.335756Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-13T10:46:25.335756Z digest=sha256:6b55edbdb6a10c76a3006aaa786f8ca98c238896903ced4993645a12c1dd3aff

Observation a67af2fe-7a3d-441c-9e6e-90de4947eaeb · inbound

Finite-Time Analysis of Q-Value Iteration for General-Sum Stackelberg Games cites this paper.

Finite-Time Analysis of Q-Value Iteration for General-Sum Stackelberg Games A Lyapunov Theory for Finite-Sample Guarantees of Asynchronous Q-Learning and TD-Learning Variants

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-10T22:30:50.085352Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T19:49:31.844061Z digest=sha256:f75d36b3f6a755f8261199f86964780d17b14cb4348ab3599f630d0a73b3d026

Observation 02176971-3da5-4684-95e8-c211ba9b87ed · inbound

Lyapunov-Certified Direct Switching Theory for Q-Learning cites this paper.

Lyapunov-Certified Direct Switching Theory for Q-Learning A Lyapunov Theory for Finite-Sample Guarantees of Asynchronous Q-Learning and TD-Learning Variants

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-11T13:06:02.465577Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T02:24:29.722541Z digest=sha256:e4265bb84323a2ba1d42a5376f76b548b97f276acfc5ceb268b8f491ad7149f8

Observation 2ac93f8b-a011-4a4a-8a12-3801873cf073 · inbound

Reinforcement Learning for Exponential Utility: Algorithms and Convergence in Discounted MDPs cites this paper.

Reinforcement Learning for Exponential Utility: Algorithms and Convergence in Discounted MDPs A Lyapunov Theory for Finite-Sample Guarantees of Asynchronous Q-Learning and TD-Learning Variants

Reference 45

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T03:50:57.361222Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-11T02:11:22.846629Z digest=sha256:cee475546f51f8d6f48eb9fbbb704424874efdb469b785d59f0cea621b1e8b11

Observation 0fb8ab85-932c-437f-a653-053f80fc5d41 · inbound

A Switching System Theory of Q-Learning with Linear Function Approximation cites this paper.

A Switching System Theory of Q-Learning with Linear Function Approximation A Lyapunov Theory for Finite-Sample Guarantees of Asynchronous Q-Learning and TD-Learning Variants

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-13T06:17:22.853925Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T06:16:23.809510Z digest=sha256:2a3a430781bc463243ffdd569fd72e7d772ceac1a73c5ab1d4d2bc9cd88e322f

Observation 3a59870b-a982-4202-a0ae-7432229b9f58 · inbound

A Switching System Theory of Q-Learning with Linear Function Approximation cites this paper.

A Switching System Theory of Q-Learning with Linear Function Approximation A Lyapunov Theory for Finite-Sample Guarantees of Asynchronous Q-Learning and TD-Learning Variants

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-20T22:34:09.258848Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-20T22:34:05.838427Z digest=sha256:0e03d365a0f49f673f5c38713ca17f6edd1711ae7b168f3d7fc49773a9e176b5

Observation 8805dc15-5c74-4b07-939a-21c19f122d0d · inbound

Sign-Separated Asymmetric Finite-Time Error Analysis of Q-Learning cites this paper.

Sign-Separated Asymmetric Finite-Time Error Analysis of Q-Learning A Lyapunov Theory for Finite-Sample Guarantees of Asynchronous Q-Learning and TD-Learning Variants

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-20T18:08:50.710327Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-20T18:06:25.278993Z digest=sha256:9551f541d60233c84b91fd19a1207afada08b9d9d51528f8cc90d30ca8b1b0b5

Observation f2d90220-2bda-4215-a106-3bf936689f5e · inbound

Concentration of General Stochastic Approximation Under Heavy-Tailed Markovian Noise cites this paper.

Concentration of General Stochastic Approximation Under Heavy-Tailed Markovian Noise A Lyapunov Theory for Finite-Sample Guarantees of Asynchronous Q-Learning and TD-Learning Variants

Reference 206

Resolution
metadata mismatch
arxiv_id, observed 2026-05-21T02:29:25.065553Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-21T02:27:24.989781Z digest=sha256:f357b2ed93e7c066e5b97bfb37b778ef1c7244b13ee02d671c2135e99d5889c8

Observation 76bf0df5-3a2a-401e-a94f-0243c08c8e9c · inbound

Finite-Time Analysis of Discounted Exponential-Utility Reinforcement Learning cites this paper.

Finite-Time Analysis of Discounted Exponential-Utility Reinforcement Learning A Lyapunov Theory for Finite-Sample Guarantees of Asynchronous Q-Learning and TD-Learning Variants

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-04T18:22:18.999794Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T18:22:18.999794Z digest=sha256:c15d044ffc0f931eca79dec107f17d6dc7175f04730c06cab83b53924e48bf59