Pith. sign in

Paper Citation Record · LEDGER

Fast Rates for the Regret of Offline Reinforcement Learning

As of 18 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 8 inbound Pith citation observations for arXiv:2102.00479.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2102.00479 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 8 of 8 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 8 of 8 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-14T04:32:28.089394Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T08:49:42.708008Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation a44a7752-1112-4cb5-bf24-a7c2b2e50c65 · inbound

Pessimism-Free Offline Learning in General-Sum Games via KL Regularization cites this paper.

Pessimism-Free Offline Learning in General-Sum Games via KL Regularization Fast Rates for the Regret of Offline Reinforcement Learning

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-11T15:21:09.640502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-09T20:10:30.920114Z digest=sha256:2df95b6071d410851fa0954a80fe362fe2c6d66394871de86c11d54d49e8646d

Observation 2f88d481-9f77-467a-80b6-938298c20b59 · inbound

Pessimism-Free Offline Learning in General-Sum Games via KL Regularization cites this paper.

Pessimism-Free Offline Learning in General-Sum Games via KL Regularization Fast Rates for the Regret of Offline Reinforcement Learning

Reference 43

Resolution
verified exact
arxiv_id, observed 2026-05-20T23:39:13.556502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-20T23:34:21.091472Z digest=sha256:2b9aac2b7e346f2028c9f7d58a350293e61a9c8393b3aeda934071edf6f3bd41

Observation 5742e6f2-989f-43a4-8dad-c904850ea9e4 · inbound

Fast Rates in $\alpha$-Potential Games via Regularized Mirror Descent cites this paper.

Fast Rates in $\alpha$-Potential Games via Regularized Mirror Descent Fast Rates for the Regret of Offline Reinforcement Learning

Reference 48

Resolution
verified exact
arxiv_id, observed 2026-05-11T15:41:39.585914Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-09T19:27:15.002190Z digest=sha256:d9e0dcad7152120ddddcf239df9ff7f843df7bac0e48e4c3280b6423a60af951

Observation 3dcbc21d-8490-4c64-9e8c-5e9b9e9c0ddc · inbound

Fast Rates in $\alpha$-Potential Games via Regularized Mirror Descent cites this paper.

Fast Rates in $\alpha$-Potential Games via Regularized Mirror Descent Fast Rates for the Regret of Offline Reinforcement Learning

Reference 53

Resolution
verified exact
arxiv_id, observed 2026-05-21T00:39:18.057716Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-21T00:38:53.592229Z digest=sha256:a2e92b964433034770272c561410d06325489f6dc14771d0781fb546f5ce6978

Observation 08d6220a-b30a-4264-9e57-2b38f4a7dddd · inbound

AstroAlertBench: Evaluating the Accuracy, Reasoning, and Honesty of Multimodal LLMs in Astronomical Classification cites this paper.

AstroAlertBench: Evaluating the Accuracy, Reasoning, and Honesty of Multimodal LLMs in Astronomical Classification Fast Rates for the Regret of Offline Reinforcement Learning

Reference 86

Resolution
verified exact
arxiv_id, observed 2026-05-11T21:31:13.839797Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T05:23:18.294600Z digest=sha256:916c9a0ea1bb1387c125b7518904b2d84767d0b54f041ab1be235cf7577e066e

Observation b4f01c68-7785-42a3-9df9-991412cbf9e9 · inbound

Offline Two-Player Zero-Sum Markov Games with KL Regularization cites this paper.

Offline Two-Player Zero-Sum Markov Games with KL Regularization Fast Rates for the Regret of Offline Reinforcement Learning

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-14T19:27:52.228522Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-14T19:24:26.326184Z digest=sha256:1ec477b2829c1e6bcfa45d57f30d2b1ab90ab44e0163abd9e65c5c018c78d182

Observation d366602d-58f9-43e9-b5ee-e58cc22fd6aa · inbound

Stationary Robust Mean-Field Games under Model Mismatches cites this paper.

Stationary Robust Mean-Field Games under Model Mismatches Fast Rates for the Regret of Offline Reinforcement Learning

Reference 233

Resolution
verified exact
arxiv_id, observed 2026-07-04T08:49:42.709364Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-06-26T10:50:40.841967Z digest=sha256:da9e9ffd187e613b384493e138b62f6afab02c66d6596670e0336d46722c31d3

Observation 27777044-4ec5-4abb-9f4a-a508d243a5d4 · inbound

A Distribution Mapping Approach to Counterfactually Fair Reinforcement Learning cites this paper.

A Distribution Mapping Approach to Counterfactually Fair Reinforcement Learning Fast Rates for the Regret of Offline Reinforcement Learning

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-14T04:32:28.089394Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:32:28.089394Z digest=sha256:01aedb4afdeb37929fec989632ad8d60235b46090c984bfb9b5e6bd5aef10873