Pith. sign in

Paper Citation Record · LEDGER

The Role of Coverage in Online Reinforcement Learning

As of 18 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 11 inbound Pith citation observations for arXiv:2210.04157.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2210.04157 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 11 of 11 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 11 of 11 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T23:56:02.552940Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-07T14:33:54.431062Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation b10298f2-4e93-40ef-864d-5fcf11b9e2d3 · inbound

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning cites this paper.

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning The Role of Coverage in Online Reinforcement Learning

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-12T21:48:16.928379Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:48:16.928379Z digest=sha256:20b44be7cc9fec74664320e143e42b34060d80ea537790dd38a639e91d507342

Observation c93e8ac9-9a88-4775-88cc-de57cc13401d · inbound

Improving Environment Novelty Quantification for Effective Unsupervised Environment Design cites this paper.

Improving Environment Novelty Quantification for Effective Unsupervised Environment Design The Role of Coverage in Online Reinforcement Learning

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-08T18:17:14.789543Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T18:17:14.789543Z digest=sha256:69fb78d7ce130f824968ca797a331b74dd820f141d45c3b34dc84446f48383df

Observation e5c44ab2-b78a-4c9d-ab7e-82832fd0adde · inbound

Actor-Critics Can Achieve Optimal Sample Efficiency cites this paper.

Actor-Critics Can Achieve Optimal Sample Efficiency The Role of Coverage in Online Reinforcement Learning

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-15T23:56:02.552940Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T23:56:02.552940Z digest=sha256:03ebaa09a9d749b3d7581807ea0cbf4422371da951cee927a13fb12a5eb3aa14

Observation c2a50976-2814-4905-8bed-c7b223b4a9a0 · inbound

Outcome-Based Online Reinforcement Learning: Algorithms and Fundamental Limits cites this paper.

Outcome-Based Online Reinforcement Learning: Algorithms and Fundamental Limits The Role of Coverage in Online Reinforcement Learning

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T14:10:43.178507Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:10:43.178507Z digest=sha256:a9acfedf5877d1b30130e1a8ba5af28de4283597d613ebab947c7a598fd26486

Observation 2c102604-556e-4131-b27d-df8e3dfe5770 · inbound

Square$\chi$PO: Differentially Private and Robust $\chi^2$-Preference Optimization in Offline Direct Alignment cites this paper.

Square$\chi$PO: Differentially Private and Robust $\chi^2$-Preference Optimization in Offline Direct Alignment The Role of Coverage in Online Reinforcement Learning

Reference 90

Resolution
unresolved
no resolver link, observed 2026-08-07T13:45:03.851174Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:45:03.851174Z digest=sha256:48dec80eefdc99d685219848daffb2568758007b0c78a2472105637bef72798e

Observation 3892cdd0-7f76-41af-b8f3-864fba48cbc5 · inbound

Towards Differentially Private Reinforcement Learning with General Function Approximation cites this paper.

Towards Differentially Private Reinforcement Learning with General Function Approximation The Role of Coverage in Online Reinforcement Learning

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-11T04:31:00.485398Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-11T01:15:59.533563Z digest=sha256:c87d9ef2ebf8ac2a26f54d8859049bd53eca67c728dcbf010096c6c2fe0bce6d

Observation 0181aa0f-1da9-47ca-b620-91c5f072847a · inbound

Online KL-Regularized Reinforcement Learning with Function Approximation under Misspecification cites this paper.

Online KL-Regularized Reinforcement Learning with Function Approximation under Misspecification The Role of Coverage in Online Reinforcement Learning

Reference 19

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T12:06:55.869090Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-28T02:31:11.200818Z digest=sha256:ad35a5b594c6ec32aeff3c8a8ad337b7706d2eff25e55473ef51270f4c8bbd14

Observation 652ebe5f-870a-42ed-ada9-2ab99a61889d · inbound

Online KL-Regularized Reinforcement Learning with Function Approximation under Misspecification cites this paper.

Online KL-Regularized Reinforcement Learning with Function Approximation under Misspecification The Role of Coverage in Online Reinforcement Learning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-07-14T18:23:21.126826Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T18:23:21.126826Z digest=sha256:cba3c770ef8316429c51539aa1d50ac7a0db281924f07b5e53efe746a2fabbba

Observation 23ae7914-0cfa-4e89-8e78-cca036e5460c · inbound

Quantile of Means: A Bonus-Free Ensemble Method for Minimax Optimal Reinforcement Learning cites this paper.

Quantile of Means: A Bonus-Free Ensemble Method for Minimax Optimal Reinforcement Learning The Role of Coverage in Online Reinforcement Learning

Reference 37

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T03:09:30.520219Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-06-26T18:19:02.314185Z digest=sha256:a6a85ee25a02361dbd925254579f78bb6dfedd51b94098eb6de212186c142508

Observation 8fe57352-e744-4e40-8c98-4e47f29dc806 · inbound

Fitted Occupancy-Ratio Evaluation without Bellman Completeness cites this paper.

Fitted Occupancy-Ratio Evaluation without Bellman Completeness The Role of Coverage in Online Reinforcement Learning

Reference 90

Resolution
metadata mismatch
local_arxiv, observed 2026-07-07T14:33:54.432242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-07-07T14:25:33.921937Z digest=sha256:6ea75e1df8b11a320546fb126422fb146a387320c660cc12f9aca3a8e92a82d4

Observation 3cc8a1c4-8378-4483-80d5-a2f527b6c14f · inbound

Fitted Occupancy-Ratio Evaluation without Bellman Completeness cites this paper.

Fitted Occupancy-Ratio Evaluation without Bellman Completeness The Role of Coverage in Online Reinforcement Learning

Reference 148

Resolution
unresolved
no resolver link, observed 2026-08-02T08:33:45.459641Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T08:33:45.459641Z digest=sha256:86f16aafbed3b14bdf106b375179444afa5c4f4a668da46f715b81d1fd02b675