Pith. sign in

Paper Citation Record · LEDGER

Exploration-exploitation trade-off for continuous-time episodic reinforcement learning with linear-convex models

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 3 inbound Pith citation observations for arXiv:2112.10264.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2112.10264 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 3 of 3 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T21:33:36.415315Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T08:19:43.712345Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 8f088e2d-74bb-468c-8c19-ee4d41f8cfcc · inbound

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems cites this paper.

Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems Exploration-exploitation trade-off for continuous-time episodic reinforcement learning with linear-convex models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T21:33:36.415315Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:33:36.415315Z digest=sha256:9ffb0e4bf4a09893653be5788d0ee4bcb6781bc2d376e5b44f7317d726bb5608

Observation 3b8eb266-6736-453f-a2a2-194c4c23cfb7 · inbound

Convergence Rates of Time Discretization in Extended Mean Field Control cites this paper.

Convergence Rates of Time Discretization in Extended Mean Field Control Exploration-exploitation trade-off for continuous-time episodic reinforcement learning with linear-convex models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-05T13:11:44.567437Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T13:11:44.567437Z digest=sha256:beb4605beb570129cd6b5e702615bfd34f4dee52e5304e89bd45282715a40bb3

Observation 67ff858d-e86e-46ef-bdcb-9c6803eaefde · inbound

PhiBE-Q-Learning: Bridging Off-Policy Reinforcement Learning and Continuous-Time Control cites this paper.

PhiBE-Q-Learning: Bridging Off-Policy Reinforcement Learning and Continuous-Time Control Exploration-exploitation trade-off for continuous-time episodic reinforcement learning with linear-convex models

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-07-04T08:19:43.714245Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-26T11:59:18.223000Z digest=sha256:cc22dd1894030c5cb1066922bf74193473ad52308eb5ddaca53473f9d1339157