Pith. sign in

Paper Citation Record · LEDGER

Acting in Delayed Environments with Non-Stationary Markov Policies

As of 13 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2101.11992.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2101.11992 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-11T12:27:14.712015Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T13:29:51.953490Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 13629a57-e29b-42a8-82e6-c075bb028faa · inbound

Enabling Realtime Reinforcement Learning at Scale with Staggered Asynchronous Inference cites this paper.

Enabling Realtime Reinforcement Learning at Scale with Staggered Asynchronous Inference Acting in Delayed Environments with Non-Stationary Markov Policies

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-11T12:27:14.712015Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:27:14.712015Z digest=sha256:fed6fde8c0bf00f64371df0778504c63de3a3639a7b7a8027358c7a080c006f8

Observation fa3ed8f9-681a-4d5f-b352-ee0340607649 · inbound

Model-Based Reinforcement Learning under Random Observation Delays cites this paper.

Model-Based Reinforcement Learning under Random Observation Delays Acting in Delayed Environments with Non-Stationary Markov Policies

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-18T14:36:28.716906Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-05-18T14:35:04.201252Z digest=sha256:a498e5195555224960eb3e36f9242ab215563c03b024ab917928b8e366a468e5

Observation 120fa4ab-f344-48a7-a34f-8c26b8651594 · inbound

Delayed homomorphic reinforcement learning for environments with delayed feedback cites this paper.

Delayed homomorphic reinforcement learning for environments with delayed feedback Acting in Delayed Environments with Non-Stationary Markov Policies

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-13T18:48:08.238818Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-13T18:44:26.633008Z digest=sha256:bb73171f47b9dec09d58ae5c18a846ef459b48be46c2cb0d7282fd1c243eff97

Observation 186a0298-fe68-447a-9aaf-b6a8a24dc1a0 · inbound

Finding the Time to Think: Learning Planning Budgets in Real-Time RL cites this paper.

Finding the Time to Think: Learning Planning Budgets in Real-Time RL Acting in Delayed Environments with Non-Stationary Markov Policies

Reference 7

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T13:29:51.954787Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-26T05:08:19.504454Z digest=sha256:977f005f0f308f31b4abedc6b18fdbd17f45902f5c007c9d14031d4df366ad44

Observation 432988ae-dd39-43ec-a030-5997ffabc2be · inbound

Finding the Time to Think: Learning Planning Budgets in Real-Time RL cites this paper.

Finding the Time to Think: Learning Planning Budgets in Real-Time RL Acting in Delayed Environments with Non-Stationary Markov Policies

Reference 7

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T09:34:34.892919Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-30T09:26:31.944405Z digest=sha256:98cda97820cb03996c4339a6b6552994e980eccfe5c2527fb2d4d3ef967f65e5

Observation a1d7c5c4-5126-4e25-8b64-c4bb067a0e5f · inbound

Dark matter environments and safeguards for spacetime inference from horizon scale interferometry cites this paper.

Dark matter environments and safeguards for spacetime inference from horizon scale interferometry Acting in Delayed Environments with Non-Stationary Markov Policies

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-02T03:10:33.775629Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T03:10:33.775629Z digest=sha256:04e7ef540eab38549ec21659f74ff46d6b48e244844edc9b6742d76821da82d0