Pith. sign in

Paper Citation Record · LEDGER

A Primal-Dual Approach to Constrained Markov Decision Processes

As of 23 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2101.10895.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2101.10895 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T00:06:35.700545Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-23T22:03:30.980566Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation b13a324d-039d-4ade-999f-cc86ce9fe26b · inbound

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form cites this paper.

Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form A Primal-Dual Approach to Constrained Markov Decision Processes

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-23T22:03:30.982953Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-23T21:58:56.180393Z digest=sha256:c908c0244b65746e79def672e72ebe52817d3886526a4f3202795270e1cd242c

Observation a4be79cc-0c75-4005-abbf-734c509fc2b1 · inbound

Inpatient Overflow Management with Proximal Policy Optimization cites this paper.

Inpatient Overflow Management with Proximal Policy Optimization A Primal-Dual Approach to Constrained Markov Decision Processes

Reference 7

Resolution
metadata mismatch
arxiv_id, observed 2026-05-23T19:18:20.764052Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-23T19:16:01.279311Z digest=sha256:b99164a5d9315b1124f07480329d8c69c35f323a039792119cf822a39e193a2f

Observation 4e11c989-1b4f-494d-b00d-2a18c3091576 · inbound

Offline Safe Reinforcement Learning Using Trajectory Classification cites this paper.

Offline Safe Reinforcement Learning Using Trajectory Classification A Primal-Dual Approach to Constrained Markov Decision Processes

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-11T11:30:06.654064Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T11:30:06.654064Z digest=sha256:722ba3b1c27efb86d46af885b21f4c2ebe32a34465ae79d60308833fc33a1760

Observation fb5c4472-5cdf-45f5-a56f-1691627ba70e · inbound

Primal-Dual Policy Optimization for Linear CMDPs with Adversarial Losses cites this paper.

Primal-Dual Policy Optimization for Linear CMDPs with Adversarial Losses A Primal-Dual Approach to Constrained Markov Decision Processes

Reference 36

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T01:52:06.123198Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-13T01:10:23.797409Z digest=sha256:655ab71cbe050a8efad9dc74a79e36d117b591e3da7384c5bb41ae917f3ba672

Observation 0567988d-bcb1-4d76-9177-105a543c1c6d · inbound

Policy Optimization in Hybrid Discrete-Continuous Action Spaces via Mixed Gradients cites this paper.

Policy Optimization in Hybrid Discrete-Continuous Action Spaces via Mixed Gradients A Primal-Dual Approach to Constrained Markov Decision Processes

Reference 87

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T02:13:30.123193Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-15T02:12:58.603012Z digest=sha256:64fe5b1c08120dda2ad5d7cfe310fa43472ef75ffd82201277ed42e59898e6fe

Observation 5c9b4304-a0c8-4ee0-9174-0d77e3fe0465 · inbound

Decentralized Multi-Player Q-Learning in Episodic Markov Decision Processes with Information Asymmetry cites this paper.

Decentralized Multi-Player Q-Learning in Episodic Markov Decision Processes with Information Asymmetry A Primal-Dual Approach to Constrained Markov Decision Processes

Reference 107

Resolution
unresolved
no resolver link, observed 2026-08-16T00:06:35.700545Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:06:35.700545Z digest=sha256:2728699e5a1e686c3d83ee3cfdbcbed6209ae1e0ad158b3a7842e0143f30a694