Pith. sign in

Paper Citation Record · LEDGER

Projection-Based Constrained Policy Optimization

As of 16 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 19 inbound Pith citation observations for arXiv:2010.03152.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2010.03152 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 19 of 19 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 19 of 19 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T20:01:08.588304Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T08:49:42.827587Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 57d2901b-701d-40f7-88a2-c35fb3078c04 · inbound

Q-learning-based Model-free Safety Filter cites this paper.

Q-learning-based Model-free Safety Filter Projection-Based Constrained Policy Optimization

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-12T05:55:14.664113Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:55:14.664113Z digest=sha256:1e9bc45d2a0a603e6d7581e5516c70d5b223c866b0884213958da730b8352f6d

Observation 4d7c7c40-858a-4ad7-90ca-42aef69426f9 · inbound

From Text to Trajectory: Exploring Complex Constraint Representation and Decomposition in Safe Reinforcement Learning cites this paper.

From Text to Trajectory: Exploring Complex Constraint Representation and Decomposition in Safe Reinforcement Learning Projection-Based Constrained Policy Optimization

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-11T17:32:14.336803Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:32:14.336803Z digest=sha256:b1b59852b8219cf5bd8c90d717bfa0e7cd39634b1587f6811816be98af64fdc8

Observation f7d4c6dd-6a6a-4971-9ab3-bf5fcb76046f · inbound

Safe Reinforcement Learning using Finite-Horizon Gradient-based Estimation cites this paper.

Safe Reinforcement Learning using Finite-Horizon Gradient-based Estimation Projection-Based Constrained Policy Optimization

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-11T15:20:56.540973Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T15:20:56.540973Z digest=sha256:b1dd057a2ab6fc3502d3be3bec1ce36397055c9b5ce40ad45530174389d8d176

Observation 0d99bba0-a693-46f3-9ef9-bf4d1c4f5b14 · inbound

Tilted Quantile Gradient Updates for Quantile-Constrained Reinforcement Learning cites this paper.

Tilted Quantile Gradient Updates for Quantile-Constrained Reinforcement Learning Projection-Based Constrained Policy Optimization

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-11T13:25:20.519329Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T13:25:20.519329Z digest=sha256:f1660981d9a2e12a154858e50f089f4af9cd7ddd225f3a9502a11db3cc5946af

Observation d9ca5f64-4792-4ad5-8b81-cca65e75d970 · inbound

Operator Splitting for Convex Constrained Markov Decision Processes cites this paper.

Operator Splitting for Convex Constrained Markov Decision Processes Projection-Based Constrained Policy Optimization

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-11T12:40:45.599554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:40:45.599554Z digest=sha256:8e714b424a71dfb025f91dfdf1a61b07cca5692b694e42c5529c95ed2397fadf

Observation 311564d5-b7e6-4960-aa54-f66b851479f6 · inbound

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning cites this paper.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Projection-Based Constrained Policy Optimization

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-15T20:01:08.588304Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T20:01:08.588304Z digest=sha256:0d2467f111e4ab2e581579f8b3fbb9b9af86c6c96a45544edfbafac6ea9a8cf3

Observation c45262ad-fe4e-41fd-b2cf-23cf9243c6ec · inbound

PNAct: Crafting Backdoor Attacks in Safe Reinforcement Learning cites this paper.

PNAct: Crafting Backdoor Attacks in Safe Reinforcement Learning Projection-Based Constrained Policy Optimization

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-06T21:20:29.464007Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:20:29.464007Z digest=sha256:512ab539be748782eb9bfac38c31198212b97856f9aab75b3837bd95346630a1

Observation 69baed02-188c-4f37-84b3-29394e97d553 · inbound

Proactive Constrained Policy Optimization with Preemptive Penalty cites this paper.

Proactive Constrained Policy Optimization with Preemptive Penalty Projection-Based Constrained Policy Optimization

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T05:26:39.944282Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T05:26:39.944282Z digest=sha256:d837c31f432289b97619b1ab58a58b4e975d9243257c380d24b8ef7bde2a4448

Observation 5bd81fc8-9bc4-4a4e-ba30-9110a2ae3d83 · inbound

Towards Adaptive External Communication in Autonomous Vehicles: A Conceptual Design Framework cites this paper.

Towards Adaptive External Communication in Autonomous Vehicles: A Conceptual Design Framework Projection-Based Constrained Policy Optimization

Reference 145

Resolution
unresolved
no resolver link, observed 2026-08-05T19:28:42.568008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:28:42.568008Z digest=sha256:db1d588b7884ba1efae6e1f84c4bd9c823c98db1134d1a19a16c21968251816f

Observation 54f2db2d-3e3b-4647-bc16-4aba70695daa · inbound

HAEPO: History-Aggregated Exploratory Policy Optimization cites this paper.

HAEPO: History-Aggregated Exploratory Policy Optimization Projection-Based Constrained Policy Optimization

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-05T16:13:08.329208Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:13:08.329208Z digest=sha256:ec407bd67f8e7404d7cee1997cd51cb305436e44af9de2b13db879588a904839

Observation 9d66733f-0c55-4b67-b2d9-06229e753e3c · inbound

Learning Reachability of Energy Storage Arbitrage cites this paper.

Learning Reachability of Energy Storage Arbitrage Projection-Based Constrained Policy Optimization

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-17T00:11:23.314646Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-17T00:09:57.342909Z digest=sha256:05ca784033550495b3af818497519ea0e6d00999e8b60d9d81cd18d3e5220bea

Observation 74cbb9dd-b34a-4e9a-a670-d1f326ae747b · inbound

Shaping Zero-Shot Coordination via State Blocking cites this paper.

Shaping Zero-Shot Coordination via State Blocking Projection-Based Constrained Policy Optimization

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-13T07:42:30.202198Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-13T07:42:23.260464Z digest=sha256:7fa7e061b61948b7e0af39addf80310e4a9f43750d14ca3bd40e143b01bd7560

Observation 3707cb9f-bf1c-4307-94ab-185a1bb1573f · inbound

Stochastic Minimum-Cost Reach-Avoid Reinforcement Learning cites this paper.

Stochastic Minimum-Cost Reach-Avoid Reinforcement Learning Projection-Based Constrained Policy Optimization

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-13T07:32:30.563478Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-13T07:28:24.455817Z digest=sha256:0279407a66e48b13812740dabc9fb6127541098adeaf44117a80528b52d38232

Observation 377cdaa2-8b56-4df6-8339-484b1c0cc31d · inbound

Stochastic Minimum-Cost Reach-Avoid Reinforcement Learning cites this paper.

Stochastic Minimum-Cost Reach-Avoid Reinforcement Learning Projection-Based Constrained Policy Optimization

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-20T22:49:10.091288Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-20T22:48:55.661356Z digest=sha256:a87fc0e4cf9fa1c1d8566fa83a524fd6b39507cd0dcaff5c75a87dea02adcb85

Observation b65ad3bb-b558-49a0-a9b5-dbc898cded5a · inbound

Safety-Constrained Reinforcement Learning with Post-Training Reachability Verification for Robot Navigation cites this paper.

Safety-Constrained Reinforcement Learning with Post-Training Reachability Verification for Robot Navigation Projection-Based Constrained Policy Optimization

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-15T04:55:03.697736Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-15T04:52:48.491036Z digest=sha256:86e5ca7cb13af5e5750b67607d9e44884f32767f89574cd485158b85e4ed4de2

Observation 5302c802-cf66-4d30-9860-41e11ba0c694 · inbound

Action-Conditioned Risk Gating for Safety-Critical Control under Partial Observability cites this paper.

Action-Conditioned Risk Gating for Safety-Critical Control under Partial Observability Projection-Based Constrained Policy Optimization

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-05-15T02:53:33.836749Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-15T02:50:40.285132Z digest=sha256:07ea405a74863ed6cec4424099d2a28b223870c858dce30bd552672eea9726f3

Observation f72318d1-fc15-43de-9c53-e04372183f8b · inbound

Stationary Robust Mean-Field Games under Model Mismatches cites this paper.

Stationary Robust Mean-Field Games under Model Mismatches Projection-Based Constrained Policy Optimization

Reference 193

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T08:49:42.828921Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-06-26T10:50:40.841967Z digest=sha256:119b3b1249616a2dc50fe2120ebd020b798173263129b7c61f9a5c3bd56b56f7

Observation 72aa2f1c-96ae-495b-a307-7c421392e84d · inbound

PPO-EAL: Exact Augmented Lagrangian Proximal Policy Optimization for Safe Robotic Control cites this paper.

PPO-EAL: Exact Augmented Lagrangian Proximal Policy Optimization for Safe Robotic Control Projection-Based Constrained Policy Optimization

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-06-29T19:13:53.387161Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-06-29T04:49:22.328169Z digest=sha256:259a33f00fb8c880c8899a648de9ffc1d4ff8ab818a8a97bbc0b80c7ccd88ece

Observation 358b8717-05b9-4301-89fc-6bd17154ec29 · inbound

Towards Value-Constrained Credit Assignment in Fully Delegated AI Cooperatives cites this paper.

Towards Value-Constrained Credit Assignment in Fully Delegated AI Cooperatives Projection-Based Constrained Policy Optimization

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T16:55:51.013454Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-06-29T04:27:10.983662Z digest=sha256:fdda008cf19883c9f6a00632e4aacdee25547bc00f24e0d64757e368c53348af