Pith. sign in

Paper Citation Record · LEDGER

Projection-Based Constrained Policy Optimization

As of 16 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 19 inbound Pith citation observations for arXiv:2010.03152.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2010.03152 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 19 of 19 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 19 of 19 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T20:01:08.588304Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T08:49:42.827587Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 57d2901b-701d-40f7-88a2-c35fb3078c04 · inbound

Q-learning-based Model-free Safety Filter cites this paper.

Q-learning-based Model-free Safety Filter Projection-Based Constrained Policy Optimization

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-12T05:55:14.664113Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:55:14.664113Z digest=sha256:fc274943a54e5ec795934cb54b767cbb6f429f6b5fa716980bc25a5ed8dcec1f

Observation 4d7c7c40-858a-4ad7-90ca-42aef69426f9 · inbound

From Text to Trajectory: Exploring Complex Constraint Representation and Decomposition in Safe Reinforcement Learning cites this paper.

From Text to Trajectory: Exploring Complex Constraint Representation and Decomposition in Safe Reinforcement Learning Projection-Based Constrained Policy Optimization

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-11T17:32:14.336803Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:32:14.336803Z digest=sha256:b4d6f70a6c367e2c671b69a99c0c79c4639811f560b3341f2df77e7a57b28f3c

Observation f7d4c6dd-6a6a-4971-9ab3-bf5fcb76046f · inbound

Safe Reinforcement Learning using Finite-Horizon Gradient-based Estimation cites this paper.

Safe Reinforcement Learning using Finite-Horizon Gradient-based Estimation Projection-Based Constrained Policy Optimization

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-11T15:20:56.540973Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T15:20:56.540973Z digest=sha256:7fcbced529d583a398db6569b82cdc45b909348f3081b37ae619bec5e40728ca

Observation 0d99bba0-a693-46f3-9ef9-bf4d1c4f5b14 · inbound

Tilted Quantile Gradient Updates for Quantile-Constrained Reinforcement Learning cites this paper.

Tilted Quantile Gradient Updates for Quantile-Constrained Reinforcement Learning Projection-Based Constrained Policy Optimization

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-11T13:25:20.519329Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T13:25:20.519329Z digest=sha256:2ce0b530ad860d30494774c2f294048f395f355054a25f647cd247ce40356c15

Observation d9ca5f64-4792-4ad5-8b81-cca65e75d970 · inbound

Operator Splitting for Convex Constrained Markov Decision Processes cites this paper.

Operator Splitting for Convex Constrained Markov Decision Processes Projection-Based Constrained Policy Optimization

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-11T12:40:45.599554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:40:45.599554Z digest=sha256:5afa0b24498e39d6ffeaa2fd00bd21fa17d278406e4811cf276aca14fd3edd64

Observation 311564d5-b7e6-4960-aa54-f66b851479f6 · inbound

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning cites this paper.

Situational-Constrained Sequential Resources Allocation via Reinforcement Learning Projection-Based Constrained Policy Optimization

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-15T20:01:08.588304Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T20:01:08.588304Z digest=sha256:aa402a33c98569014d790dfac823463365462cabbfb27fa8eddf707efbdf9ced

Observation c45262ad-fe4e-41fd-b2cf-23cf9243c6ec · inbound

PNAct: Crafting Backdoor Attacks in Safe Reinforcement Learning cites this paper.

PNAct: Crafting Backdoor Attacks in Safe Reinforcement Learning Projection-Based Constrained Policy Optimization

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-06T21:20:29.464007Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:20:29.464007Z digest=sha256:e78a6110af04cf366c702b071b19ed41b5224dbbe224b023c16325d482df4d6c

Observation 69baed02-188c-4f37-84b3-29394e97d553 · inbound

Proactive Constrained Policy Optimization with Preemptive Penalty cites this paper.

Proactive Constrained Policy Optimization with Preemptive Penalty Projection-Based Constrained Policy Optimization

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T05:26:39.944282Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T05:26:39.944282Z digest=sha256:c03419ba6e124532f27723bed15af7f649d9b9ef504388e3f4dc7e68e889e2de

Observation 5bd81fc8-9bc4-4a4e-ba30-9110a2ae3d83 · inbound

Towards Adaptive External Communication in Autonomous Vehicles: A Conceptual Design Framework cites this paper.

Towards Adaptive External Communication in Autonomous Vehicles: A Conceptual Design Framework Projection-Based Constrained Policy Optimization

Reference 145

Resolution
unresolved
no resolver link, observed 2026-08-05T19:28:42.568008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:28:42.568008Z digest=sha256:66a8b616b14820a84c71a64905fd194d4e5aaf026545f858425d317f62008f3e

Observation 54f2db2d-3e3b-4647-bc16-4aba70695daa · inbound

HAEPO: History-Aggregated Exploratory Policy Optimization cites this paper.

HAEPO: History-Aggregated Exploratory Policy Optimization Projection-Based Constrained Policy Optimization

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-05T16:13:08.329208Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:13:08.329208Z digest=sha256:68185311fb1759590154c20078cf066fd3440ae2415c272ca0cfbc673cf68232

Observation 9d66733f-0c55-4b67-b2d9-06229e753e3c · inbound

Learning Reachability of Energy Storage Arbitrage cites this paper.

Learning Reachability of Energy Storage Arbitrage Projection-Based Constrained Policy Optimization

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-17T00:11:23.314646Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-17T00:09:57.342909Z digest=sha256:1661b1e41fabb1770ee35764458800462a499d7045bbe89f56fe7258675d4593

Observation 74cbb9dd-b34a-4e9a-a670-d1f326ae747b · inbound

Shaping Zero-Shot Coordination via State Blocking cites this paper.

Shaping Zero-Shot Coordination via State Blocking Projection-Based Constrained Policy Optimization

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-13T07:42:30.202198Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-13T07:42:23.260464Z digest=sha256:010e15ded7645eab6fce9074c540a8853f25fdb400bf120e46ee0a6dea4bb0bc

Observation 3707cb9f-bf1c-4307-94ab-185a1bb1573f · inbound

Stochastic Minimum-Cost Reach-Avoid Reinforcement Learning cites this paper.

Stochastic Minimum-Cost Reach-Avoid Reinforcement Learning Projection-Based Constrained Policy Optimization

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-13T07:32:30.563478Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-13T07:28:24.455817Z digest=sha256:e9350e3f95ae80b1c416eb2522690d8d24ebd2252355269cc0d3b0a024d32ff4

Observation 377cdaa2-8b56-4df6-8339-484b1c0cc31d · inbound

Stochastic Minimum-Cost Reach-Avoid Reinforcement Learning cites this paper.

Stochastic Minimum-Cost Reach-Avoid Reinforcement Learning Projection-Based Constrained Policy Optimization

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-20T22:49:10.091288Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-20T22:48:55.661356Z digest=sha256:b017c2eb9a0dc49ecb4d6a5267fa531675574f236a2e4b9f53f100de43309c60

Observation b65ad3bb-b558-49a0-a9b5-dbc898cded5a · inbound

Safety-Constrained Reinforcement Learning with Post-Training Reachability Verification for Robot Navigation cites this paper.

Safety-Constrained Reinforcement Learning with Post-Training Reachability Verification for Robot Navigation Projection-Based Constrained Policy Optimization

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-15T04:55:03.697736Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-15T04:52:48.491036Z digest=sha256:6fbf4980101561e84dfa90719ba68a8ff30c33c8a0cced3c1f79dbc1434c8500

Observation 5302c802-cf66-4d30-9860-41e11ba0c694 · inbound

Action-Conditioned Risk Gating for Safety-Critical Control under Partial Observability cites this paper.

Action-Conditioned Risk Gating for Safety-Critical Control under Partial Observability Projection-Based Constrained Policy Optimization

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-05-15T02:53:33.836749Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-15T02:50:40.285132Z digest=sha256:a5308e3c624483ccee8ed639aa4f9ee414ac1483320498f14bcadd47c7cd1a56

Observation f72318d1-fc15-43de-9c53-e04372183f8b · inbound

Stationary Robust Mean-Field Games under Model Mismatches cites this paper.

Stationary Robust Mean-Field Games under Model Mismatches Projection-Based Constrained Policy Optimization

Reference 193

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T08:49:42.828921Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-06-26T10:50:40.841967Z digest=sha256:ddc0ef090fbb4403d48b387db09988d344adbb1ff894218d68f6c9fd33f04440

Observation 72aa2f1c-96ae-495b-a307-7c421392e84d · inbound

PPO-EAL: Exact Augmented Lagrangian Proximal Policy Optimization for Safe Robotic Control cites this paper.

PPO-EAL: Exact Augmented Lagrangian Proximal Policy Optimization for Safe Robotic Control Projection-Based Constrained Policy Optimization

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-06-29T19:13:53.387161Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-06-29T04:49:22.328169Z digest=sha256:03948127a656b2a9b6e7740213b0bc2678347c9cc5ad6b244928f279c06576e9

Observation 358b8717-05b9-4301-89fc-6bd17154ec29 · inbound

Towards Value-Constrained Credit Assignment in Fully Delegated AI Cooperatives cites this paper.

Towards Value-Constrained Credit Assignment in Fully Delegated AI Cooperatives Projection-Based Constrained Policy Optimization

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T16:55:51.013454Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-06-29T04:27:10.983662Z digest=sha256:583cd89584de506c781c27850db3f5ebca7c9371c254b1634f09504fe8e43777