Pith. sign in

Paper Citation Record · LEDGER

Conservative Safety Critics for Exploration

As of 15 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 10 inbound Pith citation observations for arXiv:2010.14497.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2010.14497 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 10 of 10 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 10 of 10 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-11T21:39:11.917036Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

32
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation bd1b9821-a7db-4cbd-9802-90cf01a2c2b9 · inbound

Action Mapping for Reinforcement Learning in Continuous Environments with Constraints cites this paper.

Action Mapping for Reinforcement Learning in Continuous Environments with Constraints Conservative Safety Critics for Exploration

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-11T21:39:11.917036Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T21:39:11.917036Z digest=sha256:d1ae287753fed2bdf6391506f8e03800944419751d6fb9d81105646e1a244a26

Observation 33693b59-16e0-45ff-8b2c-7ff49c725d71 · inbound

Confidence-Guided Human-AI Collaboration: Reinforcement Learning with Distributional Proxy Value Propagation for Autonomous Driving cites this paper.

Confidence-Guided Human-AI Collaboration: Reinforcement Learning with Distributional Proxy Value Propagation for Autonomous Driving Conservative Safety Critics for Exploration

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T11:06:47.251327Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:06:47.251327Z digest=sha256:ca46a104a915d79aece741e89b163d0b2d8af3fd8b8bc58eba84deab9a33748e

Observation 648534cc-54f6-48bc-be58-07809e82f975 · inbound

Safe and Performant Deployment of Autonomous Systems via Model Predictive Control and Hamilton-Jacobi Reachability Analysis cites this paper.

Safe and Performant Deployment of Autonomous Systems via Model Predictive Control and Hamilton-Jacobi Reachability Analysis Conservative Safety Critics for Exploration

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T21:50:27.426347Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:50:27.426347Z digest=sha256:ee2ea907a4d03787a9e749485130d04d322f25bc380b80437652ac9a71357829

Observation 334bd6ef-b2c9-4c2d-ba11-59d06e5ab319 · inbound

Optimistic Exploration for Risk-Averse Constrained Reinforcement Learning cites this paper.

Optimistic Exploration for Risk-Averse Constrained Reinforcement Learning Conservative Safety Critics for Exploration

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T18:16:20.283541Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:16:20.283541Z digest=sha256:9fb81ba5c0f3f3c227bfe00066d2177c0adc301fcba5bb5aae7126bcbf536189

Observation 71fc7cc8-6839-4a25-9b29-4ff614bf22f4 · inbound

Safe-Support Q-Learning: Learning without Unsafe Exploration cites this paper.

Safe-Support Q-Learning: Learning without Unsafe Exploration Conservative Safety Critics for Exploration

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T23:36:38.584131Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-07T16:36:35.746034Z digest=sha256:305b90483f5dea16fd480710284a58cff71329596f150b8afcf21d2997415dd3

Observation 6ad14b2d-ed9d-4171-877a-979f3d4c58cb · inbound

Pre-VLA: Preemptive Runtime Verification for Reliable Vision-Language-Action and World-Model Rollouts cites this paper.

Pre-VLA: Preemptive Runtime Verification for Reliable Vision-Language-Action and World-Model Rollouts Conservative Safety Critics for Exploration

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-05-22T07:04:41.680504Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-22T07:02:27.381182Z digest=sha256:f592654621fd2df200943c99117f342ce7746e2c88f0774d6af5bd6b42de11b2

Observation 00c8031e-9ab2-4d22-b2b1-83ada45ad16a · inbound

Safe Embodied AI for Long-horizon Tasks: A Cross-layer Analysis of Robotic Manipulation cites this paper.

Safe Embodied AI for Long-horizon Tasks: A Cross-layer Analysis of Robotic Manipulation Conservative Safety Critics for Exploration

Reference 79

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T12:56:56.755323Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-06-28T01:46:36.081851Z digest=sha256:738ddcb13483d1510391a4e8a3a918b647966e6ffb96da39f4b51169e5f268b5

Observation dc6f4182-b742-4d66-9fd8-d342581fa8e5 · inbound

An Agency-Transferring Model-Free Policy Enhancement Technique cites this paper.

An Agency-Transferring Model-Free Policy Enhancement Technique Conservative Safety Critics for Exploration

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-06-27T17:21:06.998970Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-06-27T17:11:45.240357Z digest=sha256:93645c89aebd790997b3b86fd50797755d64329f2777decaef2595aeff88a23a

Observation 1530175c-0c1e-4df3-8196-e8a36b27e1f0 · inbound

SHAPO: Sharpness-Aware Policy Optimization for Safe Exploration cites this paper.

SHAPO: Sharpness-Aware Policy Optimization for Safe Exploration Conservative Safety Critics for Exploration

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-07-03T00:57:29.846675Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-06-27T16:58:19.848530Z digest=sha256:74513cc1ada6c9f173da12ab8bf16a3ef6898ea2b7992877725fbd2c8d64e5c8

Observation d3db1e45-529a-400a-b05a-6e952f5c1b3d · inbound

UniIntervene: Agentic Intervention for Efficient Real-World Reinforcement Learning cites this paper.

UniIntervene: Agentic Intervention for Efficient Real-World Reinforcement Learning Conservative Safety Critics for Exploration

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-07-03T10:58:02.695906Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-06-27T09:46:59.746745Z digest=sha256:c7d5d4daa917e5c671860ad4851eaff0c423425fd8bc0746879e353eaf391028