Pith. sign in

Paper Citation Record · LEDGER

Datasets and Benchmarks for Offline Safe Reinforcement Learning

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 10 inbound Pith citation observations for arXiv:2306.09303.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2306.09303 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 10 of 10 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 10 of 10 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-09T13:03:17.975720Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T00:07:28.404423Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation a3b4e8e4-3f9d-43d7-93f4-6ba9370317c8 · inbound

From Uncertain to Safe: Conformal Adaptation of Diffusion Models for Safe PDE Control cites this paper.

From Uncertain to Safe: Conformal Adaptation of Diffusion Models for Safe PDE Control Datasets and Benchmarks for Offline Safe Reinforcement Learning

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-09T13:03:17.975720Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T13:03:17.975720Z digest=sha256:5e7cd7ff24f05c10e166cf19058842abfc8dafcbf4f22c84f0197856e5a036ba

Observation 97092806-0863-493d-be61-7e200cb4bf3e · inbound

Skill Expansion and Composition in Parameter Space cites this paper.

Skill Expansion and Composition in Parameter Space Datasets and Benchmarks for Offline Safe Reinforcement Learning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-08T17:26:09.304205Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:26:09.304205Z digest=sha256:a009ca2025cf86d1df096e6b94ef570babb190c1660f1228ac3755a52ed0e273

Observation 55ce338c-309c-42ad-aca1-30454c431300 · inbound

PyTupli: A Scalable Infrastructure for Collaborative Offline Reinforcement Learning Projects cites this paper.

PyTupli: A Scalable Infrastructure for Collaborative Offline Reinforcement Learning Projects Datasets and Benchmarks for Offline Safe Reinforcement Learning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T14:58:01.250621Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:58:01.250621Z digest=sha256:f415c48dec6916eb3637f1f93ccd0d52f35506a40be5a6dcad4974069f55f09a

Observation f0751b9c-ed33-46ac-bf04-49112ec2ab7d · inbound

A Provable Approach for End-to-End Safe Reinforcement Learning cites this paper.

A Provable Approach for End-to-End Safe Reinforcement Learning Datasets and Benchmarks for Offline Safe Reinforcement Learning

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T13:30:23.525259Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:30:23.525259Z digest=sha256:e96493e6bff3a803be8c60e88181975cd28813f48242b10e29430cfaa4b64bd3

Observation 1057e0d8-bc57-497c-8d62-2ba660bb9523 · inbound

Decoupled Guidance Diffusion for Adaptive Offline Safe Reinforcement Learning cites this paper.

Decoupled Guidance Diffusion for Adaptive Offline Safe Reinforcement Learning Datasets and Benchmarks for Offline Safe Reinforcement Learning

Reference 6

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T16:36:09.839172Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-09T15:44:36.262834Z digest=sha256:a3a0096059ddeb95530bbb6d4731285dfd458a2c7ccf87e7949e12a559bd7aba

Observation 93f73896-d807-493a-820c-1f05ce9435a7 · inbound

Behavior-Consistent Deep Reinforcement Learning cites this paper.

Behavior-Consistent Deep Reinforcement Learning Datasets and Benchmarks for Offline Safe Reinforcement Learning

Reference 128

Resolution
verified exact
arxiv_id, observed 2026-05-21T05:39:40.732442Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-21T05:37:29.919862Z digest=sha256:bbbc81353879cd8e5cd10c5ffdbabc5d8a37f13f0efff4cfac4fbb3d0ad41f9b

Observation d966ac6b-1eb3-4855-adb9-0b41eda9a314 · inbound

Behavior-Consistent Deep Reinforcement Learning cites this paper.

Behavior-Consistent Deep Reinforcement Learning Datasets and Benchmarks for Offline Safe Reinforcement Learning

Reference 128

Resolution
verified exact
arxiv_id, observed 2026-05-22T10:06:21.827692Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-22T10:06:04.478006Z digest=sha256:a97f1574fb3123e1365af207c65180e2957b581e1135ad112a53f79c0615acf3

Observation ff058c93-975a-4621-9f80-3e6dfff1420c · inbound

Safe-RULE: Safe Reinforcement UnLEarning cites this paper.

Safe-RULE: Safe Reinforcement UnLEarning Datasets and Benchmarks for Offline Safe Reinforcement Learning

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-07-03T00:07:28.405955Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-27T17:28:09.686163Z digest=sha256:fde6036b8ecca061ac523b4630a8cd795b4084407ac63408bc8e1973ddf4227d

Observation 141933e7-b4dd-4989-87f1-cd399241c878 · inbound

OopsieVerse: A Safety Benchmark with Damage-Aware Simulation for Robot Manipulation cites this paper.

OopsieVerse: A Safety Benchmark with Damage-Aware Simulation for Robot Manipulation Datasets and Benchmarks for Offline Safe Reinforcement Learning

Reference 25

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T10:45:43.023784Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-01T05:05:51.441647Z digest=sha256:fd5c4c403a024ec339f8d77eca57a7928798fe95e08cde093488c9bd6c31a30e

Observation 3c08cf8e-55d0-414e-b187-1e2135b5b822 · inbound

CDCP: Conditional Diffusion Model with Contextual Prompts for Multi-task Offline Safe Reinforcement Learning cites this paper.

CDCP: Conditional Diffusion Model with Contextual Prompts for Multi-task Offline Safe Reinforcement Learning Datasets and Benchmarks for Offline Safe Reinforcement Learning

Reference 68

Resolution
unresolved
no resolver link, observed 2026-07-11T23:08:40.265657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:08:40.265657Z digest=sha256:7bc81ff104bf4a9f57b96926da22304e2a9407b7e2be7ca69c24cf416fa1a551