Pith. sign in

Paper Citation Record · LEDGER

Deontically Constrained Policy Improvement in Reinforcement Learning Agents

As of 14 August 2026, this Paper Citation Record lists 30 of 30 outbound references and 0 inbound Pith citation observations for arXiv:2506.06959.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.06959 v1

Coverage vector

measured 30 of 30 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T05:51:59.377166Z

measured 30 of 30 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

30 of 30 outbound references displayed

  • verified exact3
  • verified fuzzy15
  • unresolved12
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation ef694f84-f6fd-4147-be08-60dc1791f07e · outbound

This paper cites an unresolved cited work.

Deontically Constrained Policy Improvement in Reinforcement Learning Agents Unresolved cited work

Reference 1

Resolution
verified exact
doi, observed 2026-08-07T05:51:59.431911Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T05:51:59.275128Z digest=sha256:2043ebdf897effd222b7a092f4eaaefeb9e27b59c7190ecd22f69c162d7495b1

Observation 776b32c3-0d34-4a9f-b5e4-525e8eaf151f · outbound

This paper cites MacGlashan and M.

Deontically Constrained Policy Improvement in Reinforcement Learning Agents MacGlashan and M

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:51:59.733942Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T05:51:59.279650Z digest=sha256:5e3e208c8479291e6380cfb726c73a6bdf4b9c2b1725b9b9e8546803eee045a5

Observation 93214bf8-2f8a-4625-bc76-32c06ece7a55 · outbound

This paper cites an unresolved cited work.

Deontically Constrained Policy Improvement in Reinforcement Learning Agents Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:51:59.722956Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T05:51:59.283513Z digest=sha256:95d12933d74301ff2ece4568f72e562d15b285bc252bc7b0ed940c1a9674d07d

Observation 55b61f62-4ce3-4ba2-b7b9-4ef132ec3a49 · outbound

This paper cites CONSTRAINED MARKOV DECISION PROCESSES,.

Deontically Constrained Policy Improvement in Reinforcement Learning Agents CONSTRAINED MARKOV DECISION PROCESSES,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:51:59.709523Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T05:51:59.289018Z digest=sha256:6deb95ffc4764595f3e6dd1c1dd3c90046f3c1dc250e4a370b131f90cf1c4306

Observation 7099a15a-d4a5-4ace-9ced-74e4bf3b3739 · outbound

This paper cites A Framework for Transforming Specifications in Reinforcement Learning,.

Deontically Constrained Policy Improvement in Reinforcement Learning Agents A Framework for Transforming Specifications in Reinforcement Learning,

Reference 5

Resolution
verified exact
doi, observed 2026-08-07T05:51:59.420336Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T05:51:59.293265Z digest=sha256:00c65d1538c9f0c57298cee500ee507b39a1700597f37c2fd1089ff87e3ef0b2

Observation a427ae30-f82a-442b-8594-ecb891cd7695 · outbound

This paper cites Principles of model checking,.

Deontically Constrained Policy Improvement in Reinforcement Learning Agents Principles of model checking,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:51:59.698500Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T05:51:59.297257Z digest=sha256:ae02515eddb03e3182397e37c8c1f05f4c06fd0d814df3b3ad7ae0ec9b76b1f2

Observation 9126efed-9ff3-4d66-ad0f-ee6c61a50fa2 · outbound

This paper cites 467–477, eleventh European Conference on Symbolic and Quantitative Approaches to Reasoning with Uncertainty (ECSQARU 2011).

Deontically Constrained Policy Improvement in Reinforcement Learning Agents 467–477, eleventh European Conference on Symbolic and Quantitative Approaches to Reasoning with Uncertainty (ECSQARU 2011)

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:51:59.687200Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T05:51:59.301000Z digest=sha256:83da1cdec146864f673cf31bbf76cb227c4eae33a1fd048479dc179c27a74068

Observation aafed6bc-9ddb-4b3e-b599-fe44f924c8ee · outbound

This paper cites Herzig and N.

Deontically Constrained Policy Improvement in Reinforcement Learning Agents Herzig and N

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:51:59.675892Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T05:51:59.304330Z digest=sha256:c6a8e6d635b86a774e71574c7cb958b4818bdb4be72337c260c04b55797760b9

Observation 8d40a761-ac08-4c5e-a10f-fd7a54493fe6 · outbound

This paper cites Herzig and N.

Deontically Constrained Policy Improvement in Reinforcement Learning Agents Herzig and N

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:51:59.664547Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T05:51:59.307570Z digest=sha256:dd57b2c78349bd1d2bcc68d0845924a17ed83215e7c4dde689345bf68447bfb5

Observation 1c348826-f92a-42a7-aa39-477b4affa13c · outbound

This paper cites an unresolved cited work.

Deontically Constrained Policy Improvement in Reinforcement Learning Agents Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:51:59.653571Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T05:51:59.311715Z digest=sha256:e6df9a2cb6e98c80858affd2f6040ba55a60cbf39dc45f1827ce06c900952ff3

Observation 89dd09a2-75ce-4f24-8adc-d1986a460b2e · outbound

This paper cites Ellis, M.

Deontically Constrained Policy Improvement in Reinforcement Learning Agents Ellis, M

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:51:59.642195Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T05:51:59.315080Z digest=sha256:cba1e6379d015fb4308df8490e79b35f7f6bda3dc54a52b5999d9fbf30af8b1a

Observation ee4c91fb-8835-401b-bea3-13e512b416c0 · outbound

This paper cites Agency and Deontic Logic,.

Deontically Constrained Policy Improvement in Reinforcement Learning Agents Agency and Deontic Logic,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:51:59.630502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T05:51:59.318258Z digest=sha256:56fdde07317d195e67437125b061c1fc3235e472e5a3aa3655a6c79fa8f9bb4f

Observation 8a7150f9-679e-43fc-9847-0c95f5ab478a · outbound

This paper cites an unresolved cited work.

Deontically Constrained Policy Improvement in Reinforcement Learning Agents Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:51:59.619582Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T05:51:59.322021Z digest=sha256:621c76d1b00f32a3a9b70e36c2f10a9f79ba4e4e7b7c2d6c55808f082b157fc7

Observation 338e8366-ee62-439b-b80f-f3fbb1176f0b · outbound

This paper cites an unresolved cited work.

Deontically Constrained Policy Improvement in Reinforcement Learning Agents Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:51:59.608729Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T05:51:59.325125Z digest=sha256:40b5f83f873f1e203ed2c2989b3911f8acd147933b9cfa679f2f4ee0082fc532

Observation f3d4ad82-b2db-4ad2-afff-10c4c30b927b · outbound

This paper cites Piterman and D.

Deontically Constrained Policy Improvement in Reinforcement Learning Agents Piterman and D

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:51:59.597907Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T05:51:59.328414Z digest=sha256:83c7e6be63bf71b56271989d3a00db096b24e5ec556271acc0bd17bafeb150f0

Observation 5fb7a5b2-388f-424a-b79f-fcb4fbbfb74c · outbound

This paper cites an unresolved cited work.

Deontically Constrained Policy Improvement in Reinforcement Learning Agents Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:51:59.587483Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T05:51:59.331580Z digest=sha256:03228978ca9d9aeb0e1075d0f69c8c77db4058c5ede64c73e09955520a413cd0

Observation 3f2c4751-fcc9-4542-83ba-4051015fbe45 · outbound

This paper cites an unresolved cited work.

Deontically Constrained Policy Improvement in Reinforcement Learning Agents Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:51:59.576429Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T05:51:59.334812Z digest=sha256:e5b520b37bb0336e4b1722ea798b7943a0c2ccacffbac1712b55b9136ce210b1

Observation 96e9d5b1-829f-4b5f-a4fb-e83559a48170 · outbound

This paper cites an unresolved cited work.

Deontically Constrained Policy Improvement in Reinforcement Learning Agents Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:51:59.565583Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T05:51:59.337885Z digest=sha256:2b57d9dd899e78a24f30aa63df4feae0200af0577994a5d5d84f60c662f613e0

Observation d9248b66-05a8-4b9d-a4e2-eb2b03d77b6b · outbound

This paper cites Jansen and U.

Deontically Constrained Policy Improvement in Reinforcement Learning Agents Jansen and U

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:51:59.554208Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T05:51:59.341372Z digest=sha256:9b88518fa05df60f51829865efd5ac3e386d5152c1ae47b5bcb90aa09d690cff

Observation add62f7c-d7fa-4fcb-a6eb-4eae663f5e96 · outbound

This paper cites URLhttp://www.lsv.fr/ ~markey/Teaching/ESSLLI06/ESSLLI-proc.pdf.

Deontically Constrained Policy Improvement in Reinforcement Learning Agents URLhttp://www.lsv.fr/ ~markey/Teaching/ESSLLI06/ESSLLI-proc.pdf

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:51:59.542576Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T05:51:59.344471Z digest=sha256:915a53c305f3efb0765e439b7d077416943ea243f72a1da4313954f34f46489b

Observation 01811f5a-0c47-475f-9c93-b4a971dde818 · outbound

This paper cites an unresolved cited work.

Deontically Constrained Policy Improvement in Reinforcement Learning Agents Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:51:59.531232Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T05:51:59.347827Z digest=sha256:46197e271e0c4dd787e0b3d0e38bf4d4da035915bce43be56bdefb6d1938d244

Observation d9f40715-cc08-47b5-80ec-c53063625578 · outbound

This paper cites The Alignment Problem from a Deep Learning Perspective.

Deontically Constrained Policy Improvement in Reinforcement Learning Agents The Alignment Problem from a Deep Learning Perspective

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T05:51:59.350948Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:51:59.350948Z digest=sha256:2defda69f870a33237045fb5d9a1113425d44ad1ee39e9c418e84c32976ebf71

Observation 649473c8-5b57-41dc-89cc-f2b34fa42493 · outbound

This paper cites Bhatia and J.

Deontically Constrained Policy Improvement in Reinforcement Learning Agents Bhatia and J

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:51:59.520139Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T05:51:59.354394Z digest=sha256:4e1171a26b8cd508232f26ff73f11e60b0d5f9427853fd67b616b9b5604f2abb

Observation 4e20bc2b-d882-4156-9594-177191708a1d · outbound

This paper cites Reymond, P.

Deontically Constrained Policy Improvement in Reinforcement Learning Agents Reymond, P

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:51:59.508403Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T05:51:59.357500Z digest=sha256:6919d226c65e72e0bb2ec7b4d943daf4e52ddf4ba24c8612a2500a24e68072d1

Observation d0f4ae0b-7c28-4a7e-a97a-bfd77265eeb4 · outbound

This paper cites an unresolved cited work.

Deontically Constrained Policy Improvement in Reinforcement Learning Agents Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:51:59.496912Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T05:51:59.360650Z digest=sha256:a30087e8ac41919609b7a479fefc8c1b910b6a179dadb111d2637e799ed6b264

Observation 53f10e6f-ce32-4548-8b30-6b1bf90b56bf · outbound

This paper cites an unresolved cited work.

Deontically Constrained Policy Improvement in Reinforcement Learning Agents Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:51:59.486997Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T05:51:59.363767Z digest=sha256:2fe008293f083bff6563ba866bcd6204e4304e7a43bd75323933c37c119ef834

Observation 4321b91f-5364-4073-8111-67e19769f0ed · outbound

This paper cites an unresolved cited work.

Deontically Constrained Policy Improvement in Reinforcement Learning Agents Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:51:59.477003Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T05:51:59.367067Z digest=sha256:90901c5a31db49de48d895946e583c0403e56fd6c549521a913056b8857284ac

Observation 274d09dc-98ae-4601-8153-70bb957e9981 · outbound

This paper cites Reinforcement Learning: An Introduction,.

Deontically Constrained Policy Improvement in Reinforcement Learning Agents Reinforcement Learning: An Introduction,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:51:59.466724Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T05:51:59.370124Z digest=sha256:97869afb7e155eb35fea2aa381e56ac23983fdd0742c991f727756c74a875acc

Observation dff0c386-38fe-4f47-93ee-050983d9b013 · outbound

This paper cites an unresolved cited work.

Deontically Constrained Policy Improvement in Reinforcement Learning Agents Unresolved cited work

Reference 29

Resolution
verified exact
doi, observed 2026-08-07T05:51:59.407771Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T05:51:59.373500Z digest=sha256:42640003bd83720444857e6aadb16a4865e432c0193658c904ca6dfe96130638

Observation c9491c9f-ceba-4de6-b0ce-ac65b47d9c78 · outbound

This paper cites Littman and M.

Deontically Constrained Policy Improvement in Reinforcement Learning Agents Littman and M

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:51:59.455989Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T05:51:59.377166Z digest=sha256:ee1ec395e44349347699684fa0289723134abfcface6461b692f4e4b3dac76ad

Pith citing papers

No inbound Pith citation observations are available.