Pith. sign in

Paper Citation Record · LEDGER

Deontically Constrained Policy Improvement in Reinforcement Learning Agents

As of 19 August 2026, this Paper Citation Record lists 30 of 30 outbound references and 0 inbound Pith citation observations for arXiv:2506.06959.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.06959 v1

Coverage vector

measured 30 of 30 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T05:51:59.377166Z

measured 30 of 30 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

30 of 30 outbound references displayed

  • verified exact3
  • verified fuzzy15
  • unresolved12
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation ef694f84-f6fd-4147-be08-60dc1791f07e · outbound

This paper cites an unresolved cited work.

Deontically Constrained Policy Improvement in Reinforcement Learning Agents Unresolved cited work

Reference 1

Resolution
verified exact
doi, observed 2026-08-07T05:51:59.431911Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T05:51:59.275128Z digest=sha256:365552f839d212f4a2167c9523fda5534b2407a69ecc02dd82aee5e2597c8a41

Observation 776b32c3-0d34-4a9f-b5e4-525e8eaf151f · outbound

This paper cites MacGlashan and M.

Deontically Constrained Policy Improvement in Reinforcement Learning Agents MacGlashan and M

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:51:59.733942Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T05:51:59.279650Z digest=sha256:20896cfccbbe69b20e7df3b64ebf7bdca29e380e92308f7c1c537d52c0c28d3b

Observation 93214bf8-2f8a-4625-bc76-32c06ece7a55 · outbound

This paper cites an unresolved cited work.

Deontically Constrained Policy Improvement in Reinforcement Learning Agents Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:51:59.722956Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T05:51:59.283513Z digest=sha256:e1925f31236b3eba52573a16a90065cc167b6bd9680f54dfbd34353278fc7155

Observation 55b61f62-4ce3-4ba2-b7b9-4ef132ec3a49 · outbound

This paper cites CONSTRAINED MARKOV DECISION PROCESSES,.

Deontically Constrained Policy Improvement in Reinforcement Learning Agents CONSTRAINED MARKOV DECISION PROCESSES,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:51:59.709523Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T05:51:59.289018Z digest=sha256:a999476bdcd35d3b6622f6981bd47502090d703636e9c21b223d82321168d5bd

Observation 7099a15a-d4a5-4ace-9ced-74e4bf3b3739 · outbound

This paper cites A Framework for Transforming Specifications in Reinforcement Learning,.

Deontically Constrained Policy Improvement in Reinforcement Learning Agents A Framework for Transforming Specifications in Reinforcement Learning,

Reference 5

Resolution
verified exact
doi, observed 2026-08-07T05:51:59.420336Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T05:51:59.293265Z digest=sha256:d78d9f938a4bfba5e8d10027d8795c1cd4fc7f25aa60408975e7101e79dd5083

Observation a427ae30-f82a-442b-8594-ecb891cd7695 · outbound

This paper cites Principles of model checking,.

Deontically Constrained Policy Improvement in Reinforcement Learning Agents Principles of model checking,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:51:59.698500Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T05:51:59.297257Z digest=sha256:71dd76148074a9898d4fa04a131d589bcd8f8848c1367f02b3a659c796d4f75e

Observation 9126efed-9ff3-4d66-ad0f-ee6c61a50fa2 · outbound

This paper cites 467–477, eleventh European Conference on Symbolic and Quantitative Approaches to Reasoning with Uncertainty (ECSQARU 2011).

Deontically Constrained Policy Improvement in Reinforcement Learning Agents 467–477, eleventh European Conference on Symbolic and Quantitative Approaches to Reasoning with Uncertainty (ECSQARU 2011)

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:51:59.687200Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T05:51:59.301000Z digest=sha256:d2b66142f5785aabb39534d8cf26b2cab07487143198d55794f5aa4ac8ff0652

Observation aafed6bc-9ddb-4b3e-b599-fe44f924c8ee · outbound

This paper cites Herzig and N.

Deontically Constrained Policy Improvement in Reinforcement Learning Agents Herzig and N

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:51:59.675892Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T05:51:59.304330Z digest=sha256:b9a0abc749c043fd41decce1e0f1200994d2252ab6cf61405e09696b49928a97

Observation 8d40a761-ac08-4c5e-a10f-fd7a54493fe6 · outbound

This paper cites Herzig and N.

Deontically Constrained Policy Improvement in Reinforcement Learning Agents Herzig and N

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:51:59.664547Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T05:51:59.307570Z digest=sha256:12d95dd4320efa8d4fd8737fdbd7e62b40511c99ffcef6c8faee806cc0d5e963

Observation 1c348826-f92a-42a7-aa39-477b4affa13c · outbound

This paper cites an unresolved cited work.

Deontically Constrained Policy Improvement in Reinforcement Learning Agents Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:51:59.653571Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T05:51:59.311715Z digest=sha256:8ff8cebafd7940cbf226349630d4a65c077c46c22f3e81e04e846f5f818b1caf

Observation 89dd09a2-75ce-4f24-8adc-d1986a460b2e · outbound

This paper cites Ellis, M.

Deontically Constrained Policy Improvement in Reinforcement Learning Agents Ellis, M

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:51:59.642195Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T05:51:59.315080Z digest=sha256:964e1746952de1bad9b5d3eb3dbe00d1c94af2714accbd0705f9be070fdb4370

Observation ee4c91fb-8835-401b-bea3-13e512b416c0 · outbound

This paper cites Agency and Deontic Logic,.

Deontically Constrained Policy Improvement in Reinforcement Learning Agents Agency and Deontic Logic,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:51:59.630502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T05:51:59.318258Z digest=sha256:dbf17d9b440dad62f0cbdb5bf01a0e50776eefd64b2ca9bb828ca87e387cdeb5

Observation 8a7150f9-679e-43fc-9847-0c95f5ab478a · outbound

This paper cites an unresolved cited work.

Deontically Constrained Policy Improvement in Reinforcement Learning Agents Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:51:59.619582Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T05:51:59.322021Z digest=sha256:695077d6314acc2eb835088ac912eb79e85d0f338471c636188c0814b2199c44

Observation 338e8366-ee62-439b-b80f-f3fbb1176f0b · outbound

This paper cites an unresolved cited work.

Deontically Constrained Policy Improvement in Reinforcement Learning Agents Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:51:59.608729Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T05:51:59.325125Z digest=sha256:bcbafca9dc8faf5fa73cb56a551ead4be4bd9f6c5798323796eaf3037ecab5b0

Observation f3d4ad82-b2db-4ad2-afff-10c4c30b927b · outbound

This paper cites Piterman and D.

Deontically Constrained Policy Improvement in Reinforcement Learning Agents Piterman and D

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:51:59.597907Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T05:51:59.328414Z digest=sha256:f2bbbba37af8441026d9378d148099627a040ae740ebf1b4830495043e67856a

Observation 5fb7a5b2-388f-424a-b79f-fcb4fbbfb74c · outbound

This paper cites an unresolved cited work.

Deontically Constrained Policy Improvement in Reinforcement Learning Agents Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:51:59.587483Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T05:51:59.331580Z digest=sha256:f423a0e361a2b8ab6e8e10dd6eb12226aa0620a19235b6ffb3681af88ac9d93c

Observation 3f2c4751-fcc9-4542-83ba-4051015fbe45 · outbound

This paper cites an unresolved cited work.

Deontically Constrained Policy Improvement in Reinforcement Learning Agents Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:51:59.576429Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T05:51:59.334812Z digest=sha256:2e54832000076e726620e89fb4d285d78ef3056973acca47eb3206a2ba751c91

Observation 96e9d5b1-829f-4b5f-a4fb-e83559a48170 · outbound

This paper cites an unresolved cited work.

Deontically Constrained Policy Improvement in Reinforcement Learning Agents Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:51:59.565583Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T05:51:59.337885Z digest=sha256:de48224adc803ea6bafe0c5eb3cb2ec788a0d647a0f19666e9a4412f247598f6

Observation d9248b66-05a8-4b9d-a4e2-eb2b03d77b6b · outbound

This paper cites Jansen and U.

Deontically Constrained Policy Improvement in Reinforcement Learning Agents Jansen and U

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:51:59.554208Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T05:51:59.341372Z digest=sha256:f0669be727d569e0a73161d38b1cbf421d92db5c88d34b83cb22166432a18264

Observation add62f7c-d7fa-4fcb-a6eb-4eae663f5e96 · outbound

This paper cites URLhttp://www.lsv.fr/ ~markey/Teaching/ESSLLI06/ESSLLI-proc.pdf.

Deontically Constrained Policy Improvement in Reinforcement Learning Agents URLhttp://www.lsv.fr/ ~markey/Teaching/ESSLLI06/ESSLLI-proc.pdf

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:51:59.542576Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T05:51:59.344471Z digest=sha256:3da2d408b132ab985ddaf16467ea8a44351ee09bfc9d5e79e3f5519721a277ce

Observation 01811f5a-0c47-475f-9c93-b4a971dde818 · outbound

This paper cites an unresolved cited work.

Deontically Constrained Policy Improvement in Reinforcement Learning Agents Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:51:59.531232Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T05:51:59.347827Z digest=sha256:43ddbe08dbe2aa8539121bd1f479ace58027ca6d253cd17dce17e6b10d6bef5a

Observation d9f40715-cc08-47b5-80ec-c53063625578 · outbound

This paper cites The Alignment Problem from a Deep Learning Perspective.

Deontically Constrained Policy Improvement in Reinforcement Learning Agents The Alignment Problem from a Deep Learning Perspective

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T05:51:59.350948Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:51:59.350948Z digest=sha256:9713856d2fe66e7cf4e379816317ebae2d54fad826556176a101224797cffc01

Observation 649473c8-5b57-41dc-89cc-f2b34fa42493 · outbound

This paper cites Bhatia and J.

Deontically Constrained Policy Improvement in Reinforcement Learning Agents Bhatia and J

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:51:59.520139Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T05:51:59.354394Z digest=sha256:1369c95bd237323b79418709f908d2bdb92705d64e40f449ebc957b6b54b66e1

Observation 4e20bc2b-d882-4156-9594-177191708a1d · outbound

This paper cites Reymond, P.

Deontically Constrained Policy Improvement in Reinforcement Learning Agents Reymond, P

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:51:59.508403Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T05:51:59.357500Z digest=sha256:ef1fa69c80e05844e4c3b06d70748e57dc7224878b832077f39ce1432f593e4e

Observation d0f4ae0b-7c28-4a7e-a97a-bfd77265eeb4 · outbound

This paper cites an unresolved cited work.

Deontically Constrained Policy Improvement in Reinforcement Learning Agents Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:51:59.496912Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T05:51:59.360650Z digest=sha256:061bf4b64b709cbf28de89cbc414e57514bb98f7dcbb0f7bf224f2b50c23a01c

Observation 53f10e6f-ce32-4548-8b30-6b1bf90b56bf · outbound

This paper cites an unresolved cited work.

Deontically Constrained Policy Improvement in Reinforcement Learning Agents Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:51:59.486997Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T05:51:59.363767Z digest=sha256:2c32fed77c9c69841f239ac614a33e047e7259728aaa49744ab555d037a4b522

Observation 4321b91f-5364-4073-8111-67e19769f0ed · outbound

This paper cites an unresolved cited work.

Deontically Constrained Policy Improvement in Reinforcement Learning Agents Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:51:59.477003Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T05:51:59.367067Z digest=sha256:ed6bff26b14e04d00722682dddf0b8103a3c8bf431c9f1cd3980e2921649dc21

Observation 274d09dc-98ae-4601-8153-70bb957e9981 · outbound

This paper cites Reinforcement Learning: An Introduction,.

Deontically Constrained Policy Improvement in Reinforcement Learning Agents Reinforcement Learning: An Introduction,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:51:59.466724Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T05:51:59.370124Z digest=sha256:b703069cc9cb035dbec5c1552cdb3bc5a2b9b3001390aa1153f7a5b02cd2b7b9

Observation dff0c386-38fe-4f47-93ee-050983d9b013 · outbound

This paper cites an unresolved cited work.

Deontically Constrained Policy Improvement in Reinforcement Learning Agents Unresolved cited work

Reference 29

Resolution
verified exact
doi, observed 2026-08-07T05:51:59.407771Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T05:51:59.373500Z digest=sha256:832d811e3f06c353f0a1f6e80a40f1f09ed62e3cbf5fe1687d4894a58bbc63a6

Observation c9491c9f-ceba-4de6-b0ce-ac65b47d9c78 · outbound

This paper cites Littman and M.

Deontically Constrained Policy Improvement in Reinforcement Learning Agents Littman and M

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:51:59.455989Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T05:51:59.377166Z digest=sha256:7518367ab3e67581f325999a27092dc655e2fd099aa7670f7f4c4df14d41cae4

Pith citing papers

No inbound Pith citation observations are available.