Pith. sign in

Paper Citation Record · LEDGER

HAVA: Hybrid Approach to Value-Alignment through Reward Weighing for Reinforcement Learning

As of 22 August 2026, this Paper Citation Record lists 34 of 34 outbound references and 0 inbound Pith citation observations for arXiv:2505.15011.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.15011 v1

Coverage vector

measured 34 of 34 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:29:46.948522Z

measured 34 of 34 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

34 of 34 outbound references displayed

  • verified exact7
  • verified fuzzy5
  • unresolved21
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation f2332345-46e5-4552-bc79-0810f28a4cfc · outbound

This paper cites an unresolved cited work.

HAVA: Hybrid Approach to Value-Alignment through Reward Weighing for Reinforcement Learning Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:29:50.068590Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T15:29:43.685469Z digest=sha256:6e8db4000edab62a4bbecc52b7199a202cabcd77a8282517efb48df3c063136d

Observation 92fa9ca6-4ed9-4ec6-96f3-23a674c10e81 · outbound

This paper cites an unresolved cited work.

HAVA: Hybrid Approach to Value-Alignment through Reward Weighing for Reinforcement Learning Unresolved cited work

Reference 2

Resolution
verified exact
doi, observed 2026-08-07T15:29:47.879679Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T15:29:43.728897Z digest=sha256:3f922d2bf4b03fb0d29ebaad0c20ee2b7da65c9ef34e1a149137990b2e3255b9

Observation e9adfc80-a04d-4a0a-b6eb-05e92ec58a40 · outbound

This paper cites an unresolved cited work.

HAVA: Hybrid Approach to Value-Alignment through Reward Weighing for Reinforcement Learning Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:29:49.965780Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T15:29:43.806848Z digest=sha256:d13b7e9da8801305ed1044c2c52f946a046b8fc6278cf5bc2a22a150b6609b84

Observation b63afe19-0577-4c95-b520-50dbdd65545d · outbound

This paper cites 2011.Machine Ethics.

HAVA: Hybrid Approach to Value-Alignment through Reward Weighing for Reinforcement Learning 2011.Machine Ethics

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:43.877217Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:29:43.877217Z digest=sha256:55c059ab57c3b1d8e2e9edc623ae987f90c4967c208a5f1d7d4d8aab0711c93b

Observation 27251b20-f90f-4d83-99ab-cf5909d59cde · outbound

This paper cites an unresolved cited work.

HAVA: Hybrid Approach to Value-Alignment through Reward Weighing for Reinforcement Learning Unresolved cited work

Reference 5

Resolution
verified exact
doi, observed 2026-08-07T15:29:47.757610Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T15:29:44.051349Z digest=sha256:d9640682823e7c924aecabba0f74850840e2b146341c02bc68986f4241f4c913

Observation 5d31ef8c-32c7-4f81-bebc-032c5f01fb53 · outbound

This paper cites an unresolved cited work.

HAVA: Hybrid Approach to Value-Alignment through Reward Weighing for Reinforcement Learning Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:29:49.885469Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T15:29:44.103805Z digest=sha256:1f18e197a93ffa9b2680aab57c0a7c81206d2c72b5b5868cae692e0b50c1faa8

Observation d1ed2666-4d6f-4b88-a99b-c05226295133 · outbound

This paper cites an unresolved cited work.

HAVA: Hybrid Approach to Value-Alignment through Reward Weighing for Reinforcement Learning Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:29:49.819983Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T15:29:44.273364Z digest=sha256:2fff989faec54c26ec72069fbc5161abbd9e572240768fcf1e484bd098fff7b1

Observation 7c782612-e9c2-4935-8399-85d73ef4f1c9 · outbound

This paper cites Noisy Networks for Exploration.

HAVA: Hybrid Approach to Value-Alignment through Reward Weighing for Reinforcement Learning Noisy Networks for Exploration

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:44.411449Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:29:44.411449Z digest=sha256:495f7849e765b616d3591b19c1fa5aa939aa73171b1863885bb3bc8f98041c2f

Observation 548e27e1-1903-42a6-8b79-686afb662f1d · outbound

This paper cites an unresolved cited work.

HAVA: Hybrid Approach to Value-Alignment through Reward Weighing for Reinforcement Learning Unresolved cited work

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:44.541245Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:29:44.541245Z digest=sha256:e2ff78b9003e5f33ae40279fcbfb3ace75cc9b4145f93fcc3c4e97fa2c71bf51

Observation 1b5c38e4-c5c3-4d5f-8325-3b2da96af338 · outbound

This paper cites an unresolved cited work.

HAVA: Hybrid Approach to Value-Alignment through Reward Weighing for Reinforcement Learning Unresolved cited work

Reference 11

Resolution
verified exact
doi, observed 2026-08-07T15:29:47.598569Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T15:29:44.813018Z digest=sha256:4937a947f6e51fe0769bb95898ddb7f85f6ec9eaf3d667bd867f4abf2c96fbb6

Observation 6bf8f9d4-7852-479b-aa0d-da8c51ca34b9 · outbound

This paper cites an unresolved cited work.

HAVA: Hybrid Approach to Value-Alignment through Reward Weighing for Reinforcement Learning Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:29:49.746484Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T15:29:44.931454Z digest=sha256:e3bb67bc3e6f3d216931072a7e37750bbffba4596c0c7888ff9f3ae4aecc1d3c

Observation eb319be1-0dd1-4b26-a26b-d5a732be27aa · outbound

This paper cites Hauptabt.

HAVA: Hybrid Approach to Value-Alignment through Reward Weighing for Reinforcement Learning Hauptabt

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:29:49.670291Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T15:29:45.058627Z digest=sha256:d0dad9be039a1007ba3bbfbd052e319c5ff4a45e85a661c5d259454e881adf02

Observation 7de1c634-0e51-4e0c-926b-1730fe63f9fd · outbound

This paper cites an unresolved cited work.

HAVA: Hybrid Approach to Value-Alignment through Reward Weighing for Reinforcement Learning Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:29:49.600286Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T15:29:45.167157Z digest=sha256:a0ecda3bc6c1c1a78f12db64e285f105bdd804ac58095d96949b549f31ecbf1a

Observation f8d70b68-4efb-4b33-90cb-1371c0060446 · outbound

This paper cites an unresolved cited work.

HAVA: Hybrid Approach to Value-Alignment through Reward Weighing for Reinforcement Learning Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:29:49.502077Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T15:29:45.263771Z digest=sha256:3c0e38930226e4f631d5729609eb1336a35c0fd23a457b21434f7add47d592c1

Observation 550a2488-0789-4df3-87ee-9969e6210caa · outbound

This paper cites Rusu, Joel Veness, Marc G.

HAVA: Hybrid Approach to Value-Alignment through Reward Weighing for Reinforcement Learning Rusu, Joel Veness, Marc G

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:45.357991Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:29:45.357991Z digest=sha256:253943069340b383921f0ae8a663d76cf2eb206272301b017a40e6194edf8e4c

Observation a3dbad0a-4ed1-45e7-97ba-f38945127bb1 · outbound

This paper cites Training Value-Aligned Reinforcement Learning Agents Using a Normative Prior.

HAVA: Hybrid Approach to Value-Alignment through Reward Weighing for Reinforcement Learning Training Value-Aligned Reinforcement Learning Agents Using a Normative Prior

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-08-07T15:29:48.390968Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T15:29:45.440912Z digest=sha256:fb4d77285d398b50cc2f479230e8267fb51ddf24651587bb3e66026577de25a4

Observation af378706-719e-4f27-b0e4-215f3a9c79b6 · outbound

This paper cites an unresolved cited work.

HAVA: Hybrid Approach to Value-Alignment through Reward Weighing for Reinforcement Learning Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:45.536818Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:29:45.536818Z digest=sha256:e7c026517050de804c8132d81e9d21db4655d60f334596cc51725383d8263262

Observation 31e0eabc-f2b7-47cc-a689-0f23b5163ca9 · outbound

This paper cites Neufeld, Ezio Bartocci, and Agata Ciabattoni.

HAVA: Hybrid Approach to Value-Alignment through Reward Weighing for Reinforcement Learning Neufeld, Ezio Bartocci, and Agata Ciabattoni

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:29:49.401537Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T15:29:45.659152Z digest=sha256:f30d88c37f5295d77e4ea7c239a09b5a9e865643cae62e754e719a4e252bc71d

Observation 28df5bc8-b13f-4eb6-a5f9-1b6c73d43956 · outbound

This paper cites Neufeld, Ezio Bartocci, Agata Ciabattoni, and Guido Governatori.

HAVA: Hybrid Approach to Value-Alignment through Reward Weighing for Reinforcement Learning Neufeld, Ezio Bartocci, Agata Ciabattoni, and Guido Governatori

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:29:49.215661Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T15:29:45.791671Z digest=sha256:0af14d656b527c3935a95eeb4d1c7ff5a696f1e153f1a813ec896bbbad1a3fcd

Observation 3f6dd90d-8610-495a-bfe2-0a004950ee3f · outbound

This paper cites Varshney, Murray Campbell, Moninder Singh, and Francesca Rossi.

HAVA: Hybrid Approach to Value-Alignment through Reward Weighing for Reinforcement Learning Varshney, Murray Campbell, Moninder Singh, and Francesca Rossi

Reference 21

Resolution
verified exact
doi, observed 2026-08-07T15:29:47.389279Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T15:29:46.016637Z digest=sha256:9f9d4a099b80672362a10598bc1c9fe2d0879daa87d1a03da48799409b3ca196

Observation ea220ad8-f90a-4872-a2df-eff8584e2783 · outbound

This paper cites Osoba, Benjamin Boudreaux, and Douglas Yeung.

HAVA: Hybrid Approach to Value-Alignment through Reward Weighing for Reinforcement Learning Osoba, Benjamin Boudreaux, and Douglas Yeung

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:46.133292Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:29:46.133292Z digest=sha256:97e05752fc88089a8ff513bffd05de346423d01938d1961a3d80d65106eb3b7d

Observation ef64e642-b235-4f0e-ba81-6653850dc5e2 · outbound

This paper cites an unresolved cited work.

HAVA: Hybrid Approach to Value-Alignment through Reward Weighing for Reinforcement Learning Unresolved cited work

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:46.189575Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:29:46.189575Z digest=sha256:30f14a4c7cc1b1e9d8b5a2d161d5fe06965f4a8bc55066172b94bb260637b977

Observation 4854d889-93a8-40a6-bedc-e52c39514889 · outbound

This paper cites Oliehoek, and Luciano C.

HAVA: Hybrid Approach to Value-Alignment through Reward Weighing for Reinforcement Learning Oliehoek, and Luciano C

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:29:49.069779Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T15:29:46.262577Z digest=sha256:0fcab8afa08096324b2b511d4d9f48da0305443b25b0e04ecfb42336b1cea318

Observation 8e0d87da-b920-4c2c-be6b-fe776b1211e7 · outbound

This paper cites an unresolved cited work.

HAVA: Hybrid Approach to Value-Alignment through Reward Weighing for Reinforcement Learning Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:29:48.798664Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T15:29:46.373662Z digest=sha256:fbfb0ff7cee9febcd3fbc18f8d04ed2712e7783a24d06cc2651e3f1aab14e6f7

Observation f2e25747-3b2a-4679-bd64-903669f2310d · outbound

This paper cites Shaw, Andreas Stöckel, Ryan W.

HAVA: Hybrid Approach to Value-Alignment through Reward Weighing for Reinforcement Learning Shaw, Andreas Stöckel, Ryan W

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:46.489529Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:29:46.489529Z digest=sha256:e595994afc417596db7e761bed6172cf45a6bef3463003023ad61b9f4323d3cc

Observation dee9b111-49f4-4194-8121-6ad689b612e3 · outbound

This paper cites InProceedings of the 21st International Conf.

HAVA: Hybrid Approach to Value-Alignment through Reward Weighing for Reinforcement Learning InProceedings of the 21st International Conf

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:29:48.925273Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T15:29:46.320335Z digest=sha256:a8ab712b1e7c615bd43528b8e96591d27b7d1839ad15005e166f73e60195100a

Observation 164746cd-c772-4674-847e-d9d152fddccd · outbound

This paper cites Deep Reinforcement Learning with Double Q-learning.

HAVA: Hybrid Approach to Value-Alignment through Reward Weighing for Reinforcement Learning Deep Reinforcement Learning with Double Q-learning

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:46.611548Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:29:46.611548Z digest=sha256:b3ce7ba4d81a2294c51a808f8f260c6479612aa599838f6b64c8a3d6b8af7520

Observation c5cc2ca7-15bb-4bdf-ae16-5cb09a9a0ce0 · outbound

This paper cites Czarnecki, Michaël Mathieu, An- drew Dudzik, Junyoung Chung, David H.

HAVA: Hybrid Approach to Value-Alignment through Reward Weighing for Reinforcement Learning Czarnecki, Michaël Mathieu, An- drew Dudzik, Junyoung Chung, David H

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:46.745452Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:29:46.745452Z digest=sha256:c7964ab766adb96ea7fc2a61f186c35861a866e55cb79a03490fe7c523034e11

Observation b4cb5ffb-dfe9-4aa1-a9a3-1a7754de3f93 · outbound

This paper cites an unresolved cited work.

HAVA: Hybrid Approach to Value-Alignment through Reward Weighing for Reinforcement Learning Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:29:48.726223Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T15:29:46.846243Z digest=sha256:bd61ca3e9ab312b74e8a164ddc0412d83898adc626bd95a87625649d20894b03

Observation 1d695a29-cbb7-42db-84d1-e8ec1032ec9c · outbound

This paper cites an unresolved cited work.

HAVA: Hybrid Approach to Value-Alignment through Reward Weighing for Reinforcement Learning Unresolved cited work

Reference 31

Resolution
verified exact
raw_fallback, observed 2026-08-07T15:29:48.085845Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T15:29:46.535647Z digest=sha256:0232d998a7da9aede20b85cc8d6abcf65159e2e4c4ccdd0ec67ec177407bb397

Observation 948d17fb-3f1e-4d04-b5fe-c84cdc96c77d · outbound

This paper cites an unresolved cited work.

HAVA: Hybrid Approach to Value-Alignment through Reward Weighing for Reinforcement Learning Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:29:48.634613Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T15:29:46.948522Z digest=sha256:f150a78466cfbfe7b396aeae50bad07ab79ecb3d863480b8463ab41cda52e080

Observation fc70aa8e-b68e-466b-8639-77ca9f51adef · outbound

This paper cites on Artificial Intelligence33, 01 (Jul.

HAVA: Hybrid Approach to Value-Alignment through Reward Weighing for Reinforcement Learning on Artificial Intelligence33, 01 (Jul

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:44.163712Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:29:44.163712Z digest=sha256:9f72cd51869bdf50337add83483c96cd3d6bfc60aaf20b3d380c7480b9fa6ad8

Observation 505f516f-aac2-40d8-99ee-83f6560d8370 · outbound

This paper cites https://doi.org/10.1007/s10676- 022-09665-8.

HAVA: Hybrid Approach to Value-Alignment through Reward Weighing for Reinforcement Learning https://doi.org/10.1007/s10676- 022-09665-8

Reference 2022

Resolution
malformed identifier
no resolver link, observed 2026-08-07T15:29:45.894600Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:29:45.894600Z digest=sha256:67a50676b8038b372b36454fc5db2e2594fc5ab2f58c41f29955b4aca4190b34

Observation 4fb8835b-2bb1-4ef5-8d8d-ea135d9da299 · outbound

This paper cites an unresolved cited work.

HAVA: Hybrid Approach to Value-Alignment through Reward Weighing for Reinforcement Learning Unresolved cited work

Reference 9789

Resolution
verified exact
doi, observed 2026-08-07T15:29:47.165029Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T15:29:46.434752Z digest=sha256:cfdcf9dd8a75ca7845453e097647a4b7f7064c150fe053d5b6d6f05057978fca

Pith citing papers

No inbound Pith citation observations are available.