Pith. sign in

Paper Citation Record · LEDGER

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models

As of 10 August 2026, this Paper Citation Record lists 75 of 75 outbound references and 0 inbound Pith citation observations for arXiv:2607.13172.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.13172 v1

Coverage vector

measured 75 of 75 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-02T06:06:17.543845Z

measured 75 of 75 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

75 of 75 outbound references displayed

  • verified exact2
  • verified fuzzy0
  • unresolved73
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 5872679e-69fb-4871-8442-faea6ecf55b5 · outbound

This paper cites In: International Conference on Machine Learning.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: International Conference on Machine Learning

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:10.407075Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:10.407075Z digest=sha256:9af9f546ccd37ecdbe34c98ff2e8ab7247097cfa5796b29d079a597fc2eab631

Observation 645e471f-4904-4ee9-99fb-f067f3f6c5c9 · outbound

This paper cites CRC Press (1999).

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models CRC Press (1999)

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:10.482844Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:10.482844Z digest=sha256:b6e532bc6f36808ed9188bbe3a717a63df20e359027f091d38fd74eacd5a56c9

Observation 635f6ba8-d718-4b37-944c-2bde0560d2d1 · outbound

This paper cites Constrained Policy Optimization via Bayesian World Models.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Constrained Policy Optimization via Bayesian World Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:10.559482Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:10.559482Z digest=sha256:538fa29eaadf3fe208b9066f01d01a6c3c5580d0b51b66a5a8ecf7bd54cc52c2

Observation 57b6cea2-4dba-4902-a553-fedd4f52f7b6 · outbound

This paper cites V-JEPA 2: Self-Supervised Video Models Enable Understanding, Prediction and Planning.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models V-JEPA 2: Self-Supervised Video Models Enable Understanding, Prediction and Planning

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:10.637436Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:10.637436Z digest=sha256:fe5583230867c73e843629498680ec8471d83c514659ab8f3588a26a22120409

Observation 358affb7-4d68-4030-ad32-1c145fad9b7a · outbound

This paper cites an unresolved cited work.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Unresolved cited work

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:10.779702Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:10.779702Z digest=sha256:424edaf05eca3597f7f11aafd6cd87dc26aa504ab145fe3e8efb7589985ce2d8

Observation 21ae2d28-a47f-4457-b042-6a54ef43b148 · outbound

This paper cites In: Conference on Robot Learning.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: Conference on Robot Learning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:10.861850Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:10.861850Z digest=sha256:c65f3bacc2d8ae64e8ab8d7137807d004c3b1bbc47bfd0a4ca0441ad1fe5b393

Observation 3482d3ab-8f75-4c3b-8830-71a2cb140a5f · outbound

This paper cites In: Handbook of statistics, vol.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: Handbook of statistics, vol

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:10.943797Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:10.943797Z digest=sha256:033305c7905c0930d5af7ce6972bc3235841f8424b2d574cd7a7c432a67c5fd8

Observation aca5c713-9cad-4fe8-8fce-72390c668a99 · outbound

This paper cites The method of paired comparisons.Biometrika39(3/4), 324–345 (1952).

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models The method of paired comparisons.Biometrika39(3/4), 324–345 (1952)

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:11.032054Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:11.032054Z digest=sha256:bd9d92d240674285454dc8cc8aada87d6d3e99ac7687c9be9cf7dd28b4e8f109

Observation 5016fb8d-b9da-4ee7-b6cd-c63212b73bc5 · outbound

This paper cites OpenAI Gym.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models OpenAI Gym

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:11.152578Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:11.152578Z digest=sha256:c964e95e7d7c6b809b6aa0d2a8bfec9414ded8c9a8a421f1b6d3c6003c69e185

Observation 8f90a88f-d5d7-4b16-bcdf-454ff0b300cf · outbound

This paper cites In: International Conference on Machine Learning.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: International Conference on Machine Learning

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:11.262813Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:11.262813Z digest=sha256:fd326e7a73ca5965839e110ebf1581ca669e1d5751b76cae8317521c8644f2c8

Observation 02e51200-a361-4bce-bf94-993176c2b2c6 · outbound

This paper cites In: Conference on Robot Learning.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: Conference on Robot Learning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:11.371963Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:11.371963Z digest=sha256:1fcdf53a3d4e9ea0e6eeaaea3cb381c6188975a8ca3b85e6db2721c66cb3261f

Observation 13e8c098-c7b3-4f83-a797-a91f43c26da0 · outbound

This paper cites In: Proceedings of the AAAI Conference on Artificial Intelligence.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: Proceedings of the AAAI Conference on Artificial Intelligence

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:11.572284Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:11.572284Z digest=sha256:aafc4671b8ba840e1e87edf7ac77ae227dbcb4cd6b241f8ffc3c7d6ac80a3376

Observation a4ad9ae3-6f99-4754-b6a8-deff92f61d24 · outbound

This paper cites In: Advances in Neural Information Processing Systems.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: Advances in Neural Information Processing Systems

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:11.680141Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:11.680141Z digest=sha256:0ef9ec9982b3c27e44f0c3fd8dd1a410daa9c5b5ff149626686ec78be4817f35

Observation d7e97992-2418-4da3-b72c-fa52d17f9619 · outbound

This paper cites In: Advances in Neural Information Processing Systems.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: Advances in Neural Information Processing Systems

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:11.803531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:11.803531Z digest=sha256:eaf9436e4310b124097d94aedd7cbf54fa2dd2133f0e51e81971845518b5edd9

Observation 3e99214a-5d7b-4b27-9085-d06858a07401 · outbound

This paper cites In: The Twelfth International Conference on Learning Representations (2024).

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: The Twelfth International Conference on Learning Representations (2024)

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:11.914470Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:11.914470Z digest=sha256:06d986e18e724ab9584f85b15d55ba264c8ed3585bb41c925e30bd830272ee4a

Observation 003ee4ba-50a2-4782-81a2-488d3e064660 · outbound

This paper cites In: Pacific Rim Inter- national Conference on Artificial Intelligence.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: Pacific Rim Inter- national Conference on Artificial Intelligence

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:11.985701Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:11.985701Z digest=sha256:6c2efe63b1c886e40953007d664dc6105c2bc5c775f2dcd35029447fcdd98da7

Observation a4c24d43-d88f-4a7c-9c83-2358e42272e6 · outbound

This paper cites an unresolved cited work.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Unresolved cited work

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:12.044195Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:12.044195Z digest=sha256:89c50dacaa92d341dcbec011b7a90d2e9e4fe00c2bd350ba20d8e18cde759aeb

Observation d04bc555-6b81-4871-af4f-e414589ed257 · outbound

This paper cites Parenting: Safe Reinforcement Learning from Human Input.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Parenting: Safe Reinforcement Learning from Human Input

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:12.149527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:12.149527Z digest=sha256:c6999f27258a65e23a9c6b4164e829cd1e5abe61080e8de6fa879d2d8e56ef03

Observation 18b43076-150c-4c7d-b569-2580dd35b431 · outbound

This paper cites Automatica25(3), 335–348 (1989).

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Automatica25(3), 335–348 (1989)

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:12.208187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:12.208187Z digest=sha256:f6cd3f259a48f723d1d6223a524ffc8b9c65539804bc045ca24852afffcb1226

Observation 7509fb55-fde4-4126-b2ea-a2f6b0179fa0 · outbound

This paper cites Journal of Machine Learning Research16(1), 1437–1480 (2015).

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Journal of Machine Learning Research16(1), 1437–1480 (2015)

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:12.261956Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:12.261956Z digest=sha256:550022bea1c7aa4576779e3388e99685bc31a347bf26b5258f27fd9f28269d15

Observation 7fe89d64-3774-41e7-b02b-f175fe6d44c0 · outbound

This paper cites IEEE Access7, 165007–165017 (2019).

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models IEEE Access7, 165007–165017 (2019)

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:12.301697Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:12.301697Z digest=sha256:f13801a0a35d439ad91d82c16d8ea3565d79a507c7cae9947afc4316bf9a90d3

Observation bc6f1a8b-7806-4288-8df3-8424ba291284 · outbound

This paper cites In: Proceedings of the AAAI Conference on Artificial Intelligence.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: Proceedings of the AAAI Conference on Artificial Intelligence

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:12.410489Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:12.410489Z digest=sha256:78f150a843c124fad0d3dfb8dac629f86bda530e7e9563410c957efc847ad376

Observation 53896ce6-c72f-468f-a090-9c6c9ccbec8b · outbound

This paper cites Frontiers in Neurorobotics17, 1280341 (2023).

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Frontiers in Neurorobotics17, 1280341 (2023)

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:12.461979Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:12.461979Z digest=sha256:a7172232d80d91d4b301874d1d872b31248243be329ed104bb5dcfcc41cd697f

Observation 4fda8422-b934-4416-b670-1ccb4585d4b8 · outbound

This paper cites IEEE Transactions on Pattern Analysis and Machine Intelligence (2024).

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models IEEE Transactions on Pattern Analysis and Machine Intelligence (2024)

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:12.513304Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:12.513304Z digest=sha256:4574859ea4f45d94aac1bac643f9eefeb3cedb95810cbc21572e245289b66f9d

Observation 2d319fb4-ab68-45ab-bf88-a348bdfff61d · outbound

This paper cites In: International Conference on Principles and Practice of Multi-Agent Systems.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: International Conference on Principles and Practice of Multi-Agent Systems

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:12.563032Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:12.563032Z digest=sha256:0fbea3bd466664c1ea8b3ba6b342609d2a3d35b851a07225ab63ed471d2b0fcb

Observation c6b6c9fc-913c-436f-b37e-247da2248c9e · outbound

This paper cites In: Proceedings of the 32nd International Conference on Neural Information Process- ing Systems.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: Proceedings of the 32nd International Conference on Neural Information Process- ing Systems

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:12.615304Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:12.615304Z digest=sha256:38d72e7d2baa6c2962d6fde5902feec15f180cdf1d0024330be0c32a4d09c0ce

Observation 6e415062-a077-4100-88da-d721a118e24d · outbound

This paper cites Dream to Control: Learning Behaviors by Latent Imagination.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Dream to Control: Learning Behaviors by Latent Imagination

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:12.660894Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:12.660894Z digest=sha256:14d70d49ac567f25b61b2414bb80685ff0f77e4ff7a101f8e8716158ced26a5b

Observation 56aad13e-629d-4680-994f-87849a73a0f7 · outbound

This paper cites In: International Conference on Machine Learning.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: International Conference on Machine Learning

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:12.724121Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:12.724121Z digest=sha256:cf13382b4e28ad2c377f216bdbdca1dbf2bf79a576bbe8967dd0a3baa2a9365b

Observation 943c0f4b-72e2-4ba6-8660-fa074d8bb6d0 · outbound

This paper cites Mastering Atari with Discrete World Models.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Mastering Atari with Discrete World Models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:12.788064Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:12.788064Z digest=sha256:f8d5b6e004d8a5bef0429e6b6b2774a0254cab18a840bd9f0c610e2ac33f1989

Observation 0ca26ccb-06a1-42ec-a9ea-d7fafb19234a · outbound

This paper cites Nature640(8059), 647–653 (2025).

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Nature640(8059), 647–653 (2025)

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:12.900210Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:12.900210Z digest=sha256:976854cbaa886846d171a31ff31fb9e59da632568abe5a5673c0434bb045ffd8

Observation fa3e44c1-3dd1-4b81-bb01-b02016f7f152 · outbound

This paper cites Evolutionary Computation9(2), 159–195 (2001).

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Evolutionary Computation9(2), 159–195 (2001)

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:12.958241Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:12.958241Z digest=sha256:b3b921d4aee623547d1b752728de65ab6792e59933a924faf5eb155a83cba142

Observation f757f143-a240-4854-bd5f-c0c208396c7e · outbound

This paper cites Neural Computation 9(8), 1735–1780 (1997).

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Neural Computation 9(8), 1735–1780 (1997)

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:13.048117Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:13.048117Z digest=sha256:21be2b4ae5ef14680ef79d795332a2f40277f7f9fdf46675f015b309c7268e38

Observation 950401ec-3c38-4b08-98ad-faaad4904bb9 · outbound

This paper cites 65–70 (1979).

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models 65–70 (1979)

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:13.146254Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:13.146254Z digest=sha256:e7c660c493f3be9069706838459bd52473aef67428d813e0e31aefad2d6203be

Observation f56b354d-fb38-49c7-8a45-b4d22fd99058 · outbound

This paper cites In: International Conference on Learning Represen- tations.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: International Conference on Learning Represen- tations

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:13.201380Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:13.201380Z digest=sha256:c9160cebbe78bf8d5fdc516bea527fb5464286e35443e40d65afda8e7f12f9df

Observation f2ea76fe-7146-47a8-80f5-a3c04e4d718c · outbound

This paper cites IEEE Access (2025).

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models IEEE Access (2025)

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:13.274243Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:13.274243Z digest=sha256:03cbcb0f0056ff3a85c31ca4e75100ce866b9f22db1d994b51ec906116665aaa

Observation 1a0bf936-427a-4af9-ae7a-9d934677ae93 · outbound

This paper cites an unresolved cited work.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Unresolved cited work

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:13.294083Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:13.294083Z digest=sha256:6ba4a123d6c90f0e10820cf9e5ad7bc706506ffe2fb3c1a862be45dbfbc0fd2d

Observation cf643320-0fe1-44d1-ae41-7106519c751a · outbound

This paper cites In: Proceedings of the 21st International Conference on Autonomous Agents and Multiagent Systems.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: Proceedings of the 21st International Conference on Autonomous Agents and Multiagent Systems

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:13.415296Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:13.415296Z digest=sha256:f1909d0ceecb992f09faadfa154775f61a33e9f3f09d895b4f7ad180ca58269f

Observation 4c6c5dc8-7912-4fdc-98bb-884db763a9c0 · outbound

This paper cites an unresolved cited work.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Unresolved cited work

Reference 38

Resolution
verified exact
doi, observed 2026-08-02T06:09:24.475711Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-02T06:06:13.517450Z digest=sha256:13d409470adf772a9c180b3309f465c3c16618bab0f7fc66c5dd3b1f0e0c4b54

Observation 7c188e54-b49a-433a-bdf0-6619003d1220 · outbound

This paper cites In: Proceedings of the 18th International Conference on Agents and Artificial Intelligence.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: Proceedings of the 18th International Conference on Agents and Artificial Intelligence

Reference 39

Resolution
verified exact
doi, observed 2026-08-02T06:09:24.301079Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-02T06:06:13.622953Z digest=sha256:4fa11ab0175e99bfcca6ebff56bbbd2b9ec60dfca1258d9c4e800eb4f020fe3c

Observation f49963fc-ef84-4b43-a42b-96086028000d · outbound

This paper cites In: International Confer- ence on Learning Representations (2023).

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: International Confer- ence on Learning Representations (2023)

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:13.763095Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:13.763095Z digest=sha256:6fdbf0f87cfb8a37d29f335e0a6de9af24ae43d6ca845627205026c9c0fa04a2

Observation d09fb77e-9940-4330-9a07-8ccdd633b873 · outbound

This paper cites Auto-Encoding Variational Bayes.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Auto-Encoding Variational Bayes

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:13.913636Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:13.913636Z digest=sha256:68495e8dc91f46da861dc35323d16aa5eae3ea313f11b66da38baca15f329e60

Observation f9441392-03eb-47ea-88a2-4281bd825553 · outbound

This paper cites In: Proceedings of the fifth International Conference on Knowledge Capture.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: Proceedings of the fifth International Conference on Knowledge Capture

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:14.043032Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:14.043032Z digest=sha256:9a5f23c4f598eb643e3869dbe630ed8ab788ee3a641d8e2d275ebf17fc7e8db3

Observation e1330e3f-97ac-4d75-97ab-7f057f5b8c36 · outbound

This paper cites Jour- nal of the American statistical Association47(260), 583–621 (1952).

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Jour- nal of the American statistical Association47(260), 583–621 (1952)

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:14.089341Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:14.089341Z digest=sha256:be460872124582bf4375743b629befbb1f7b0689616a34bcb28f4e67ba07b64d

Observation 8397e02c-a90f-4027-8cbd-54bbb6192d5c · outbound

This paper cites 2, 2022-06-27.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models 2, 2022-06-27

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:14.194382Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:14.194382Z digest=sha256:da94cc3a093540bfff8a0ebf5549dc775cf02cdc7f5ae2bdf9a9fda103a48f9c

Observation e1848782-22fe-4e43-8502-a06ec1a2c853 · outbound

This paper cites In: Inter- national Conference on Machine Learning.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: Inter- national Conference on Machine Learning

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:14.325447Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:14.325447Z digest=sha256:082be03aac6317bea09baefc19676db271e2ab317063884a2a4af752e9e31718

Observation dcfcf58d-0322-4365-a1e0-dcd6d8635859 · outbound

This paper cites B-Pref: Benchmarking Preference-Based Reinforcement Learning.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models B-Pref: Benchmarking Preference-Based Reinforcement Learning

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:14.354404Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:14.354404Z digest=sha256:09ce1f9a23db996a7d654a5462404243604622889e58654d3f9b4fe5cd97e11a

Observation ee0177bf-6ff1-4f8b-9caa-d4173551a443 · outbound

This paper cites Scalable agent alignment via reward modeling: a research direction.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Scalable agent alignment via reward modeling: a research direction

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:14.400874Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:14.400874Z digest=sha256:5c1191630d5f03f7c1ddb69344ba70a66309fccec6a3e13437f19b6c1aa6e4eb

Observation 24c366b6-3ec2-4869-ad26-a94856ca2f87 · outbound

This paper cites Cognitive Computation17(5), 1–16 (2025).

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Cognitive Computation17(5), 1–16 (2025)

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:14.457876Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:14.457876Z digest=sha256:52d261a49b0925d6f495d3c0f8d81dfcef97aa32ac837d458840660041d32299

Observation 026496bb-fa81-43a1-9971-17339f61fec6 · outbound

This paper cites IEEE Transactions on Vehicular Technology (2026).

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models IEEE Transactions on Vehicular Technology (2026)

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:14.505942Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:14.505942Z digest=sha256:77fbf612209594888d8768f1d5316a0f1639088f7d7390395ac9f302b3310129

Observation 0d9527da-7e3d-4c08-8d4e-ff9b119897a6 · outbound

This paper cites In: 2023 IEEE International Conference on Robotics and Automation (ICRA).

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: 2023 IEEE International Conference on Robotics and Automation (ICRA)

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:14.562784Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:14.562784Z digest=sha256:62a85d1412f607da3ad92879324f89a879f673ced9348f8bfadfd40d06832434

Observation 1160be37-5f4b-4398-84d8-04913db487b9 · outbound

This paper cites In: Proceedings of the 23rd International Conference on Autonomous Agents and Mul- tiagent Systems.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: Proceedings of the 23rd International Conference on Autonomous Agents and Mul- tiagent Systems

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:14.618551Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:14.618551Z digest=sha256:28b240773503328404f5c2bd621f4c38385fd9fc0a727fb8e7b6b8e9a43acd8c

Observation 4c63f1e5-a823-43a7-bfe6-55c99a5222bb · outbound

This paper cites Advances in Neural Information Processing Systems pp.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Advances in Neural Information Processing Systems pp

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:14.667389Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:14.667389Z digest=sha256:8ffc4bbe2fc17b3c3ddbedb7f023aa14ac858f4885e21929c0c00ea9e7e1fe05

Observation 776cf99f-9ddf-4ee4-8df3-3f8ee777ce29 · outbound

This paper cites Journal of Machine Learning research9(Nov), 2579–2605 (2008).

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Journal of Machine Learning research9(Nov), 2579–2605 (2008)

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:14.786967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:14.786967Z digest=sha256:b59dffd61a5c1f8b3dd719c7e2ae5180e51faa8fd69870cfd0453f966de9bf2f

Observation 957e170d-daf2-41a7-8f9a-86a380d4fbe1 · outbound

This paper cites In: The Fourteenth International Conference on Learning Representations (2026).

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: The Fourteenth International Conference on Learning Representations (2026)

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:14.845611Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:14.845611Z digest=sha256:eb293ecff931f0b2e6df2c4c7e256a95b5374ff89f1a8eed606b2938ec7b5171

Observation 0fc26d47-98bb-47c6-b80c-6a3d8d2b6478 · outbound

This paper cites 50–60 (1947).

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models 50–60 (1947)

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:14.967182Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:14.967182Z digest=sha256:a0063d0105e9e91a861eb4a436bbdfba3ee43c58da35f0c016d4e606cd26a520

Observation 3442ae48-deb0-4e1e-bd19-aa877764b219 · outbound

This paper cites an unresolved cited work.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Unresolved cited work

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:15.086757Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:15.086757Z digest=sha256:3e78abd838b0b05334dc1b59c27cd6be86589b07b79e98d4e984f6932f37c685

Observation af83dc68-0a49-4aec-b7bf-e314cb8d7e4d · outbound

This paper cites In: Pro- ceedings of the Seventeenth International Conference on Machine Learning.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: Pro- ceedings of the Seventeenth International Conference on Machine Learning

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:15.267967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:15.267967Z digest=sha256:6f6e57b5967a698d0be377ff9ab4b8018ab954bf486e6a80d480608c4401a6fe

Observation 3885f59c-497c-4dc9-ad4c-283c9b24e31a · outbound

This paper cites Advances in Neural Information Processing Systems pp.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Advances in Neural Information Processing Systems pp

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:15.330406Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:15.330406Z digest=sha256:7575856d94c515d6e9212040c120907d461e6bf3437a3b8b057241e85c83ff2b

Observation 0aaabe6b-ca77-476c-97bc-8734150627b3 · outbound

This paper cites In: 10th International Conference on Learning Representa- tions, ICLR 2022 (2022).

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: 10th International Conference on Learning Representa- tions, ICLR 2022 (2022)

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:15.433542Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:15.433542Z digest=sha256:0edfc9296327ba0d5cc481efa0dac553ad04109214e0612bd7427a86f8c48489

Observation 074b221b-1036-4f30-aa36-8b18886db844 · outbound

This paper cites Advances in neural information processing systems36, 53728–53741 (2023).

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Advances in neural information processing systems36, 53728–53741 (2023)

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:15.548909Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:15.548909Z digest=sha256:d46860bd51406c4531ab639a937230d525dcccdfbb602822db6412b5ba2acb47

Observation 90203423-ee24-41a8-b5a0-79d12db101fe · outbound

This paper cites In: Learning for Dynamics and Control.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: Learning for Dynamics and Control

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:15.674885Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:15.674885Z digest=sha256:ff6e44143fb420c961a447cb0fe2024ceeeea3b97426bcba766b746383cc0c97

Observation b552c757-6c46-45d1-99d0-e6dd2803d23b · outbound

This paper cites Safe Deep RL in 3D Environments using Human Feedback.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Safe Deep RL in 3D Environments using Human Feedback

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:15.779669Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:15.779669Z digest=sha256:69162636ca3aedc8c8b6d3a05a3fd39f72bff77969fd34aa91d20db008b84410

Observation 3250ad68-b736-426a-81d7-5b520946a97e · outbound

This paper cites In: International Conference on Machine Learn- ing.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: International Conference on Machine Learn- ing

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:15.872266Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:15.872266Z digest=sha256:b57393f8af2be50a1daefa7ac0d3ccf7843cc67676c56e9c80c2bfe795f40c01

Observation c3a928fa-c4bb-47e8-a354-044fada585b5 · outbound

This paper cites In: Proceedings of the 17th Interna- tional Conference on Autonomous Agents and MultiAgent Systems.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: Proceedings of the 17th Interna- tional Conference on Autonomous Agents and MultiAgent Systems

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:16.034204Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:16.034204Z digest=sha256:86f28bb0239ba37f2b7a33cfb15613b61fe2bc6164d9b6f28f53934b5ec4bf2a

Observation 92f2741f-5159-475e-8371-d7a4aaa3fcfd · outbound

This paper cites In: Proceedings of the AAAI Conference on Artificial Intelligence.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: Proceedings of the AAAI Conference on Artificial Intelligence

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:16.157222Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:16.157222Z digest=sha256:58a45f27b75e87b14fc293129d2822ff04d1696f4176fd686fee6429b2745127

Observation d76b9425-38b3-4529-aeeb-cb462a2feff1 · outbound

This paper cites an unresolved cited work.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Unresolved cited work

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:16.287126Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:16.287126Z digest=sha256:3374f8ad0027d991557b8b76ec81998fa6ea7b13ddf22207e1651dbfba4a9366

Observation ff335f8f-a034-48d1-9cce-9f925edaf422 · outbound

This paper cites 12151–12162 (2020).

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models 12151–12162 (2020)

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:16.398947Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:16.398947Z digest=sha256:46814c3e4681fa7a57afbf0584cf22b3e28a7c67517c1102e17fe9ecaacad476

Observation 20b95e36-b785-4b08-ad8e-5b64e3732c0b · outbound

This paper cites an unresolved cited work.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Unresolved cited work

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:16.496029Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:16.496029Z digest=sha256:16552deaac0063d9446a8b47b2979120e9a78da99fb5d09e1d7ea8918b11fac3

Observation e991026e-41f1-4826-b093-b6be673af8c7 · outbound

This paper cites Advances in Neural Information Processing Systems36, 29252–29272 (2023).

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Advances in Neural Information Processing Systems36, 29252–29272 (2023)

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:16.638621Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:16.638621Z digest=sha256:adeabb7678a95542ea7ec5e43dc9813ebcf656dc62bb379a2c68ef60dfbc1504

Observation 08edd2ea-965a-4942-b0fb-1261872a5187 · outbound

This paper cites Advances in Neural Information Processing Systems34, 20759–20771 (2021).

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Advances in Neural Information Processing Systems34, 20759–20771 (2021)

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:16.836785Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:16.836785Z digest=sha256:dcdcc8387cddaaa989976f575fa751865ef03cc801b50bebe06b55fcbf8f0b93

Observation f3072b85-af77-423a-892d-e7035bf9f231 · outbound

This paper cites In: Proceedings of the AAAI conference on artificial intelligence.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: Proceedings of the AAAI conference on artificial intelligence

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:16.994533Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:16.994533Z digest=sha256:65671999226942a9e17df78397a90ae62747b10c4233b8f6c6e839e31bad32bf

Observation e7cbe25a-5236-4692-a1e6-7365ea47ed3f · outbound

This paper cites In: 9th International Conference on Learning Representations, ICLR 2021 (2021).

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: 9th International Conference on Learning Representations, ICLR 2021 (2021)

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:17.157234Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:17.157234Z digest=sha256:7954c20b0dd4a50640bf42d057c9aecb79b8b1953260ede2c005ef87cc16ad67

Observation 693bc6dd-be15-4198-bb61-a348d4c92b05 · outbound

This paper cites In: Deep RL Workshop NeurIPS 2021 (2021).

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: Deep RL Workshop NeurIPS 2021 (2021)

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:17.253761Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:17.253761Z digest=sha256:128d07a2da0457cf3c6020ec94125657c36cb53f784e2899f85181b8f8643ba5

Observation 173032a1-72fe-4220-aeee-4099c2a0edf1 · outbound

This paper cites Advances in Neural Information Processing Systems35, 2608–2621 (2022).

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Advances in Neural Information Processing Systems35, 2608–2621 (2022)

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:17.437226Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:17.437226Z digest=sha256:74ae4906fb5287c9a5a80e2de94c21cbb75ab0477dcdcbc3d389da07a383d375

Observation ff9edad7-bbf7-4f5f-bd6a-993e64a61a78 · outbound

This paper cites In: International Conference on Machine Learning.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: International Conference on Machine Learning

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:17.543845Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:17.543845Z digest=sha256:e992f0c2ccb8d4b1746cf5e4cfe5a1fc8223094cf889e01be53c500143d73c19

Pith citing papers

No inbound Pith citation observations are available.