Pith. sign in

Paper Citation Record · LEDGER

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework

As of 17 August 2026, this Paper Citation Record lists 100 of 230 outbound references and 2 inbound Pith citation observations for arXiv:2411.11761.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.11761 v2

Coverage vector

measured 100 of 230 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T18:15:15.675983Z

measured 102 of 102 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T04:49:43.466155Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-07T13:53:42.619268Z

Reference resolution

100 of 230 outbound references displayed

  • verified exact6
  • verified fuzzy0
  • unresolved91
  • parse uncertain0
  • malformed identifier2
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 2d752c3b-4af4-4f00-b275-65f66a3e1032 · outbound

This paper cites Adadi and M.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Adadi and M

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.258937Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.258937Z digest=sha256:5a13f5d8a64c6e04ad67dd8955c5a60490d3b6912ffae946c7688e33b6643bcb

Observation 36d29749-6c96-4768-aa18-49df12a208c0 · outbound

This paper cites an unresolved cited work.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Unresolved cited work

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.263763Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.263763Z digest=sha256:3fe6781b71ab9d666bbeb530fe13b49f72d1fb2c310c983a29ca001c40af0941

Observation c744cfb5-98ee-4085-b2a7-0300021c3d4f · outbound

This paper cites an unresolved cited work.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Unresolved cited work

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.268086Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.268086Z digest=sha256:196fec3b65de4e7ee6255b3165bd4b618207f3ff2497d16ff6dc2f9718b340c3

Observation db9a6119-f6a3-4d68-805c-3e3e11417fd2 · outbound

This paper cites an unresolved cited work.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Unresolved cited work

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.275852Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.275852Z digest=sha256:c3ce9aa9a469588bad582528bb87c7be75f79a7728152f717bb588ea246dd874

Observation 13bbccc6-de27-46bd-a91c-bbe019bf6ea7 · outbound

This paper cites an unresolved cited work.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Unresolved cited work

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.279608Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.279608Z digest=sha256:ea6e5ac938bbc668b9124d9a91227c8e3e47c313796c2da6f2fc31cb9448e1e0

Observation 592dbe1c-b2fd-4899-9bc2-e8401a19f5ec · outbound

This paper cites an unresolved cited work.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Unresolved cited work

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.283943Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.283943Z digest=sha256:bb1bd484f472d6841ab0fcaac537e397c0b9fa469f7558005a6d65412c96cc3a

Observation 740dc0cc-0548-4261-913b-0478e9da3000 · outbound

This paper cites an unresolved cited work.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Unresolved cited work

Reference 7

Resolution
malformed identifier
no resolver link, observed 2026-08-12T18:15:15.287643Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.287643Z digest=sha256:6324f9e185a4ef8f74d5571e8589a1669aedfe9903bce3fa218da0e41f821142

Observation b0a8958e-c780-4598-8055-c817b9e408e4 · outbound

This paper cites an unresolved cited work.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Unresolved cited work

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.291546Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.291546Z digest=sha256:0930c7019d8ad46fdcf5bc59b23d79b48c7b6cf1d909ce466134413720e1357d

Observation bd4a84aa-fd08-4199-8ee8-92b6c71ab8f7 · outbound

This paper cites DQN-TAMER: Human-in-the-Loop Reinforcement Learning with Intractable Feedback.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework DQN-TAMER: Human-in-the-Loop Reinforcement Learning with Intractable Feedback

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.295424Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.295424Z digest=sha256:8a8a818e31a892df7d3d449dfe1e12d0ce2263bf40cb6c2c2c8f504f07539428

Observation 905d39ff-fb8e-4b02-bcde-f5ec00f581b0 · outbound

This paper cites an unresolved cited work.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Unresolved cited work

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.299331Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.299331Z digest=sha256:f9f7633521e646733c6e4f51ce561d41dea8fe4b3bfdcc94aed99d0bd6ffbc6e

Observation db2cd15f-48d0-49fb-97d8-8bcf48878317 · outbound

This paper cites an unresolved cited work.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Unresolved cited work

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.302932Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.302932Z digest=sha256:90f38b002fc9606433ae5f91f8b4a4acf577345bca07178a2185dc879edf3bb6

Observation 6de264a6-cda9-4a00-b558-30f3dc4f5026 · outbound

This paper cites an unresolved cited work.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Unresolved cited work

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.306793Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.306793Z digest=sha256:9b8a8977ae648de96d86ee84013d4e57f99f5647da8fd2c9cc4c91681afc2db6

Observation 23962868-4990-4e29-a0a0-a5cd9387bf58 · outbound

This paper cites an unresolved cited work.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Unresolved cited work

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.314980Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.314980Z digest=sha256:a9ea7fcf0a9e6c6f0c27b2f8b40385688e81af66ec923990de62861fdc85bd7f

Observation 4ee6cdf7-7508-4c60-a157-e2c7becbdaf3 · outbound

This paper cites Brown, Jack Clark, Sam McCandlish, Chris Olah, and Jared Kaplan.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Brown, Jack Clark, Sam McCandlish, Chris Olah, and Jared Kaplan

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.318816Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.318816Z digest=sha256:6b892e7d1fc209710c9141c0d0d97a150d28eec8c093fa61ac074af723706b56

Observation 34b61ad4-3b72-45ce-9f37-e448032c107b · outbound

This paper cites an unresolved cited work.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Unresolved cited work

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.327265Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.327265Z digest=sha256:f7b1b70e91235bca560673478b181b4018b3e4b4c891903416f6926e629371f3

Observation d440e91f-6329-49e5-b887-d28c43542209 · outbound

This paper cites Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.331348Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.331348Z digest=sha256:6bd615a8e2b4ce2405d7454e25dbac7b92dbba90023ed7e9af66308f7e14a97c

Observation 2ba36619-3802-424d-8632-ecac70de1879 · outbound

This paper cites Bowman, Zac Hatfield-Dodds, Ben Mann, Dario Amodei, Nicholas Joseph, Sam McCandlish, Tom Brown, and Jared Kaplan.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Bowman, Zac Hatfield-Dodds, Ben Mann, Dario Amodei, Nicholas Joseph, Sam McCandlish, Tom Brown, and Jared Kaplan

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.335705Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.335705Z digest=sha256:028d73f75c728358a0168ed91e54024379e25686a002fdf80b778a52063e71fa

Observation d954d0a2-456f-40a9-b1e2-272a80e3d77c · outbound

This paper cites an unresolved cited work.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.344407Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.344407Z digest=sha256:fec4d787abc6a3eb585a74deffff501c7461bab0e0d040163577ed624fa89e93

Observation 5b6e0fa8-a1cd-4c28-8b38-54da9874cea9 · outbound

This paper cites an unresolved cited work.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Unresolved cited work

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.348418Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.348418Z digest=sha256:c5be2b076a7c0fbd8a2ee5ca824fddfc7d1848ed6bef5565f20a57a6cd89a5e1

Observation d126d785-1c3a-4ba3-a0e6-b4c3ae9b182c · outbound

This paper cites an unresolved cited work.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Unresolved cited work

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.352225Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.352225Z digest=sha256:82d4d887005d391d43bb2bff2f0c8a4b1a2dfa010f86cd3a33b7842fde552db8

Observation 1da7c018-86e7-4336-9842-29615477044e · outbound

This paper cites an unresolved cited work.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Unresolved cited work

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.355976Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.355976Z digest=sha256:f44cb49a5200bb7da26a1cb8d5384baba939d9cc24637af0b2a33b40f3241e19

Observation 2c7e5665-774f-4c4f-8957-a796713b0364 · outbound

This paper cites an unresolved cited work.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Unresolved cited work

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.359518Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.359518Z digest=sha256:b241b3a8da219253b708e684e4eed05af1fd41817512c596ccb08881edb6dd00

Observation b0578416-43fc-48ff-88fe-f856c8ea42f3 · outbound

This paper cites an unresolved cited work.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Unresolved cited work

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.363125Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.363125Z digest=sha256:b8e4876102fff90f5018d616651f0f92b2fcd9f37af7d645360e64b689809dbf

Observation a948aefc-a329-4426-b398-7355029ecc9f · outbound

This paper cites an unresolved cited work.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Unresolved cited work

Reference 24

Resolution
malformed identifier
no resolver link, observed 2026-08-12T18:15:15.367687Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.367687Z digest=sha256:4b33ee85c42c0cb87f4162bcef782abfc761bb28a59144678e9c02148e12a80c

Observation 6f5ea255-f732-408a-8c8c-ecbccda994d2 · outbound

This paper cites an unresolved cited work.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Unresolved cited work

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.371612Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.371612Z digest=sha256:9accc72e975bd5f5761d472dc4d1b42ebcff07cd112e1f972faabaee482e63d6

Observation 9a579aed-2e01-44ce-8dd0-61d706c1cd06 · outbound

This paper cites an unresolved cited work.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Unresolved cited work

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.375186Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.375186Z digest=sha256:765d7bc4cbc0a15db535116d79dabe9b051fb58b1a9c64c4c75d0a148fc65ff7

Observation 826a0476-31ee-495d-ae49-e9e4b5c7579d · outbound

This paper cites an unresolved cited work.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Unresolved cited work

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.379059Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.379059Z digest=sha256:210a5093dd1df4cd5fecc5c05eb91bac462f98c912c43cccbec71b1464f2f006

Observation 50d6eddd-fc39-4a80-9213-bfd7c86056f1 · outbound

This paper cites an unresolved cited work.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Unresolved cited work

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.383373Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.383373Z digest=sha256:b4bd9b14ac18ef4937e5b52c1baec63905445d1cc7b9fdde0760644cd742ee1b

Observation 2afc0f67-8c83-4a4e-b931-cd2bf5874f62 · outbound

This paper cites Bradley Knox and Peter Stone.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Bradley Knox and Peter Stone

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.386633Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.386633Z digest=sha256:362f7ddb14117ed587d803c935b894a305a42381e281e5f7212523bae8e65a43

Observation 236b5316-fb2e-449d-a914-67a1bd83195e · outbound

This paper cites an unresolved cited work.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Unresolved cited work

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.390454Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.390454Z digest=sha256:b6ad50decbad392cc9a052336f0f5b4807646e2782d88f0375242abf59d24fe7

Observation 38053e51-ec49-4493-94ba-34a653078492 · outbound

This paper cites an unresolved cited work.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Unresolved cited work

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.393863Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.393863Z digest=sha256:a676a252977ac49dbdcfb9172e25e8869b99349a39a381d654257e107d783275

Observation f16851a9-c73e-459c-afd7-1a2ec0fcc02e · outbound

This paper cites Extrapolating Beyond Suboptimal Demonstrations via Inverse Reinforcement Learning from Observations.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Extrapolating Beyond Suboptimal Demonstrations via Inverse Reinforcement Learning from Observations

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.397911Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.397911Z digest=sha256:069485fd4cf7e935f6254a5c1864ab500e3dd282e544db38e358f2ba8b364d76

Observation 8fab0c89-e996-4125-84e5-0f89efd1e21f · outbound

This paper cites an unresolved cited work.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Unresolved cited work

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.402559Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.402559Z digest=sha256:52f693f5b0d5690db3012ca9dc920f71db1b125bb94e931459df502938e3d644

Observation c505e5b3-85d8-46a6-84d1-fae0a2a3490e · outbound

This paper cites an unresolved cited work.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Unresolved cited work

Reference 34

Resolution
verified exact
doi, observed 2026-08-12T18:15:16.322649Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T18:15:15.406732Z digest=sha256:9b674840509a3fe5aa8241ac3662a6ef5263ef937d8d52e2cd94eecd902c0aad

Observation b1b461c7-ccef-47fc-bdad-d4cc0172a7ed · outbound

This paper cites an unresolved cited work.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Unresolved cited work

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.411079Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.411079Z digest=sha256:94b415b3b3609459c9ace0deda56d563c19f3632198efa1f72c26c295784ef2b

Observation eae2e6d3-c826-4ff8-817a-b9fb153dc9c2 · outbound

This paper cites Open Problems and Fundamental Limitations of Reinforcement Learning from Human Feedback.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Open Problems and Fundamental Limitations of Reinforcement Learning from Human Feedback

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.416006Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.416006Z digest=sha256:21ae325973f0e16954bbaf913991f47c60ed02c848ced186001bf68977249f43

Observation e0ceac96-28be-4c9e-aa20-228beabe503b · outbound

This paper cites an unresolved cited work.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Unresolved cited work

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.420911Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.420911Z digest=sha256:bc59e3259e475724977954322271bda4da201eb88d0893c4f8499069c4450e90

Observation fc4425e0-ea4f-42e5-886a-a1a963056c6f · outbound

This paper cites an unresolved cited work.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Unresolved cited work

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.424902Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.424902Z digest=sha256:ead7f5e5653bbc773a9b6124a335e0192317a9d2e725f928e22d00877de37481

Observation 36d8aedb-c8ea-4ede-af28-61b8102dd00d · outbound

This paper cites BabyAI: A Platform to Study the Sample Efficiency of Grounded Language Learning.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework BabyAI: A Platform to Study the Sample Efficiency of Grounded Language Learning

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.428719Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.428719Z digest=sha256:84479ad363d23dc77d1e07475ae0ba673bec12a98737e6ca0f268a9becbbc15b

Observation 3d141350-6a57-4050-a175-d00bc0df049b · outbound

This paper cites an unresolved cited work.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Unresolved cited work

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.433497Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.433497Z digest=sha256:197e7b227fecb56550c1d5d002add49155f9042060bf70f69b2b39cdc7f323f3

Observation 403a22a5-17f6-408b-84e8-a315bd384959 · outbound

This paper cites an unresolved cited work.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Unresolved cited work

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.436841Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.436841Z digest=sha256:551182e124279ddfa897160403f94a976d2ca8e9aba2a3fef46523bd48dc824c

Observation 22faeb14-3a0f-481a-b690-ff5ca53cfe7b · outbound

This paper cites an unresolved cited work.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Unresolved cited work

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.441619Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.441619Z digest=sha256:63344169737f7289da287ae580963b1960cbc4fbc5c3218f387be6f776bc4e72

Observation b42517da-a5a0-497a-b10e-35326d6dcb3c · outbound

This paper cites Using Machine Teaching to Investigate Human Assumptions when Teaching Reinforcement Learners.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Using Machine Teaching to Investigate Human Assumptions when Teaching Reinforcement Learners

Reference 43

Resolution
verified exact
local_arxiv, observed 2026-08-12T18:15:17.503411Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T18:15:15.446029Z digest=sha256:b59ad1434a9fc0cd7ccc3664d479cd17346e99de9ecc41ebd2c360f237bb07f7

Observation c456124d-b35c-426d-b99d-18971ec5dd30 · outbound

This paper cites an unresolved cited work.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Unresolved cited work

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.449995Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.449995Z digest=sha256:bdcc0e94b353aee3f5fe43195438d3b5330f2e94620a20cb613897e31b2011fa

Observation c887c3fd-fdad-4d88-89a9-185e1cbb5972 · outbound

This paper cites an unresolved cited work.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Unresolved cited work

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.453645Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.453645Z digest=sha256:4c33b31b07872067f61b6ea836e90a5422d015e311d1218d5382785b3eaac437

Observation b53fb3c5-d9df-4863-bdca-79e6bbd58194 · outbound

This paper cites Broad-persistent Advice for Interactive Reinforcement Learning Scenarios.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Broad-persistent Advice for Interactive Reinforcement Learning Scenarios

Reference 46

Resolution
verified exact
local_arxiv, observed 2026-08-12T18:15:17.486576Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T18:15:15.457040Z digest=sha256:91de83bb37a222a518327ea0dd1ce179ca6a983ec08277edeee50bde5e9ff63a

Observation cb0eaad1-5adf-47c4-8876-9fc4e8912278 · outbound

This paper cites Parisi, and Stefan Wermter.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Parisi, and Stefan Wermter

Reference 47

Resolution
metadata mismatch
raw_fallback, observed 2026-08-12T18:15:17.469255Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T18:15:15.461751Z digest=sha256:ac3d2d58f750e5561354faba5eccd3812654092e6291d5a4c4568012b112b969

Observation 1d1779e8-eb2d-4665-bd6f-8df5c2d5211f · outbound

This paper cites an unresolved cited work.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Unresolved cited work

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.465277Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.465277Z digest=sha256:b4ab9c0809d6630271e25a1e450d2051709e322133bdbeabd3ae86efce4e0aa8

Observation f17038cd-ed4c-4876-9a2b-b22c339cad24 · outbound

This paper cites an unresolved cited work.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Unresolved cited work

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.469194Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.469194Z digest=sha256:42fba39333757dcb0868394cadf9b3d3c3b0c567032cce0931a555d02f0def2a

Observation 72ac4b67-0907-4a47-a959-ed8a1279871c · outbound

This paper cites an unresolved cited work.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Unresolved cited work

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.473608Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.473608Z digest=sha256:d72546a247f2e1e2254058357a42d141ed77cee25ea98bc651f316482c99b46a

Observation 31423784-e544-4742-9dfe-40caa8ebe970 · outbound

This paper cites Deshpande, Jeff Schneider, Deepak Pathak, David Held, and Benjamin Eysenbach.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Deshpande, Jeff Schneider, Deepak Pathak, David Held, and Benjamin Eysenbach

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.478148Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.478148Z digest=sha256:231695df057c49c1d7aec5a297581e740d5ad2387ad7e72b53c7deb0c1d8765d

Observation e2217a00-6830-49ef-814b-375063034c84 · outbound

This paper cites Benchmarking Deep Reinforcement Learning for Continuous Control.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Benchmarking Deep Reinforcement Learning for Continuous Control

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.481768Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.481768Z digest=sha256:466b447f0254cc2fbe2890966c83e70f01291487d219c546e4effe78add8bae7

Observation 9d9ec737-a588-4e2e-ac96-60b613aae526 · outbound

This paper cites Dudley and Per Ola Kristensson.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Dudley and Per Ola Kristensson

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.486352Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.486352Z digest=sha256:ccf420d2b0644cc5e123480e84c9e053b4c8c02f0ee5f239d05486fbe5d83bcc

Observation 6d5600ad-4b30-412f-a14f-581376d3b9fc · outbound

This paper cites an unresolved cited work.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Unresolved cited work

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.490173Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.490173Z digest=sha256:7dea9eaaa6bd01cc2fd831dd3dceb495c8817c3226982b71c4b1da5c3db6bc85

Observation 85207a17-bd72-48cc-b808-c428b44ff3e5 · outbound

This paper cites an unresolved cited work.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Unresolved cited work

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.494567Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.494567Z digest=sha256:ec23a0a65822ecabd806c72e4b08e5e80cd71c73c2d0cbea3270c7d9707c61c4

Observation abcd76fe-1937-40db-a8b3-61d5df8d7be8 · outbound

This paper cites an unresolved cited work.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Unresolved cited work

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.498105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.498105Z digest=sha256:4945b71549d9e650ce0cfd87dadce35550714d8b7b42b899938f5cf89ef1548a

Observation df9c4a6b-c8da-4d27-8d25-5fb8a43da6a2 · outbound

This paper cites an unresolved cited work.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Unresolved cited work

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.501801Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.501801Z digest=sha256:a8a45e65246de2770e6ad012e865fab610bb33e3e54679e23923bd2553a5aee3

Observation a87e0388-aac1-4f74-b659-7003fdfbe71a · outbound

This paper cites an unresolved cited work.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Unresolved cited work

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.505497Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.505497Z digest=sha256:93067074b4cccb14579f2c40898c5a77229f689f056495dd0aff8e8d8ec31900

Observation d7e3483d-58e5-4aaa-84c3-91182185e39d · outbound

This paper cites an unresolved cited work.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Unresolved cited work

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.509926Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.509926Z digest=sha256:12b11fc66b14b386e34e159c9709014b06167033846e95e6ed1997b68f6bf189

Observation 18b47739-c21b-483a-ae57-d74a7efa4201 · outbound

This paper cites an unresolved cited work.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Unresolved cited work

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.514327Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.514327Z digest=sha256:cd3329b67d3a6849f4a618140227d71d345388e61ad8d43a4c39bebbfe99b103

Observation f7504670-a013-4801-bd64-273da66cc7bc · outbound

This paper cites an unresolved cited work.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Unresolved cited work

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.518319Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.518319Z digest=sha256:090462c0ce6445ad7c5cb7537f62dc6b7c02bda302e522fca14f9d8c1458e326

Observation 399551cb-0ddf-42a6-8e8f-9c8e372ddfbb · outbound

This paper cites an unresolved cited work.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Unresolved cited work

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.522257Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.522257Z digest=sha256:2bd64393c204811cdf6b5fedafc79594b0eeb5696d363608425f750706b56b50

Observation 90955e1d-f4a5-48ea-8ae2-565d661e9c7d · outbound

This paper cites an unresolved cited work.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Unresolved cited work

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.527057Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.527057Z digest=sha256:8571b3ff96ad2aa48571f52d9945e9e47226792a90ed42537a0f3df6b2a1a98c

Observation 98fa9339-e45f-4463-8968-61c5f1eca636 · outbound

This paper cites an unresolved cited work.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Unresolved cited work

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.530756Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.530756Z digest=sha256:10a494eae1c98405d65493a5d6c911d98a7961aa819aa7279445c388c1f15232

Observation 38280829-3121-4a14-a401-71c4d8deb72f · outbound

This paper cites Choice Set Misspecification in Reward Inference.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Choice Set Misspecification in Reward Inference

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.535401Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.535401Z digest=sha256:efa44764631831a06ed52b49a24940cf4d339ef761e64ab82877e6aab033ba31

Observation a70f7910-1e70-47dd-8257-e3949cae8dc5 · outbound

This paper cites Red Teaming Language Models to Reduce Harms: Methods, Scaling Behaviors, and Lessons Learned.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Red Teaming Language Models to Reduce Harms: Methods, Scaling Behaviors, and Lessons Learned

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.539134Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.539134Z digest=sha256:91a786872532e79ac937e7f6332d52a1b2fd96b2adfddfb24a90c4f137aaf2df

Observation cc0915c8-e1e2-4b7e-9f29-311906a8b431 · outbound

This paper cites APRIL: Interactively Learning to Summarise by Combining Active Preference Learning and Reinforcement Learning.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework APRIL: Interactively Learning to Summarise by Combining Active Preference Learning and Reinforcement Learning

Reference 67

Resolution
verified exact
local_arxiv, observed 2026-08-12T18:15:17.312037Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T18:15:15.543119Z digest=sha256:79c965e05000d79258d6e656a8ba9ac5f950a630a6b4f4cf0a8b2fb06ad23641

Observation 959001cf-1533-4d5b-81cb-164c9a358211 · outbound

This paper cites The Effect of Modeling Human Rationality Level on Learning Rewards from Multiple Feedback Types.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework The Effect of Modeling Human Rationality Level on Learning Rewards from Multiple Feedback Types

Reference 68

Resolution
verified exact
local_arxiv, observed 2026-08-12T18:15:17.294535Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T18:15:15.547368Z digest=sha256:2eb2db39de33a1d17e17d03146b336b6a91942ff92dfd834900ef246a2469506

Observation 677e93f4-aab2-4329-8f45-c761e71cee31 · outbound

This paper cites an unresolved cited work.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Unresolved cited work

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.551329Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.551329Z digest=sha256:9ebde9b18672299f9ebbd2f50f572b14a96cec05974a0da00adb7c3b1a0b5e81

Observation 086b3363-fa59-4a44-abf8-e6665cba8458 · outbound

This paper cites an unresolved cited work.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Unresolved cited work

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.554896Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.554896Z digest=sha256:4e03d7da48717bbf3d2f4dd9e73c59a1bbf472f01037c45a50d0511a0439b211

Observation 1b2779a1-80aa-44b3-b3bc-48e8fc79b287 · outbound

This paper cites an unresolved cited work.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Unresolved cited work

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.558855Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.558855Z digest=sha256:ffb0a1bfd3a79023a3f84022181efe26387b5a98ceadd0dfe6764833afd45b37

Observation 8c87efd5-9bc6-4862-932e-96c7caca7973 · outbound

This paper cites an unresolved cited work.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Unresolved cited work

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.562881Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.562881Z digest=sha256:c41e3646451dd1e4d90be5328030cf9170d71f91b1c043a8a5467c6a7b266b94

Observation 781e3788-4e30-41c1-8c23-7a1904671e42 · outbound

This paper cites an unresolved cited work.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Unresolved cited work

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.566355Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.566355Z digest=sha256:7569906a81a648299c2c0b20d00dc9cf1b1015522936ff3da98a8977516799ee

Observation 52f8ca4b-b864-46a9-b15e-274c6c1c986c · outbound

This paper cites Guiding Reinforcement Learning Exploration Using Natural Language.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Guiding Reinforcement Learning Exploration Using Natural Language

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.569994Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.569994Z digest=sha256:33def408f505f0063ea6f8889ecaaf82a326ce4ca4b888e97bd0510ec665aa47

Observation 9f3ca8fb-a359-4e0b-9dd7-bdfc903f89d7 · outbound

This paper cites an unresolved cited work.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Unresolved cited work

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.574353Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.574353Z digest=sha256:f8494f3e741220641fe2583eb76d14bbe5854b46902731ef929119f0530e1a96

Observation 7f4a70ad-7565-4ce6-b5dc-148ef4ca7484 · outbound

This paper cites an unresolved cited work.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Unresolved cited work

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.577947Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.577947Z digest=sha256:d92c3634e397ef1589fafed543918646633934058e820730f11ae718bd657de3

Observation ecd917d6-12ad-4e78-81fe-f54b023edd09 · outbound

This paper cites an unresolved cited work.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Unresolved cited work

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.581791Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.581791Z digest=sha256:0b5dccf941434dfb6df6bcfd614134e31e99f2959e120ea0dca5b42365696ac7

Observation 90414646-2e10-4404-9d7b-fefd054b81ce · outbound

This paper cites an unresolved cited work.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Unresolved cited work

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.586469Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.586469Z digest=sha256:f83d4b5dd4cfe56afcdc82d08bb418e38d7ad3062d46569b72146f064de9a58a

Observation 10cda54b-8be0-47d6-b408-5fe6c27217d4 · outbound

This paper cites an unresolved cited work.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Unresolved cited work

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.589936Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.589936Z digest=sha256:961a4c2e0c4d805c39efde71663b3c2f34d71dd6b9d70839be8f0421ba0bc9d2

Observation 460235c7-d02d-4c01-b567-90d8451b4c7d · outbound

This paper cites Reward-rational (implicit) choice: A unifying formalism for reward learning.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Reward-rational (implicit) choice: A unifying formalism for reward learning

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.593768Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.593768Z digest=sha256:da7ecbfb64350ec02593957cf49e07dc2a1da09f0099cb77a0e8a7cb77736283

Observation 7036c7d0-9413-47a1-95c5-49fd2b504709 · outbound

This paper cites Reinforcement Learning from Human Feedback with Active Queries.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Reinforcement Learning from Human Feedback with Active Queries

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.598159Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.598159Z digest=sha256:ea3631a461091c87e1ed663e9e2d797813e0729f0cbe399fa91e9776658f4edd

Observation bbbd3ca6-8544-4c93-b4ed-cad2e0daad4c · outbound

This paper cites an unresolved cited work.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Unresolved cited work

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.602141Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.602141Z digest=sha256:fc24c5409988b8c6fe8bb9dbe9147142a4ecd1367d58de127849c8fc703b1f99

Observation c77ae9f2-60dc-44f9-95fb-6ab9316319b7 · outbound

This paper cites Unity: A General Platform for Intelligent Agents.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Unity: A General Platform for Intelligent Agents

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.606433Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.606433Z digest=sha256:650644e9f5a4a5b7d78c1d848393d08429836c41b0f13b470f033b194731c840

Observation 9152dc1f-8198-42a5-9725-5b3d25deedf7 · outbound

This paper cites an unresolved cited work.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Unresolved cited work

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.610358Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.610358Z digest=sha256:77fd9e0086bf240f0c1fa0b61722a00fac3013a0188eb4021841dc0535907605

Observation 1a5e3a24-1f6b-4ee2-b526-ce70f5d5aaf8 · outbound

This paper cites an unresolved cited work.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Unresolved cited work

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.614077Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.614077Z digest=sha256:79a0bb693f83dbdea8a150e4a11eaa4cf060d6766b93ee5728db880b026f5c4f

Observation b6a5b2ad-a0f3-4074-93b2-4f77659522f6 · outbound

This paper cites an unresolved cited work.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Unresolved cited work

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.617960Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.617960Z digest=sha256:21106a0963044a72372dd67092cc5f013eb9473eeb88c9c7f5d6c17e59729dd2

Observation 7d0c4ba9-5f89-41fc-baa0-0917ca0df6a3 · outbound

This paper cites an unresolved cited work.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Unresolved cited work

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.621309Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.621309Z digest=sha256:8f5d7b656af1de36aa2aef735e7e80c81e0633d9c6157c349bf9b28771743f64

Observation 28a8f9d2-ef66-4feb-b706-0b4a43bd62b4 · outbound

This paper cites an unresolved cited work.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Unresolved cited work

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.624743Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.624743Z digest=sha256:58e64f857ab5f7e71d7d06c779db7f3744647adb8d9c33afed4743fda6fa88f4

Observation 97b6bf3a-2340-4d28-8df5-5af889e7d89d · outbound

This paper cites an unresolved cited work.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Unresolved cited work

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.629155Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.629155Z digest=sha256:12e0f3f34b0f4141bc29065bd4bbfb014e61f726d0c6cbf896c1540c0f8f8c2f

Observation 56ca0346-2d9e-44f8-8d5b-9cb8d1ca6412 · outbound

This paper cites an unresolved cited work.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Unresolved cited work

Reference 90

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.633116Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.633116Z digest=sha256:aed637abd3b70dd59c3cd67075ef14ca7b5bb7f1a8f62376edfe0f35573de035

Observation 6c0d7e9b-c666-43c8-871a-20cae22bde0b · outbound

This paper cites an unresolved cited work.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Unresolved cited work

Reference 91

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.636948Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.636948Z digest=sha256:f20195eb65ce62733e5b1449956cbe5b3a711c1c760f3fb64db49ea86d4a6843

Observation ff8c5cf9-8561-4054-95f8-db97ca544c7e · outbound

This paper cites an unresolved cited work.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Unresolved cited work

Reference 92

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.640997Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.640997Z digest=sha256:3379432016006af01ccdf7d3c796c75b26f32825a6e8f4842a46371f9a182d57

Observation 25ddf0f6-d3a3-489a-8adb-d24ef185bf13 · outbound

This paper cites Newtonian Action Advice: Integrating Human Verbal Instruction with Reinforcement Learning.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Newtonian Action Advice: Integrating Human Verbal Instruction with Reinforcement Learning

Reference 93

Resolution
verified exact
local_arxiv, observed 2026-08-12T18:15:17.170546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T18:15:15.644741Z digest=sha256:e1efade0c5da8974cf94667c70650d876d3998a55e57e81266b3c3d59f9678f5

Observation 1edb2320-9c7f-4296-bd4a-f986f6cb4c0d · outbound

This paper cites an unresolved cited work.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Unresolved cited work

Reference 94

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.648600Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.648600Z digest=sha256:70f2a1d459d17fba012c5de9c57826576966c3737d9d464a8c0d2f8de5d8496b

Observation a7d13dff-6008-4b40-855c-20eeff157093 · outbound

This paper cites an unresolved cited work.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Unresolved cited work

Reference 95

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.652311Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.652311Z digest=sha256:bfa28d34d7defb9d19101aebcd7e165a5caae60ab89aa75524d44bd11dcb964f

Observation 916bad6c-62a5-49af-b8f6-f6b48381e6ad · outbound

This paper cites an unresolved cited work.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Unresolved cited work

Reference 96

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.655707Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.655707Z digest=sha256:7b0112d936d51937682ccc377ad40a519c050d8aae84e867d3795501448fa1b3

Observation 6e8e549f-d12f-4466-81fd-0e72a7c171b3 · outbound

This paper cites an unresolved cited work.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Unresolved cited work

Reference 97

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.659541Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.659541Z digest=sha256:74104917b51ab718400a0dc95458352c4f202594d1454d7836a09f2950600ba8

Observation c0434bbf-afd9-46d1-bab9-797afd16f6fe · outbound

This paper cites an unresolved cited work.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Unresolved cited work

Reference 99

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.668131Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.668131Z digest=sha256:1e0d3424eeec79a7e1f9106ac71c099439c7fd14f2b2933c8814adc581acf30c

Observation b4e6b976-6b24-49e4-9c2e-bc8d9b8dfa3c · outbound

This paper cites an unresolved cited work.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Unresolved cited work

Reference 100

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.671740Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.671740Z digest=sha256:bbb5f9a9ca532687d64e0b2bf73bfcd76a179e9d38748dfa20a4e6b7eacbb0d6

Observation 63f0c3ea-8a03-49e1-a3b7-d3c65a034cf5 · outbound

This paper cites an unresolved cited work.

Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Unresolved cited work

Reference 101

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:15.675983Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:15:15.675983Z digest=sha256:07e02dccde572ee59eda236321117ce6f027275469ab6e096fd6290c73557144

Pith citing papers

Observation 937736bc-582b-4837-9cce-bfb969b4967d · inbound

Optimal Interactive Learning on the Job via Facility Location Planning cites this paper.

Optimal Interactive Learning on the Job via Facility Location Planning Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-16T04:49:43.466155Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:49:43.466155Z digest=sha256:b6a89baaa104c209cca1af53b1846723a11723c89df68950b1f3d6568f93b3f4

Observation db71e279-304a-40eb-babd-6ef30e623a54 · inbound

LLM-Guided Reinforcement Learning: Addressing Training Bottlenecks through Policy Modulation cites this paper.

LLM-Guided Reinforcement Learning: Addressing Training Bottlenecks through Policy Modulation Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-08-07T13:53:42.823242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T13:53:39.462927Z digest=sha256:69e582984f80d0b290dcb117ac8faf405ac99d360465fddccb9e029abb61c709