Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 12 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 11 inbound Pith citation observations for arXiv:2105.14111.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:29:10.515127Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
22
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation 05521e01-6e5c-40e1-840e-92724b009584 · inbound
Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback Goal Misgeneralization in Deep Reinforcement Learning
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 96bb726c-0a5a-4782-ab76-524ff08148ab · inbound
GPT-NeoX-20B: An Open-Source Autoregressive Language Model Goal Misgeneralization in Deep Reinforcement Learning
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 0d4a7ec3-9db1-4ead-86f7-b2e096ffb788 · inbound
Language Models (Mostly) Know What They Know Goal Misgeneralization in Deep Reinforcement Learning
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 3f4146b5-8cd0-45d1-999b-b8d79be511eb · inbound
Machine Theory of Mind and the Structure of Human Values Goal Misgeneralization in Deep Reinforcement Learning
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dbcccdfa-f105-488b-8733-f836fcf509a4 · inbound
Safety, Security, and Cognitive Risks in World Models Goal Misgeneralization in Deep Reinforcement Learning
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation c2cc7945-baf2-429d-9553-0cc48a0bab08 · inbound
Risk Reporting for Developers' Internal AI Model Use Goal Misgeneralization in Deep Reinforcement Learning
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 1bc17925-2df3-40e4-b0d0-53c169e8e098 · inbound
Safe Multi-Agent Behavior Must Be Maintained, Not Merely Asserted: Constraint Drift in LLM-Based Multi-Agent Systems Goal Misgeneralization in Deep Reinforcement Learning
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation b6825105-12e9-47a2-ac9e-e0bd6fb01f87 · inbound
Correcting Influence: Unboxing LLM Outputs with Orthogonal Latent Spaces Goal Misgeneralization in Deep Reinforcement Learning
Reference 218
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 9ded97c0-7996-493e-8af5-17cec1c19e58 · inbound
Understanding Goal Generalisation in Sequential Reinforcement Learning Goal Misgeneralization in Deep Reinforcement Learning
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation d61fb680-97ee-4070-b2e7-bfd16cc6631b · inbound
When Does Reward Teach State? A Hidden-Automaton Instrument and a Group-Language Warning Signal Goal Misgeneralization in Deep Reinforcement Learning
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8fd585a6-bbfe-4b2f-92e5-961255091755 · inbound
Draining the Energy Commons: Self-Defeating Over-Appropriation as a Coordination Failure in Agentic LLM Collectives Goal Misgeneralization in Deep Reinforcement Learning
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.