Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 14 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 50 inbound Pith citation observations for arXiv:2209.00626.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-14T04:17:59.786686Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
63
pith, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation fe5749d7-f45c-4d5b-ac2d-231543115114 · inbound
Scaling Laws for Reward Model Overoptimization The Alignment Problem from a Deep Learning Perspective
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation becee435-783a-4ffe-8552-f9f561b98347 · inbound
Sparse Autoencoders Find Highly Interpretable Features in Language Models The Alignment Problem from a Deep Learning Perspective
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 80067b9f-aa3c-4a40-b1c4-1dc44c00c48a · inbound
Data-Centric Foundation Models in Computational Healthcare: A Survey The Alignment Problem from a Deep Learning Perspective
Reference 210
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 3fdd8724-59e7-4815-a791-096d637ecd14 · inbound
Sparse Feature Circuits: Discovering and Editing Interpretable Causal Graphs in Language Models The Alignment Problem from a Deep Learning Perspective
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 7f903738-a102-44b9-93bd-b147b5d04099 · inbound
Sycophancy to Subterfuge: Investigating Reward-Tampering in Large Language Models The Alignment Problem from a Deep Learning Perspective
Reference 187
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation ba9ab7f7-df82-4e25-b00d-d10996e49f89 · inbound
Safety case template for frontier AI: A cyber inability argument The Alignment Problem from a Deep Learning Perspective
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2d055cf9-5346-4e77-be64-37b7d71acfff · inbound
Beyond the Safety Bundle: Auditing the Helpful and Harmless Dataset The Alignment Problem from a Deep Learning Perspective
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 88cbea12-254c-4fed-b589-e85e52bdf171 · inbound
Can an AI Agent Safely Run a Government? Existence of Probably Approximately Aligned Policies The Alignment Problem from a Deep Learning Perspective
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9275644b-34da-4f51-bbb5-5d4707c86caa · inbound
Open Problems in Machine Unlearning for AI Safety The Alignment Problem from a Deep Learning Perspective
Reference 100
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 35b96a1d-149d-4731-a7c2-12cf6d6dc780 · inbound
Gradual Disempowerment: Systemic Existential Risks from Incremental AI Development The Alignment Problem from a Deep Learning Perspective
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e0b1aa67-3e4e-4af0-b51d-799c32909168 · inbound
Compromising Honesty and Harmlessness in Language Models via Deception Attacks The Alignment Problem from a Deep Learning Perspective
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 69e0e3d4-2040-4c0b-8b68-1c21e8c6a2e8 · inbound
From Directions to Cones: Exploring Multidimensional Representations of Propositional Facts in LLMs The Alignment Problem from a Deep Learning Perspective
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8e0d26c6-80e1-46c1-9551-0cc53537e299 · inbound
Evaluating LLM Agent Adherence to Hierarchical Safety Principles: A Lightweight Benchmark for Probing Foundational Controllability Components The Alignment Problem from a Deep Learning Perspective
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation de07081a-edd7-4381-b169-5490e630346d · inbound
Will artificial agents pursue power by default? The Alignment Problem from a Deep Learning Perspective
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d9f40715-cc08-47b5-80ec-c53063625578 · inbound
Deontically Constrained Policy Improvement in Reinforcement Learning Agents The Alignment Problem from a Deep Learning Perspective
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 972a95bd-1328-4030-a171-307a97df5f0c · inbound
Out of Control -- Why Alignment Needs Formal Control Theory (and an Alignment Control Stack) The Alignment Problem from a Deep Learning Perspective
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c8ab42da-82a3-43ea-a7ff-b988c3df64ca · inbound
Evolving Prompts In-Context: An Open-ended, Self-replicating Perspective The Alignment Problem from a Deep Learning Perspective
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c32a77d8-2d1a-4f83-8f61-4c5584cf15a4 · inbound
Agent-to-Agent Theory of Mind: Testing Interlocutor Awareness among Large Language Models The Alignment Problem from a Deep Learning Perspective
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 29342e78-4d90-4680-9744-bc205f44e80d · inbound
Thinking Beyond Tokens: From Brain-Inspired Intelligence to Cognitive Foundations for Artificial General Intelligence and its Societal Impact The Alignment Problem from a Deep Learning Perspective
Reference 219
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation efb17172-59eb-40dd-a6ba-dc2aaa5117ae · inbound
Lessons from a Chimp: AI "Scheming" and the Quest for Ape Language The Alignment Problem from a Deep Learning Perspective
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7579f224-aec7-4c7a-b527-bf7b68bc9725 · inbound
Manipulation Attacks by Misaligned AI: Risk Analysis and Safety Case Framework The Alignment Problem from a Deep Learning Perspective
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9290f995-4ac4-4161-91df-229157a44fe6 · inbound
On the Inevitability of Left-Leaning Political Bias in Aligned Language Models The Alignment Problem from a Deep Learning Perspective
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 33df2d56-405c-4584-b48d-f0f05d6b4a74 · inbound
Mechanistic Exploration of Backdoored Large Language Model Attention Patterns The Alignment Problem from a Deep Learning Perspective
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e369c779-39c4-4540-be9b-f88018797938 · inbound
Human-AI Complementarity: A Goal for Amplified Oversight The Alignment Problem from a Deep Learning Perspective
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bad157e3-c7d8-4752-9875-64b18cc0ca4e · inbound
Language Model Circuits Are Sparse in the Neuron Basis The Alignment Problem from a Deep Learning Perspective
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2b2b6fd6-4280-4d26-866d-b1890dbd293a · inbound
An Onto-Relational-Sophic Framework for Governing Synthetic Minds The Alignment Problem from a Deep Learning Perspective
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation e1c90d48-7394-4748-8c3a-a1470b975659 · inbound
Framing Effects in Independent-Agent Large Language Models: A Cross-Family Behavioral Analysis The Alignment Problem from a Deep Learning Perspective
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 47bb1925-c2b9-463a-800c-14aee472f533 · inbound
Safety, Security, and Cognitive Risks in World Models The Alignment Problem from a Deep Learning Perspective
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 8ba639fd-1849-487f-bc2b-ae7d9f8abb23 · inbound
Cognitive Comparability and the Limits of Governance: Evaluating Authority Under Radical Capability Asymmetry The Alignment Problem from a Deep Learning Perspective
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 1c211139-d19d-4576-a4c8-b30f3669bb99 · inbound
Terminal Wrench: A Dataset of 331 Reward-Hackable Environments and 3,632 Exploit Trajectories The Alignment Problem from a Deep Learning Perspective
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation b9088d91-92c3-476b-b250-2d987df8f3f9 · inbound
Relative Principals, Pluralistic Alignment, and the Structural Value Alignment Problem The Alignment Problem from a Deep Learning Perspective
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation dd349a6b-154b-45e5-8c58-c9bdb20405b2 · inbound
Spurious Correlation Learning in Preference Optimization: Mechanisms, Consequences, and Mitigation via Tie Training The Alignment Problem from a Deep Learning Perspective
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 5e13343c-e132-49c4-a373-33625775d819 · inbound
Who Owns This Agent? Tracing AI Agents Back to Their Owners The Alignment Problem from a Deep Learning Perspective
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 17b2db02-300d-4e4e-bc92-da4f815dc981 · inbound
Understanding Goal Generalisation in Sequential Reinforcement Learning The Alignment Problem from a Deep Learning Perspective
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation eec751e3-a391-4e71-8792-7a21e330318b · inbound
Temporal Preference Concepts and their Functions in a Large Language Model The Alignment Problem from a Deep Learning Perspective
Reference 81
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation cba9c881-9c24-4a4d-a12b-239c951161cf · inbound
Temporal Preference Concepts and their Functions in a Large Language Model The Alignment Problem from a Deep Learning Perspective
Reference 81
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5e41e396-90c9-4ddf-9ede-91913fab68b7 · inbound
Misaligned AI as a New Insider Risk The Alignment Problem from a Deep Learning Perspective
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation f216dbdc-5887-40b9-8baf-b10e1ef4225c · inbound
Enhancing AI Interpretability with Localised Architectures The Alignment Problem from a Deep Learning Perspective
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 1f89b684-e6f5-48a8-91a1-ef694013a808 · inbound
The Agentic Web Requires New Normative Infrastructure The Alignment Problem from a Deep Learning Perspective
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation f7ed8e7a-2c63-4b96-9d4a-aaf0ffaec409 · inbound
Decoding Hidden Deception in Reasoning LLMs: Activation Explainers for Deception Auditing The Alignment Problem from a Deep Learning Perspective
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 48be50e9-231b-48f4-a39b-3da11af23294 · inbound
Safety from Honesty in a Disinterested AI Predictor The Alignment Problem from a Deep Learning Perspective
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 8aa9d0fd-2017-4d11-8eaf-cfb0e2b86185 · inbound
Safety from Honesty in a Disinterested AI Predictor The Alignment Problem from a Deep Learning Perspective
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d36bd4a5-e4af-4fdc-be91-b373b5c89ff1 · inbound
A Scalable Approach to Evaluating Moral Sensitivity in LLMs The Alignment Problem from a Deep Learning Perspective
Reference 135
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 610c9187-8023-460f-92c3-0dab0ec79e0c · inbound
User identity conditions moral wrongness ratings in non-reasoning large language models The Alignment Problem from a Deep Learning Perspective
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 48ff409a-ace0-4170-99e4-352dff6e8235 · inbound
Persuasion Attacks Can Decrease Effectiveness of CoT Monitoring The Alignment Problem from a Deep Learning Perspective
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 53f8cf84-4132-43b9-b21f-e9005a1172ba · inbound
Hardware Mechanisms to Dynamically Throttle AI Performance The Alignment Problem from a Deep Learning Perspective
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f26ac867-d237-4dfe-8199-f00c18ea6625 · inbound
S2T-RLHF: Hierarchical Credit Assignment for Stable Preference-Based RLHF The Alignment Problem from a Deep Learning Perspective
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8a82ba0b-24ce-44b2-bb39-f914401fc253 · inbound
Draining the Energy Commons: Self-Defeating Over-Appropriation as a Coordination Failure in Agentic LLM Collectives The Alignment Problem from a Deep Learning Perspective
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1814b975-231d-4f06-9a20-ce47150dea6f · inbound
Why Study Emergent Behavior When You Can Regulate It? Aligning Multi-Agent Systems with Reward Prediction The Alignment Problem from a Deep Learning Perspective
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 745f7d37-64b8-4d71-ac37-90849517e8c3 · inbound
Evaluation-Conditioned Training: Teaching Models to Generalize to Stronger Oversight Regimes The Alignment Problem from a Deep Learning Perspective
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.