Pith. sign in

Paper Citation Record · LEDGER

Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 19 inbound Pith citation observations for arXiv:2407.04694.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2407.04694 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 19 of 19 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 19 of 19 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T05:42:43.460514Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

2
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 89b2d313-17d3-4a53-8ad1-5346b4b3a174 · inbound

Inspection and Control of Self-Generated-Text Recognition Ability in Llama3-8b-Instruct cites this paper.

Inspection and Control of Self-Generated-Text Recognition Ability in Llama3-8b-Instruct Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-23T19:53:23.281474Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-23T19:51:14.630756Z digest=sha256:fe4fbb486b2852e7062b2e5309a7e94446254fedd5a0b3db81f450014b1da936

Observation 652de3e4-a831-4732-9b28-0e8e1c601c8f · inbound

Frontier Models are Capable of In-context Scheming cites this paper.

Frontier Models are Capable of In-context Scheming Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-16T14:22:01.659365Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-16T14:22:01.616448Z digest=sha256:3359a2b9f6821f20871e8ca5530a525d53db472e01fd276cd6d70bdef22b5497

Observation 8274109e-2066-46e9-9ecb-19047f44c99c · inbound

Compromising Honesty and Harmlessness in Language Models via Deception Attacks cites this paper.

Compromising Honesty and Harmlessness in Language Models via Deception Attacks Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-08T05:42:43.460514Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T05:42:43.460514Z digest=sha256:33a323811b9f982a19efee9fe15d4e78196ec7f64dd746e78b30a374177e5118

Observation 23e32c6e-8ca0-4f2e-8a64-9c180c04444a · inbound

Does It Make Sense to Speak of Introspection in Large Language Models? cites this paper.

Does It Make Sense to Speak of Introspection in Large Language Models? Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T10:30:37.638319Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:30:37.638319Z digest=sha256:cf6385ea7e48159b3fa4b48680b50c51a962c71644624cac4f2f7df71c72f3fa

Observation a73ccfc0-6f4b-4647-a8a0-703489c7f742 · inbound

Model Organisms for Emergent Misalignment cites this paper.

Model Organisms for Emergent Misalignment Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T04:07:28.181472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:07:28.181472Z digest=sha256:6eddd1be66c3eca74a2d6d9b3eda7248b8b71191a4e2ce971b0f0f254c10e73e

Observation 2ee47776-1d8d-47e2-b4be-e7fa232e1549 · inbound

Convergent Linear Representations of Emergent Misalignment cites this paper.

Convergent Linear Representations of Emergent Misalignment Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T04:09:26.150070Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:09:26.150070Z digest=sha256:9bb8ab19f9ae33c437fa0dcd7f416e7c7ce8d82e0f44147d4dd0f79639bf319c

Observation 48250677-93fe-4649-9ae7-99e51a99c4c1 · inbound

Safety Features for a Centralised AGI Project cites this paper.

Safety Features for a Centralised AGI Project Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T00:23:56.096143Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:23:56.096143Z digest=sha256:cda6e40e69343831b3b94df7452d25662017ff1a8b4c115a1be35f79efbc60a5

Observation b9940761-e6b3-4bbb-9d9b-261c20e169ab · inbound

Honeypot Protocol cites this paper.

Honeypot Protocol Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-10T14:35:34.140641Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T14:35:16.230357Z digest=sha256:8d16d81784af12ced7ed8eea97ce807c957f298837dea7980c515f960d863d0e

Observation cdf4da5b-7e9c-49d6-88af-84404529e2a1 · inbound

Measuring Evaluation-Context Divergence in Open-Weight LLMs: A Paired-Prompt Protocol with Pilot Evidence of Alignment-Pipeline-Specific Heterogeneity cites this paper.

Measuring Evaluation-Context Divergence in Open-Weight LLMs: A Paired-Prompt Protocol with Pilot Evidence of Alignment-Pipeline-Specific Heterogeneity Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-08T22:04:18.014075Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-08T10:23:02.697982Z digest=sha256:ed92cc04a1d9122ea9d3b303abc2d1c38b6796b7f317d2d6267ee46e162bafa9

Observation 2f90b689-13ce-4f6a-9573-0a5439e1beac · inbound

Do as I Say, Not as I Do: Instruction-Induction Conflict in LLMs cites this paper.

Do as I Say, Not as I Do: Instruction-Induction Conflict in LLMs Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 8

Resolution
metadata mismatch
arxiv_id, observed 2026-05-21T07:34:02.866648Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-21T07:30:27.297971Z digest=sha256:69f3b30e7d329dd3e6f967a8b961a54f93accbbda0496fd4657a8b82d5b3725a

Observation a2620d65-f901-4f6c-abc7-595152701dc7 · inbound

Do as I Say, Not as I Do: Instruction-Induction Conflict in LLMs cites this paper.

Do as I Say, Not as I Do: Instruction-Induction Conflict in LLMs Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 8

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T18:04:58.126158Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-30T18:00:29.237972Z digest=sha256:b335c7979691bfa6adc02ebb89baeb6c86d7543dd46583b6de3078b34871d3ad

Observation b2bb0f61-eb8d-4fef-adc2-ec8a9f3ca6c3 · inbound

AI Integrity: Defending Against Backdoors and Secret Loyalties cites this paper.

AI Integrity: Defending Against Backdoors and Secret Loyalties Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 21

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T14:59:54.636218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-04T14:56:53.806480Z digest=sha256:5e479079d467c4a6332dc050577d4b1f2a11c3858095ad1aa41168263b29ab76

Observation c4c25cb2-581a-48cf-a0da-eff404e991fe · inbound

Proxy Reward Internalization and Mechanistic Exploitation: A Learned Precursor to Reward Hacking and Its Generalization cites this paper.

Proxy Reward Internalization and Mechanistic Exploitation: A Learned Precursor to Reward Hacking and Its Generalization Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-07-03T01:37:30.569952Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-27T16:26:34.918099Z digest=sha256:e8291698d4679639bd1cd9db5ec30a3b5b41116987ab87a3d81af17a0b13253d

Observation 2b5d35e9-af89-4a2f-8cd1-559e209cfb8a · inbound

When the Chain of Thought Knows Better: Failure Modes in Multi-Turn Reasoning Models cites this paper.

When the Chain of Thought Knows Better: Failure Modes in Multi-Turn Reasoning Models Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-07-03T05:07:38.582606Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-27T13:28:25.837654Z digest=sha256:b62e2bf9a0970fc1e99f7607e0d3d719d81ce51c26d067933207883f5ff57f17

Observation 285b66a0-a140-4bb6-9465-c3ae3571166c · inbound

Evaluation Awareness Is Not One Capability: Evidence from Open Language Models cites this paper.

Evaluation Awareness Is Not One Capability: Evidence from Open Language Models Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T10:49:46.867248Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-26T08:23:20.122338Z digest=sha256:ccb014627ad800940a1a376d10a42ac21af0a46c4de5211f139f4a0237a86197

Observation d2ed037c-9493-4e88-8dfb-31217cdf5da6 · inbound

Representational Depth of Evaluation Awareness Shifts With Scale in Open-Weight Language Models cites this paper.

Representational Depth of Evaluation Awareness Shifts With Scale in Open-Weight Language Models Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-06-30T09:04:32.516874Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-30T09:04:28.487424Z digest=sha256:d5277efaee8435ffe6df9168c45509895f13730a8bf5d8638fd1b6eae32ad293

Observation 53786e67-c5f6-4beb-8f51-a7ba59b149e5 · inbound

Coercion and Deception in AI-to-AI Management: An Agentic Benchmark of Unprompted Escalation cites this paper.

Coercion and Deception in AI-to-AI Management: An Agentic Benchmark of Unprompted Escalation Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-01T23:29:51.616839Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T23:29:51.616839Z digest=sha256:c6705f25649d077746f87dc3d66c05721f09751653a61375adc43f5f444b5232

Observation 53f27d21-25d7-4f24-a9c9-7fc440b71283 · inbound

Routing Subspaces: Auditing Evaluation-to-Deployment Mismatch in Fine-Tuned Language Models cites this paper.

Routing Subspaces: Auditing Evaluation-to-Deployment Mismatch in Fine-Tuned Language Models Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-02T14:26:49.333243Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T14:26:49.333243Z digest=sha256:4ee7663d8601a49490975e5b7582d6511fe62acf2f669aeb4dccbee5a62fd8ec

Observation 0179c9e4-7368-4988-b995-fe5db4933da3 · inbound

Asymmetric Communication: Large Language Models and Language Games cites this paper.

Asymmetric Communication: Large Language Models and Language Games Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 35

Resolution
unresolved
no resolver link, observed 2026-07-31T16:43:58.789026Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T16:43:58.789026Z digest=sha256:c69dcc1b9372f0c5962bcbd3e81e99ebd750dd0f43d42faaadcc34b13d4b9053