Pith. sign in

Paper Citation Record · LEDGER

Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 19 inbound Pith citation observations for arXiv:2407.04694.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2407.04694 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 19 of 19 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 19 of 19 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T05:42:43.460514Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

2
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 89b2d313-17d3-4a53-8ad1-5346b4b3a174 · inbound

Inspection and Control of Self-Generated-Text Recognition Ability in Llama3-8b-Instruct cites this paper.

Inspection and Control of Self-Generated-Text Recognition Ability in Llama3-8b-Instruct Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-23T19:53:23.281474Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-23T19:51:14.630756Z digest=sha256:80461634878802aa5f351f867e5a8f0975aa7586abed38f5317dfa5090f461a4

Observation 652de3e4-a831-4732-9b28-0e8e1c601c8f · inbound

Frontier Models are Capable of In-context Scheming cites this paper.

Frontier Models are Capable of In-context Scheming Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-16T14:22:01.659365Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-16T14:22:01.616448Z digest=sha256:f289c263fe89e33e89ca8e98a953a94a374adb075e36c148d95255ded45e48ce

Observation 8274109e-2066-46e9-9ecb-19047f44c99c · inbound

Compromising Honesty and Harmlessness in Language Models via Deception Attacks cites this paper.

Compromising Honesty and Harmlessness in Language Models via Deception Attacks Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-08T05:42:43.460514Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T05:42:43.460514Z digest=sha256:364527b478aa78fe4ab9e2992afaa66ef0f2a8a73af1635e4b8e3c53868f1326

Observation 23e32c6e-8ca0-4f2e-8a64-9c180c04444a · inbound

Does It Make Sense to Speak of Introspection in Large Language Models? cites this paper.

Does It Make Sense to Speak of Introspection in Large Language Models? Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T10:30:37.638319Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:30:37.638319Z digest=sha256:cf6385ea7e48159b3fa4b48680b50c51a962c71644624cac4f2f7df71c72f3fa

Observation a73ccfc0-6f4b-4647-a8a0-703489c7f742 · inbound

Model Organisms for Emergent Misalignment cites this paper.

Model Organisms for Emergent Misalignment Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T04:07:28.181472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:07:28.181472Z digest=sha256:6eddd1be66c3eca74a2d6d9b3eda7248b8b71191a4e2ce971b0f0f254c10e73e

Observation 2ee47776-1d8d-47e2-b4be-e7fa232e1549 · inbound

Convergent Linear Representations of Emergent Misalignment cites this paper.

Convergent Linear Representations of Emergent Misalignment Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T04:09:26.150070Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:09:26.150070Z digest=sha256:9bb8ab19f9ae33c437fa0dcd7f416e7c7ce8d82e0f44147d4dd0f79639bf319c

Observation 48250677-93fe-4649-9ae7-99e51a99c4c1 · inbound

Safety Features for a Centralised AGI Project cites this paper.

Safety Features for a Centralised AGI Project Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T00:23:56.096143Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:23:56.096143Z digest=sha256:cda6e40e69343831b3b94df7452d25662017ff1a8b4c115a1be35f79efbc60a5

Observation b9940761-e6b3-4bbb-9d9b-261c20e169ab · inbound

Honeypot Protocol cites this paper.

Honeypot Protocol Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-10T14:35:34.140641Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T14:35:16.230357Z digest=sha256:855574b8f39475386a84528770c639b621a85ff22910c297d45df2ac62e6779c

Observation cdf4da5b-7e9c-49d6-88af-84404529e2a1 · inbound

Measuring Evaluation-Context Divergence in Open-Weight LLMs: A Paired-Prompt Protocol with Pilot Evidence of Alignment-Pipeline-Specific Heterogeneity cites this paper.

Measuring Evaluation-Context Divergence in Open-Weight LLMs: A Paired-Prompt Protocol with Pilot Evidence of Alignment-Pipeline-Specific Heterogeneity Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-08T22:04:18.014075Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-08T10:23:02.697982Z digest=sha256:0cc1465d6f3f22132aae1e0ae782d198014b10729838da097eda2bafec8d03d9

Observation 2f90b689-13ce-4f6a-9573-0a5439e1beac · inbound

Do as I Say, Not as I Do: Instruction-Induction Conflict in LLMs cites this paper.

Do as I Say, Not as I Do: Instruction-Induction Conflict in LLMs Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 8

Resolution
metadata mismatch
arxiv_id, observed 2026-05-21T07:34:02.866648Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T07:30:27.297971Z digest=sha256:94b374e26d9dd06a6e57a59e546b3dc20fed0b846412f1529d4b8689dd8fe39e

Observation a2620d65-f901-4f6c-abc7-595152701dc7 · inbound

Do as I Say, Not as I Do: Instruction-Induction Conflict in LLMs cites this paper.

Do as I Say, Not as I Do: Instruction-Induction Conflict in LLMs Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 8

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T18:04:58.126158Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T18:00:29.237972Z digest=sha256:952d824a9ead70e998218e541c09e7dce58c416aaad461d43e4676cb3fd17e5e

Observation b2bb0f61-eb8d-4fef-adc2-ec8a9f3ca6c3 · inbound

AI Integrity: Defending Against Backdoors and Secret Loyalties cites this paper.

AI Integrity: Defending Against Backdoors and Secret Loyalties Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 21

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T14:59:54.636218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-07-04T14:56:53.806480Z digest=sha256:7b37b95fc69899a1791c96d10f0c65388ad6e4c9296aed32858798463885c7b6

Observation c4c25cb2-581a-48cf-a0da-eff404e991fe · inbound

Proxy Reward Internalization and Mechanistic Exploitation: A Learned Precursor to Reward Hacking and Its Generalization cites this paper.

Proxy Reward Internalization and Mechanistic Exploitation: A Learned Precursor to Reward Hacking and Its Generalization Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-07-03T01:37:30.569952Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-27T16:26:34.918099Z digest=sha256:54d9af2289c7f3ec43b238df56cd901110abcb9b21e80dca562a0c387bc62255

Observation 2b5d35e9-af89-4a2f-8cd1-559e209cfb8a · inbound

When the Chain of Thought Knows Better: Failure Modes in Multi-Turn Reasoning Models cites this paper.

When the Chain of Thought Knows Better: Failure Modes in Multi-Turn Reasoning Models Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-07-03T05:07:38.582606Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-27T13:28:25.837654Z digest=sha256:5c4d469b33a2232dcbf75a0c29c7a25e6052a139947d2a45d88688a9aaacda91

Observation 285b66a0-a140-4bb6-9465-c3ae3571166c · inbound

Evaluation Awareness Is Not One Capability: Evidence from Open Language Models cites this paper.

Evaluation Awareness Is Not One Capability: Evidence from Open Language Models Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T10:49:46.867248Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-26T08:23:20.122338Z digest=sha256:f51ef15dfac8975c24fd1a9788d57eb63ba350942c3e22c90b4190ae9cc8bca2

Observation d2ed037c-9493-4e88-8dfb-31217cdf5da6 · inbound

Representational Depth of Evaluation Awareness Shifts With Scale in Open-Weight Language Models cites this paper.

Representational Depth of Evaluation Awareness Shifts With Scale in Open-Weight Language Models Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-06-30T09:04:32.516874Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T09:04:28.487424Z digest=sha256:2c7df7cf26ba6bdc5afd04e6aa3107753586ae8f4d62f73c0ad0ad5d95e42d58

Observation 53786e67-c5f6-4beb-8f51-a7ba59b149e5 · inbound

Coercion and Deception in AI-to-AI Management: An Agentic Benchmark of Unprompted Escalation cites this paper.

Coercion and Deception in AI-to-AI Management: An Agentic Benchmark of Unprompted Escalation Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-01T23:29:51.616839Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T23:29:51.616839Z digest=sha256:c6705f25649d077746f87dc3d66c05721f09751653a61375adc43f5f444b5232

Observation 53f27d21-25d7-4f24-a9c9-7fc440b71283 · inbound

Routing Subspaces: Auditing Evaluation-to-Deployment Mismatch in Fine-Tuned Language Models cites this paper.

Routing Subspaces: Auditing Evaluation-to-Deployment Mismatch in Fine-Tuned Language Models Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-02T14:26:49.333243Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T14:26:49.333243Z digest=sha256:4ee7663d8601a49490975e5b7582d6511fe62acf2f669aeb4dccbee5a62fd8ec

Observation 0179c9e4-7368-4988-b995-fe5db4933da3 · inbound

Asymmetric Communication: Large Language Models and Language Games cites this paper.

Asymmetric Communication: Large Language Models and Language Games Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 35

Resolution
unresolved
no resolver link, observed 2026-07-31T16:43:58.789026Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T16:43:58.789026Z digest=sha256:c69dcc1b9372f0c5962bcbd3e81e99ebd750dd0f43d42faaadcc34b13d4b9053