Pith. sign in

Paper Citation Record · LEDGER

Deception in LLMs: Self-Preservation and Autonomous Goals in Large Language Models

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 10 inbound Pith citation observations for arXiv:2501.16513.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.16513 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 10 of 10 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 10 of 10 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:27:03.390674Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 466f5ac9-9c5a-48fe-80b1-f32da3f1006f · inbound

Security Concerns for Large Language Models: A Survey cites this paper.

Security Concerns for Large Language Models: A Survey Deception in LLMs: Self-Preservation and Autonomous Goals in Large Language Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T14:27:03.390674Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:27:03.390674Z digest=sha256:9629b19f09e359f58321d3fdd786efb46c92f5ff77ec241622af8e9752237c16

Observation a6d46fdf-4ef0-4def-9356-fd2bbe53bc74 · inbound

Kaleidoscopic Teaming in Multi Agent Simulations cites this paper.

Kaleidoscopic Teaming in Multi Agent Simulations Deception in LLMs: Self-Preservation and Autonomous Goals in Large Language Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T23:35:47.296378Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:35:47.296378Z digest=sha256:cb4985db9099d50570e73faece89db1cf6925774e5cdedf4d2570b196710d941

Observation a909a741-651d-4fb4-8881-87be77fe645a · inbound

The Impact of Off-Policy Training Data on Probe Generalisation cites this paper.

The Impact of Off-Policy Training Data on Probe Generalisation Deception in LLMs: Self-Preservation and Autonomous Goals in Large Language Models

Reference 4

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T20:30:11.686788Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-17T20:26:37.914522Z digest=sha256:1537c40f2f1074456e789288a964fc85a27406b413a66867469a166640c291dd

Observation d81916fb-4a16-44eb-b2af-f1ec69d5e0bb · inbound

The Possibility of Artificial Intelligence Becoming a Subject and the Alignment Problem cites this paper.

The Possibility of Artificial Intelligence Becoming a Subject and the Alignment Problem Deception in LLMs: Self-Preservation and Autonomous Goals in Large Language Models

Reference 4

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T10:44:37.792152Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T10:41:14.970156Z digest=sha256:20b4fae2945ec4fc652d97e3317420a5e80f057f6610b32f0fefb0a9ac58a539

Observation deb53dd0-7256-4122-85ac-367c627218cc · inbound

The Possibility of Artificial Intelligence Becoming a Subject and the Alignment Problem cites this paper.

The Possibility of Artificial Intelligence Becoming a Subject and the Alignment Problem Deception in LLMs: Self-Preservation and Autonomous Goals in Large Language Models

Reference 2014

Resolution
unresolved
no resolver link, observed 2026-08-02T16:13:12.025695Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T16:13:12.025695Z digest=sha256:0b9a35da30cf62a22428445b590a0173cf667e4523c3527d8b914c9a4c6867c5

Observation cb5e3c16-b45a-41eb-9c2d-51ecdd545d7b · inbound

Do Linear Probes Generalize Better in Persona Coordinates? cites this paper.

Do Linear Probes Generalize Better in Persona Coordinates? Deception in LLMs: Self-Preservation and Autonomous Goals in Large Language Models

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-05-12T04:51:23.120483Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-12T04:47:47.214726Z digest=sha256:6d4ed21243e9e2eff5e4ca69d9eb61b685e7c64432580f2c2ff4ef64de38c9a6

Observation 81f9221d-d2de-4835-8ec7-c998ff2375cb · inbound

Do Linear Probes Generalize Better in Persona Coordinates? cites this paper.

Do Linear Probes Generalize Better in Persona Coordinates? Deception in LLMs: Self-Preservation and Autonomous Goals in Large Language Models

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-05-19T17:07:40.540616Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-19T17:06:22.481341Z digest=sha256:681848ee9fc3809de3531aeaa12eb4cb06e4c0105b9f478955904a9d259c3982

Observation 8f79cedc-96ae-4ea2-9f35-309c64e4bfc3 · inbound

The Hidden Bias of Process Reward Models:PRISM for Rewarding the Right Reasoning cites this paper.

The Hidden Bias of Process Reward Models:PRISM for Rewarding the Right Reasoning Deception in LLMs: Self-Preservation and Autonomous Goals in Large Language Models

Reference 10

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T00:37:30.100882Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-27T17:07:06.227106Z digest=sha256:8fdeb2997ba41f67f281fad0c9a911e0189f3ee6f6a8b6c36a3dd40b6c0c32b5

Observation 3502dda8-1e4e-4a71-9a42-cbe73d02f0fd · inbound

Theory of Mind and Persuasion Beyond Conversation: Assessing the Capacity of LLMs to Induce Belief States via Planning and Action cites this paper.

Theory of Mind and Persuasion Beyond Conversation: Assessing the Capacity of LLMs to Induce Belief States via Planning and Action Deception in LLMs: Self-Preservation and Autonomous Goals in Large Language Models

Reference 87

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T10:15:44.681180Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-07-01T05:40:54.002702Z digest=sha256:6d04c483ff9faba166ea1bfd92cfc0026bdbf4e2e591fabb35eb76550240c2cf

Observation 2c4b6ae3-14a9-4b38-87df-147106ec41a2 · inbound

Transcoders for Investigating Deception in Language Models cites this paper.

Transcoders for Investigating Deception in Language Models Deception in LLMs: Self-Preservation and Autonomous Goals in Large Language Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-02T01:06:13.817475Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T01:06:13.817475Z digest=sha256:8a2152a03a144b0b9d7297718b90537e747305093608a954b55537a402b1d5a1