Pith. sign in

Paper Citation Record · LEDGER

Subversion Strategy Eval: Can language models statelessly strategize to subvert control protocols?

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 7 inbound Pith citation observations for arXiv:2412.12480.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.12480 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 7 of 7 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 7 of 7 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T23:26:57.294628Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T08:16:47.896473Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 3fdd020e-e755-40fa-90bc-d03e4a508367 · inbound

Out of Control -- Why Alignment Needs Formal Control Theory (and an Alignment Control Stack) cites this paper.

Out of Control -- Why Alignment Needs Formal Control Theory (and an Alignment Control Stack) Subversion Strategy Eval: Can language models statelessly strategize to subvert control protocols?

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T23:26:57.294628Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:26:57.294628Z digest=sha256:59adcc20684eeaae71fd89a2bc8e618b1442a6ac694b24d9899b6afb5cdf4c2b

Observation 7b890ddb-8501-4c5c-8ebe-256fbb103f47 · inbound

Subversion via Focal Points: Investigating Collusion in LLM Monitoring cites this paper.

Subversion via Focal Points: Investigating Collusion in LLM Monitoring Subversion Strategy Eval: Can language models statelessly strategize to subvert control protocols?

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T20:52:29.606862Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:52:29.606862Z digest=sha256:0e7a6cf57c2c0430a7a7936be117b587daf96d565df9ee92e81e9979a0e8d2cc

Observation 82d8d8d1-f84e-4a27-b3c1-ece7c72ba992 · inbound

Manipulation Attacks by Misaligned AI: Risk Analysis and Safety Case Framework cites this paper.

Manipulation Attacks by Misaligned AI: Risk Analysis and Safety Case Framework Subversion Strategy Eval: Can language models statelessly strategize to subvert control protocols?

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T16:39:36.071018Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:39:36.071018Z digest=sha256:4344527660a60c91a65e1ae0186a3a9cc9a02b9837d062ab052069902e26ab31

Observation 770be177-6000-4f5a-8c36-0d232562c226 · inbound

LinuxArena: A Control Setting for AI Agents in Live Production Software Environments cites this paper.

LinuxArena: A Control Setting for AI Agents in Live Production Software Environments Subversion Strategy Eval: Can language models statelessly strategize to subvert control protocols?

Reference 9

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T11:50:20.898113Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T11:40:58.120504Z digest=sha256:1274e362e48ad7462a45885b03cfe3521913a2c8cfb437f61650ee24fde7811f

Observation f4386413-f186-4cdc-b866-736d3c8a7961 · inbound

Ensemble Monitoring for AI Control: Diverse Signals Outweigh More Compute cites this paper.

Ensemble Monitoring for AI Control: Diverse Signals Outweigh More Compute Subversion Strategy Eval: Can language models statelessly strategize to subvert control protocols?

Reference 55

Resolution
verified exact
arxiv_id, observed 2026-05-20T20:13:43.422934Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-20T20:12:03.715605Z digest=sha256:0a070ed1106073c5b213d64d5fdb8bef0ae37515dc3607aeb249b9d7e5a11b68

Observation 39b1371d-14fa-4eb6-ae4b-cf9eef710a14 · inbound

Attack Selection in Agentic AI Control Evaluations Meaningfully Decreases Safety cites this paper.

Attack Selection in Agentic AI Control Evaluations Meaningfully Decreases Safety Subversion Strategy Eval: Can language models statelessly strategize to subvert control protocols?

Reference 12

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T08:16:47.897907Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-28T06:15:32.762877Z digest=sha256:92234cd042527a19e135cfa58000c127db4955ce733d354b49b8e77f6c5023d3

Observation 95a36e71-de68-452d-93ad-c64a3c7e6cc7 · inbound

GDM AI Control Roadmap cites this paper.

GDM AI Control Roadmap Subversion Strategy Eval: Can language models statelessly strategize to subvert control protocols?

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-02T06:53:53.134779Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T06:53:53.134779Z digest=sha256:c7a193cfc10bdb24a031a6c479b4dc3f9b8c8601ecc6c9369c42bd2171c34585