Pith. sign in

Paper Citation Record · LEDGER

Stress-Testing Capability Elicitation With Password-Locked Models

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 8 inbound Pith citation observations for arXiv:2405.19550.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2405.19550 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 8 of 8 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 8 of 8 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:11:40.474093Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T10:48:02.313626Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 5d8cadfb-b360-4778-8a7e-28404a85ddb1 · inbound

Frontier Models are Capable of In-context Scheming cites this paper.

Frontier Models are Capable of In-context Scheming Stress-Testing Capability Elicitation With Password-Locked Models

Reference 15

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T14:22:01.639206Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-16T14:22:01.616448Z digest=sha256:d70240d0de6740fd442921d84a48b2efb9acf36f816e9fd097e8910c5f2fb983

Observation 81792aed-f23f-4c4b-b15f-3b7ae48bef03 · inbound

Adversarial Attacks on Robotic Vision Language Action Models cites this paper.

Adversarial Attacks on Robotic Vision Language Action Models Stress-Testing Capability Elicitation With Password-Locked Models

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-07T11:11:40.474093Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:11:40.474093Z digest=sha256:7e3918985daa9ee3fb70005b9304b68d844580838f1e4242d95e7d13f7e15f38

Observation e4d114e6-ce53-4126-a1ab-e3887415032e · inbound

Ensemble Monitoring for AI Control: Diverse Signals Outweigh More Compute cites this paper.

Ensemble Monitoring for AI Control: Diverse Signals Outweigh More Compute Stress-Testing Capability Elicitation With Password-Locked Models

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-05-20T20:13:43.427596Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-20T20:12:03.715605Z digest=sha256:20ecd1fcc25a0c34571f618e9eaecea4edec346dd669dcdfbed013cbd523302b

Observation 5247466e-5473-45ad-8301-eddc3cce436b · inbound

Evaluation Cards: An Interpretive Layer for AI Evaluation Reporting cites this paper.

Evaluation Cards: An Interpretive Layer for AI Evaluation Reporting Stress-Testing Capability Elicitation With Password-Locked Models

Reference 46

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T02:07:33.261481Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-27T16:11:36.483820Z digest=sha256:23cf810a35667a2124d6e96e649aa00ab268345422c6357611c01305980fae34

Observation 730a255e-3e40-4597-a585-32246e01e0fa · inbound

MIRAGE: A Polarity-Flipping Encoding Subspace in LLM Agents cites this paper.

MIRAGE: A Polarity-Flipping Encoding Subspace in LLM Agents Stress-Testing Capability Elicitation With Password-Locked Models

Reference 30

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T04:57:38.144361Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-27T13:33:24.087333Z digest=sha256:76ba97fbcbab6dbbef4466af823d635eb1fb2d75b17e9b49e27b045debf3e917

Observation 74230aa1-49cf-4dc3-a9e6-40dda6aa6a55 · inbound

Generalization Hacking: Models Can Game Reinforcement Learning by Preventing Behavioral Generalization cites this paper.

Generalization Hacking: Models Can Game Reinforcement Learning by Preventing Behavioral Generalization Stress-Testing Capability Elicitation With Password-Locked Models

Reference 27

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T08:17:44.951229Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-27T10:59:09.372559Z digest=sha256:d7584b4f319512150625eb370cfd23d2c7c7c6f0ba7efc3f123d1d3014049ed3

Observation 0ee6cac9-8b12-4d12-b68a-c3d048d1f5b7 · inbound

"Did you lie?" Evaluating Lie Detectors across Model Scale and Belief-Verified Model Organisms cites this paper.

"Did you lie?" Evaluating Lie Detectors across Model Scale and Belief-Verified Model Organisms Stress-Testing Capability Elicitation With Password-Locked Models

Reference 73

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T10:48:02.315041Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-27T09:51:16.969884Z digest=sha256:4c146a170914e74cd14ddf6635ca2b152df45b05be6a9aca1601058cf13973bb

Observation 2bd3a690-3ff5-4390-98c4-ce098f4f5997 · inbound

Routing Subspaces: Auditing Evaluation-to-Deployment Mismatch in Fine-Tuned Language Models cites this paper.

Routing Subspaces: Auditing Evaluation-to-Deployment Mismatch in Fine-Tuned Language Models Stress-Testing Capability Elicitation With Password-Locked Models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-02T14:26:47.323505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T14:26:47.323505Z digest=sha256:2d0622edf76babc43feb94ee59c4fda4d6b9307429d7905ba8c8d0141f962acd