Pith. sign in

Paper Citation Record · LEDGER

Stress-Testing Capability Elicitation With Password-Locked Models

As of 12 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 9 inbound Pith citation observations for arXiv:2405.19550.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2405.19550 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 9 of 9 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00

measured 9 of 9 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-09T14:47:15.106879Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T10:48:02.313626Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 5d8cadfb-b360-4778-8a7e-28404a85ddb1 · inbound

Frontier Models are Capable of In-context Scheming cites this paper.

Frontier Models are Capable of In-context Scheming Stress-Testing Capability Elicitation With Password-Locked Models

Reference 15

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T14:22:01.639206Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-16T14:22:01.616448Z digest=sha256:73267219b8cf3632ca04a2a68b858f82ac27317d18a84e82b1f8c4f6f18990ab

Observation ef386264-db9e-4fb4-bfb3-572786246e45 · inbound

Model Tampering Attacks Enable More Rigorous Evaluations of LLM Capabilities cites this paper.

Model Tampering Attacks Enable More Rigorous Evaluations of LLM Capabilities Stress-Testing Capability Elicitation With Password-Locked Models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-09T14:47:15.106879Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T14:47:15.106879Z digest=sha256:cf2b79cd8a09f33245ab40a6c11570d582e833c6bee0ab040ebea645936963c0

Observation 81792aed-f23f-4c4b-b15f-3b7ae48bef03 · inbound

Adversarial Attacks on Robotic Vision Language Action Models cites this paper.

Adversarial Attacks on Robotic Vision Language Action Models Stress-Testing Capability Elicitation With Password-Locked Models

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-07T11:11:40.474093Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:11:40.474093Z digest=sha256:904f77b0e97cc0da449e852975937a64b0e396e375ea958ef7d70a0df9f5b7ad

Observation e4d114e6-ce53-4126-a1ab-e3887415032e · inbound

Ensemble Monitoring for AI Control: Diverse Signals Outweigh More Compute cites this paper.

Ensemble Monitoring for AI Control: Diverse Signals Outweigh More Compute Stress-Testing Capability Elicitation With Password-Locked Models

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-05-20T20:13:43.427596Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-20T20:12:03.715605Z digest=sha256:e22260d13da9c7bb8bb9c5698eb4e7e3207c370ae66325e6a165c5401ef942ae

Observation 5247466e-5473-45ad-8301-eddc3cce436b · inbound

Evaluation Cards: An Interpretive Layer for AI Evaluation Reporting cites this paper.

Evaluation Cards: An Interpretive Layer for AI Evaluation Reporting Stress-Testing Capability Elicitation With Password-Locked Models

Reference 46

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T02:07:33.261481Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-06-27T16:11:36.483820Z digest=sha256:29d45a76f0327761c3e37cce64099422c975bc2015e4ddfaaf1f576e02c82769

Observation 730a255e-3e40-4597-a585-32246e01e0fa · inbound

MIRAGE: A Polarity-Flipping Encoding Subspace in LLM Agents cites this paper.

MIRAGE: A Polarity-Flipping Encoding Subspace in LLM Agents Stress-Testing Capability Elicitation With Password-Locked Models

Reference 30

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T04:57:38.144361Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-06-27T13:33:24.087333Z digest=sha256:ebe7feedb5637cbc99bcc0ad9c31cdf41e9e388dae5a9d34ddd79d0c35b59305

Observation 74230aa1-49cf-4dc3-a9e6-40dda6aa6a55 · inbound

Generalization Hacking: Models Can Game Reinforcement Learning by Preventing Behavioral Generalization cites this paper.

Generalization Hacking: Models Can Game Reinforcement Learning by Preventing Behavioral Generalization Stress-Testing Capability Elicitation With Password-Locked Models

Reference 27

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T08:17:44.951229Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-06-27T10:59:09.372559Z digest=sha256:be2656b7241035f42ac5660f4723b2ddfba0e3d5d3f687c3c96d923661e1af5b

Observation 0ee6cac9-8b12-4d12-b68a-c3d048d1f5b7 · inbound

"Did you lie?" Evaluating Lie Detectors across Model Scale and Belief-Verified Model Organisms cites this paper.

"Did you lie?" Evaluating Lie Detectors across Model Scale and Belief-Verified Model Organisms Stress-Testing Capability Elicitation With Password-Locked Models

Reference 73

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T10:48:02.315041Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-06-27T09:51:16.969884Z digest=sha256:688f5c9a62c0f86e740169ea80095c1a3057155948c7bff21a871734596b995d

Observation 2bd3a690-3ff5-4390-98c4-ce098f4f5997 · inbound

Routing Subspaces: Auditing Evaluation-to-Deployment Mismatch in Fine-Tuned Language Models cites this paper.

Routing Subspaces: Auditing Evaluation-to-Deployment Mismatch in Fine-Tuned Language Models Stress-Testing Capability Elicitation With Password-Locked Models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-02T14:26:47.323505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T14:26:47.323505Z digest=sha256:58a78f1bac49aeb11059f0034cffee8e5488cd4c7d962b7fdbcbc1c6a22845a9