Pith. sign in

Paper Citation Record · LEDGER

Predicting LLM Safety Before Release by Simulating Deployment

As of 15 August 2026, this Paper Citation Record lists 22 of 22 outbound references and 0 inbound Pith citation observations for arXiv:2607.07184.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.07184 v1

Coverage vector

measured 22 of 22 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-07-09T18:07:42.556329Z

measured 22 of 22 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

22 of 22 outbound references displayed

  • verified exact9
  • verified fuzzy10
  • unresolved1
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a59c605b-f5b8-4386-9a4f-7586d3f74aab · outbound

This paper cites Large Language Models Often Know When They Are Being Evaluated.

Predicting LLM Safety Before Release by Simulating Deployment Large Language Models Often Know When They Are Being Evaluated

Reference 1

Resolution
metadata mismatch
local_arxiv, observed 2026-07-09T18:16:25.813826Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-07-09T18:07:42.556329Z digest=sha256:3d692f7658310b217639c832c0fa1e79ef258199b4c93668f8eb273151e498b6

Observation b5d9dc23-76d9-4dc4-9f7b-ae44ec6d38c9 · outbound

This paper cites Charles, D.

Predicting LLM Safety Before Release by Simulating Deployment Charles, D

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T18:16:26.196007Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-07-09T18:07:42.556329Z digest=sha256:0ecf3ea57670a5b3c83d311d55db2d146b19e55f55ff2b8727e6e9b415b49d60

Observation b4936888-4d53-4cde-ba24-de439737870f · outbound

This paper cites Stress Testing Deliberative Alignment for Anti-Scheming Training , url =.

Predicting LLM Safety Before Release by Simulating Deployment Stress Testing Deliberative Alignment for Anti-Scheming Training , url =

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-07-09T18:16:25.709877Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-07-09T18:07:42.556329Z digest=sha256:16e1e219551108c2e4b04247baa16e63afe253a6d226dfe8785c23649053594b

Observation 77701459-7074-482d-85cb-f00d98eb6a0b · outbound

This paper cites Metagaming matters for training, evaluation, and oversight.

Predicting LLM Safety Before Release by Simulating Deployment Metagaming matters for training, evaluation, and oversight

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T18:16:26.199722Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-07-09T18:07:42.556329Z digest=sha256:33ab355e64dd569e57e27d084673f0ca8bc3ffb783d86b9a13f22528b3a6d7d4

Observation b00ce237-5f4a-4f94-9342-7f0245368c96 · outbound

This paper cites Training llms for honesty via confessions.

Predicting LLM Safety Before Release by Simulating Deployment Training llms for honesty via confessions

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-07-09T18:16:25.811289Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-07-09T18:07:42.556329Z digest=sha256:868cb981a590de48a7c75902a1d724e4bff1eeebaaa854caac0ae2ed9de30805

Observation 5cb2503b-e511-41eb-bee7-4e5c1b1c7b4b · outbound

This paper cites Guan, Miles Wang, Micah Carroll, Zehao Dou, Annie Y.

Predicting LLM Safety Before Release by Simulating Deployment Guan, Miles Wang, Micah Carroll, Zehao Dou, Annie Y

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-07-09T18:16:25.818454Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-07-09T18:07:42.556329Z digest=sha256:d3798ccda0a55c569b5ee3cd8d747bda29adfd54bf9066562cde9713321df0d4

Observation 388ac83d-9218-4a29-af7e-2e47d17f13ed · outbound

This paper cites Jimenez, John Yang, Alexander Wettig, Shunyu Yao, Kexin Pei, Ofir Press, and Karthik R.

Predicting LLM Safety Before Release by Simulating Deployment Jimenez, John Yang, Alexander Wettig, Shunyu Yao, Kexin Pei, Ofir Press, and Karthik R

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T18:16:26.190201Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-07-09T18:07:42.556329Z digest=sha256:333241335158b0cc3d4dde7320846683fb7e02fc2202a07bb7389e400bee5bb5

Observation 468a3f8d-d66b-462a-8bc0-3ca5086d4e30 · outbound

This paper cites American Invitational Mathematics Examination (AIME), 2026.

Predicting LLM Safety Before Release by Simulating Deployment American Invitational Mathematics Examination (AIME), 2026

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T18:16:26.192460Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-07-09T18:07:42.556329Z digest=sha256:54042ac651aa6fffe5a816a725709ded90b632ea848c95bf0a1ced78409bac2a

Observation 3b894894-9bb9-4f4e-bc3e-b6248a9ee775 · outbound

This paper cites OpenAI GPT-5 System Card.https://openai.com/index/gpt-5-system-card/,.

Predicting LLM Safety Before Release by Simulating Deployment OpenAI GPT-5 System Card.https://openai.com/index/gpt-5-system-card/,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T18:16:26.187623Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-07-09T18:07:42.556329Z digest=sha256:bfde53d058191486934ff492a3d60ddb26eb93d12bf96af0f0b395da27d9b8d7

Observation c0d4fad0-08a8-45a8-877d-833aad3ee043 · outbound

This paper cites Maddison, and Tatsunori Hashimoto.

Predicting LLM Safety Before Release by Simulating Deployment Maddison, and Tatsunori Hashimoto

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T18:16:26.206196Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-07-09T18:07:42.556329Z digest=sha256:2aa0174c068e711b1433b082fb79eb882f1d18c8b2fc1245126fdaf3e7034b0f

Observation d3d25615-2e5e-4e54-9043-7fbfeae3d608 · outbound

This paper cites WildChat: 1M ChatGPT Interaction Logs in the Wild.

Predicting LLM Safety Before Release by Simulating Deployment WildChat: 1M ChatGPT Interaction Logs in the Wild

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-07-09T18:16:25.809290Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-07-09T18:07:42.556329Z digest=sha256:a091667c4ec59dce03c1e8a1bb485a20ffbadf0ca542a6845af0b0df8d907dd4

Observation 3a8df77f-a393-4a11-a01a-5f512aa34551 · outbound

This paper cites Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety.

Predicting LLM Safety Before Release by Simulating Deployment Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety

Reference 13

Resolution
verified exact
local_arxiv, observed 2026-07-09T18:16:25.820530Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-07-09T18:07:42.556329Z digest=sha256:fdbb5325b07bb5302d40637fc03100e7ff80c5087e098f57184221c4215db3db

Observation fc8328c5-7d60-43cb-b6a9-a6443bf452ce · outbound

This paper cites Estimating the probabilities of rare outputs in language models,.

Predicting LLM Safety Before Release by Simulating Deployment Estimating the probabilities of rare outputs in language models,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T18:16:26.174954Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-07-09T18:07:42.556329Z digest=sha256:36180565ef090d1a9392ead62ba275487bb59fb4288f194c09a0c2655fda36c5

Observation 6f51c6d5-a8db-463e-808c-42f4bc33c9a7 · outbound

This paper cites Estimating the Probabilities of Rare Outputs in Language Models.

Predicting LLM Safety Before Release by Simulating Deployment Estimating the Probabilities of Rare Outputs in Language Models

Reference 15

Resolution
verified exact
local_arxiv, observed 2026-07-09T18:16:25.821334Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-07-09T18:07:42.556329Z digest=sha256:cd107b896099cc0b858e4e3b90578669653e3b314988509b1cbad6406f0a3090

Observation e3f0a6b3-9976-4bf5-ba13-734202235ffd · outbound

This paper cites an unresolved cited work.

Predicting LLM Safety Before Release by Simulating Deployment Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-07-09T18:16:26.178134Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-07-09T18:07:42.556329Z digest=sha256:5dd2862851a4cf747555f8b6304f93c5c6307a4bc1e2a0c12f37fb96c5932372

Observation 2f698325-05ea-4d29-a0f7-33574204f881 · outbound

This paper cites Forecasting Rare Language Model Behaviors.

Predicting LLM Safety Before Release by Simulating Deployment Forecasting Rare Language Model Behaviors

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-07-09T18:16:25.803524Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-07-09T18:07:42.556329Z digest=sha256:d08965a4eef9a8cecfaa084d383613dc7b022ce3664e0d052eeb31dc6523f910

Observation 80d42c63-ef18-4105-9c25-d389bf74ade5 · outbound

This paper cites Estimating Tail Risks in Language Model Output Distributions.

Predicting LLM Safety Before Release by Simulating Deployment Estimating Tail Risks in Language Model Output Distributions

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-07-09T18:16:25.817785Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-07-09T18:07:42.556329Z digest=sha256:df84eaa056add3a5cabd83cb7001570dfe8c797ea9401c8c98c78bd61f490e23

Observation 978f1a04-1541-4227-b309-a66997023dce · outbound

This paper cites Evaluating predictions of model behaviour.https://www.governance.ai/anal ysis/evaluating-predictions-of-model-behaviour, April 2024.

Predicting LLM Safety Before Release by Simulating Deployment Evaluating predictions of model behaviour.https://www.governance.ai/anal ysis/evaluating-predictions-of-model-behaviour, April 2024

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T18:16:26.180627Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-07-09T18:07:42.556329Z digest=sha256:23f2a4a4a7aa2474c8c89a61de75c6cc286797c80824de345da7114cc78fd669

Observation cffe79e8-3d2c-4d35-8e8d-cbc23825a94e · outbound

This paper cites Foster and Ariel Deardorff.

Predicting LLM Safety Before Release by Simulating Deployment Foster and Ariel Deardorff

Reference 20

Resolution
verified exact
doi, observed 2026-07-09T18:16:25.704383Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-07-09T18:07:42.556329Z digest=sha256:b5e89660558a653940635b1acd0c1fbe738208b3214daa15c5ad7ef09903e1f8

Observation 7aa1ed9b-7d15-4854-85ee-ca79172e2ff3 · outbound

This paper cites Estimating model behavior before deployment with representative prompts for gpt-5.4 thinking, Mar 2026.

Predicting LLM Safety Before Release by Simulating Deployment Estimating model behavior before deployment with representative prompts for gpt-5.4 thinking, Mar 2026

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T18:16:26.184558Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-07-09T18:07:42.556329Z digest=sha256:640e96c3e9815a33b56b1075b5f5f17703462290db8a082f09b1e659053ceeba

Observation 1bc9b69d-5572-4288-b948-ef2d967dffcf · outbound

This paper cites OpenAI GPT-5.4 Thinking System Card.

Predicting LLM Safety Before Release by Simulating Deployment OpenAI GPT-5.4 Thinking System Card

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T18:16:26.210612Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-07-09T18:07:42.556329Z digest=sha256:7de6d02c06c32b99c87fdd5be253c5e1d6f39e7b0df643f2de2b6fdfadf92d7b

Observation 99448bf6-a9d8-4ef7-9500-cf82ba76e7af · outbound

This paper cites DS lower NLL.

Predicting LLM Safety Before Release by Simulating Deployment DS lower NLL

Reference 23

Resolution
malformed identifier
raw_fallback, observed 2026-07-09T18:16:26.203164Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-07-09T18:07:42.556329Z digest=sha256:4231eee7c8083ce7edb31f52fdd57d09ef75bfa76a4aa7c4cc5e3f81c5a7819e

Pith citing papers

No inbound Pith citation observations are available.