Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-07-09T18:07:42.556329Z
Paper Citation Record · LEDGER
As of 15 August 2026, this Paper Citation Record lists 22 of 22 outbound references and 0 inbound Pith citation observations for arXiv:2607.07184.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-07-09T18:07:42.556329Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
22 of 22 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation a59c605b-f5b8-4386-9a4f-7586d3f74aab · outbound
Predicting LLM Safety Before Release by Simulating Deployment Large Language Models Often Know When They Are Being Evaluated
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation b5d9dc23-76d9-4dc4-9f7b-ae44ec6d38c9 · outbound
Predicting LLM Safety Before Release by Simulating Deployment Charles, D
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation b4936888-4d53-4cde-ba24-de439737870f · outbound
Predicting LLM Safety Before Release by Simulating Deployment Stress Testing Deliberative Alignment for Anti-Scheming Training , url =
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 77701459-7074-482d-85cb-f00d98eb6a0b · outbound
Predicting LLM Safety Before Release by Simulating Deployment Metagaming matters for training, evaluation, and oversight
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation b00ce237-5f4a-4f94-9342-7f0245368c96 · outbound
Predicting LLM Safety Before Release by Simulating Deployment Training llms for honesty via confessions
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 5cb2503b-e511-41eb-bee7-4e5c1b1c7b4b · outbound
Predicting LLM Safety Before Release by Simulating Deployment Guan, Miles Wang, Micah Carroll, Zehao Dou, Annie Y
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 388ac83d-9218-4a29-af7e-2e47d17f13ed · outbound
Predicting LLM Safety Before Release by Simulating Deployment Jimenez, John Yang, Alexander Wettig, Shunyu Yao, Kexin Pei, Ofir Press, and Karthik R
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 468a3f8d-d66b-462a-8bc0-3ca5086d4e30 · outbound
Predicting LLM Safety Before Release by Simulating Deployment American Invitational Mathematics Examination (AIME), 2026
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 3b894894-9bb9-4f4e-bc3e-b6248a9ee775 · outbound
Predicting LLM Safety Before Release by Simulating Deployment OpenAI GPT-5 System Card.https://openai.com/index/gpt-5-system-card/,
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation c0d4fad0-08a8-45a8-877d-833aad3ee043 · outbound
Predicting LLM Safety Before Release by Simulating Deployment Maddison, and Tatsunori Hashimoto
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation d3d25615-2e5e-4e54-9043-7fbfeae3d608 · outbound
Predicting LLM Safety Before Release by Simulating Deployment WildChat: 1M ChatGPT Interaction Logs in the Wild
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 3a8df77f-a393-4a11-a01a-5f512aa34551 · outbound
Predicting LLM Safety Before Release by Simulating Deployment Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation fc8328c5-7d60-43cb-b6a9-a6443bf452ce · outbound
Predicting LLM Safety Before Release by Simulating Deployment Estimating the probabilities of rare outputs in language models,
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 6f51c6d5-a8db-463e-808c-42f4bc33c9a7 · outbound
Predicting LLM Safety Before Release by Simulating Deployment Estimating the Probabilities of Rare Outputs in Language Models
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation e3f0a6b3-9976-4bf5-ba13-734202235ffd · outbound
Predicting LLM Safety Before Release by Simulating Deployment Unresolved cited work
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 2f698325-05ea-4d29-a0f7-33574204f881 · outbound
Predicting LLM Safety Before Release by Simulating Deployment Forecasting Rare Language Model Behaviors
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 80d42c63-ef18-4105-9c25-d389bf74ade5 · outbound
Predicting LLM Safety Before Release by Simulating Deployment Estimating Tail Risks in Language Model Output Distributions
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 978f1a04-1541-4227-b309-a66997023dce · outbound
Predicting LLM Safety Before Release by Simulating Deployment Evaluating predictions of model behaviour.https://www.governance.ai/anal ysis/evaluating-predictions-of-model-behaviour, April 2024
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation cffe79e8-3d2c-4d35-8e8d-cbc23825a94e · outbound
Predicting LLM Safety Before Release by Simulating Deployment Foster and Ariel Deardorff
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 7aa1ed9b-7d15-4854-85ee-ca79172e2ff3 · outbound
Predicting LLM Safety Before Release by Simulating Deployment Estimating model behavior before deployment with representative prompts for gpt-5.4 thinking, Mar 2026
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 1bc9b69d-5572-4288-b948-ef2d967dffcf · outbound
Predicting LLM Safety Before Release by Simulating Deployment OpenAI GPT-5.4 Thinking System Card
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 99448bf6-a9d8-4ef7-9500-cf82ba76e7af · outbound
Predicting LLM Safety Before Release by Simulating Deployment DS lower NLL
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
No inbound Pith citation observations are available.