Pith. sign in

Paper Citation Record · LEDGER

An Example Safety Case for Safeguards Against Misuse

As of 8 August 2026, this Paper Citation Record lists 28 of 28 outbound references and 0 inbound Pith citation observations for arXiv:2505.18003.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.18003 v1

Coverage vector

measured 28 of 28 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:41:41.097514Z

measured 28 of 28 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

28 of 28 outbound references displayed

  • verified exact1
  • verified fuzzy11
  • unresolved16
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation d66ee5d5-2095-4980-be2e-7cb74b5fcf68 · outbound

This paper cites Bowman, Ethan Perez, Roger Baker Grosse, and David Duvenaud.

An Example Safety Case for Safeguards Against Misuse Bowman, Ethan Perez, Roger Baker Grosse, and David Duvenaud

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:41:43.363295Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:41:38.016972Z digest=sha256:80954e0e54911b93857be7d0f57423fca9ec69477fc4d2d14acfe0846c8292df

Observation cdd72cf9-9fa9-4a73-b8cb-21c7647da3b2 · outbound

This paper cites Responsible scaling policy evaluations report - claude 3 opus.

An Example Safety Case for Safeguards Against Misuse Responsible scaling policy evaluations report - claude 3 opus

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:41:43.223956Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:41:38.111176Z digest=sha256:b4c69f7896d4500ba346b4a8847cbd6287fa5cbc814352e1169af8dcfd827813

Observation c7dc25e9-e5b4-47ec-93c1-ac00f790029b · outbound

This paper cites Claude 3.7 sonnet system card.

An Example Safety Case for Safeguards Against Misuse Claude 3.7 sonnet system card

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:41:43.063608Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:41:38.214842Z digest=sha256:08d2a9a509bd53e25146fde9d5aac980bb515938c350dc4d757678a70ad96bc3

Observation 0752dca4-f42e-44c0-9618-9c5525a20ede · outbound

This paper cites Activating ai safety level 3 protections.

An Example Safety Case for Safeguards Against Misuse Activating ai safety level 3 protections

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:41:42.906642Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:41:38.322428Z digest=sha256:b7eacf8a7a89dd0a56b50331bea65bccf0dc82e97891566c2d4654529bb7ba9c

Observation a5066116-86d9-4b43-92dc-c417c28d272a · outbound

This paper cites Constitutional AI: Harmlessness from AI Feedback.

An Example Safety Case for Safeguards Against Misuse Constitutional AI: Harmlessness from AI Feedback

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T14:41:38.423224Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:41:38.423224Z digest=sha256:7b5440072e10a6561c0e2a6b8720c34afb8b22b6c08bcb5ff1fe1ca894ba1b2d

Observation c1a6242e-c1da-4a20-b242-d659f65ada4c · outbound

This paper cites Balesni, A.

An Example Safety Case for Safeguards Against Misuse Balesni, A

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:41:42.741164Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:41:38.578597Z digest=sha256:a0fc3cd795478596cc93fba7ad40d3b3b6d1db4d735957229b9885e4189ae707

Observation 8b4d89ca-9110-4f62-bb7b-0ad8574a1154 · outbound

This paper cites Societal Adaptation to Advanced AI.

An Example Safety Case for Safeguards Against Misuse Societal Adaptation to Advanced AI

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T14:41:38.693339Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:41:38.693339Z digest=sha256:e06acd9af8fb5a4c0aadbf86d26e9aad44679010003dc61ca63fb58071030305

Observation c1ba3a80-6061-4e84-9798-e27db1fdf7f0 · outbound

This paper cites Safety cases for frontier AI.

An Example Safety Case for Safeguards Against Misuse Safety cases for frontier AI

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T14:41:38.796591Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:41:38.796591Z digest=sha256:e296b02d602f9bf4ad811e2032c78d353f5b759b5475b756f9f1a254d667693b

Observation ccda1e91-a6f2-46b1-b0f3-c64170a11e95 · outbound

This paper cites How can safety cases be used to help with frontier AI safety? Technical report, AI Security Institute, February 2025.

An Example Safety Case for Safeguards Against Misuse How can safety cases be used to help with frontier AI safety? Technical report, AI Security Institute, February 2025

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:41:42.549388Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:41:38.909037Z digest=sha256:55557c921ef1484e91066e157a9942f2a74237d2216c3be42775c3ed5bd33b72

Observation 6790c829-70d6-4fd0-8576-ab17b5b14dc2 · outbound

This paper cites Deep reinforcement learning from human preferences.

An Example Safety Case for Safeguards Against Misuse Deep reinforcement learning from human preferences

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T14:41:39.026169Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:41:39.026169Z digest=sha256:1b2aea14d08ef21ffa70d2d1dec30654387c0975f91a77aa82a2f92eb8a0f029

Observation b1d411bd-9eb1-47bc-ae27-0de0a641b9aa · outbound

This paper cites Safety Cases: How to Justify the Safety of Advanced AI Systems.

An Example Safety Case for Safeguards Against Misuse Safety Cases: How to Justify the Safety of Advanced AI Systems

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T14:41:39.162667Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:41:39.162667Z digest=sha256:ad6ebf30ac9a93ed3e8836b48ac80ed8d05ce492e1164f9ba177ba624b5b7b95

Observation 50d78cfe-22c5-4e03-8230-578b8e45e0af · outbound

This paper cites Extending control evaluations to non-scheming threats, jan 2025.

An Example Safety Case for Safeguards Against Misuse Extending control evaluations to non-scheming threats, jan 2025

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:41:42.378003Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:41:39.278205Z digest=sha256:0e5897d97a4037d2b115fe438847427c214248240b3697d1ea248086d84c7b7a

Observation 8f30d400-60e3-4ff9-a5a5-c0fcfe46ef9b · outbound

This paper cites Does Unlearning Truly Unlearn? A Black Box Evaluation of LLM Unlearning Methods.

An Example Safety Case for Safeguards Against Misuse Does Unlearning Truly Unlearn? A Black Box Evaluation of LLM Unlearning Methods

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T14:41:39.409032Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:41:39.409032Z digest=sha256:f5af91f89a196617b00aece48d42fe2b9741cc562e8298c1b7c6109adb04719e

Observation 71992eba-19a2-4c33-8066-9922db239e57 · outbound

This paper cites Safety case template for frontier AI: A cyber inability argument.

An Example Safety Case for Safeguards Against Misuse Safety case template for frontier AI: A cyber inability argument

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T14:41:39.558791Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:41:39.558791Z digest=sha256:59bf2b15196144b0263676995e62da6bb088ff091df02ce140a494131d3b7f64

Observation 80032a35-0929-414f-a755-7f8afe57a355 · outbound

This paper cites Gray swan arena, 2025.

An Example Safety Case for Safeguards Against Misuse Gray swan arena, 2025

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:41:42.211978Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:41:39.677876Z digest=sha256:01b8f7197423b534926099277af2206e89dafa921955bc6ae5c5ff3de8fe3657

Observation 001444e2-1725-41ad-ac4d-8119978ad9cc · outbound

This paper cites Alignment faking in large language models.

An Example Safety Case for Safeguards Against Misuse Alignment faking in large language models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T14:41:39.815422Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:41:39.815422Z digest=sha256:efa6cf18f6ddaf41fc47b895eba19b12cd162006ffaeaf3ac28d4229c1a96e5a

Observation 32f2bbe1-000b-499e-944a-72cae53d7fe7 · outbound

This paper cites Best-of-N Jailbreaking.

An Example Safety Case for Safeguards Against Misuse Best-of-N Jailbreaking

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T14:41:39.930783Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:41:39.930783Z digest=sha256:7fc7e1ee7d277bcf54cb48b4229aaaf0f2335902cafd11aa30acc709c960adfb

Observation 82cf8a93-0d39-4176-8680-8947d77416c1 · outbound

This paper cites A sketch of an AI control safety case.

An Example Safety Case for Safeguards Against Misuse A sketch of an AI control safety case

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T14:41:40.068055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:41:40.068055Z digest=sha256:0a12b2d8cb2ec899becf8711021c5afb7c848e9283151222faeab16e53473c81

Observation fbb84434-3f2e-492b-b6df-b7c628dee74b · outbound

This paper cites The WMDP Benchmark: Measuring and Reducing Malicious Use With Unlearning.

An Example Safety Case for Safeguards Against Misuse The WMDP Benchmark: Measuring and Reducing Malicious Use With Unlearning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T14:41:40.218546Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:41:40.218546Z digest=sha256:1ed4f20418654d78b263bc3c1dbe56a27bbadb05de1df4ca53dc087702f3198f

Observation 0279391a-c922-4b9c-8d55-3403b2325c8d · outbound

This paper cites Tree of Attacks: Jailbreaking Black-Box LLMs Automatically.

An Example Safety Case for Safeguards Against Misuse Tree of Attacks: Jailbreaking Black-Box LLMs Automatically

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T14:41:40.333893Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:41:40.333893Z digest=sha256:591dc16bbc14add9c512443f675a169601e124f2b5deb5b49d8f5eb3b10cd0d7

Observation 898cbfea-2683-4725-9463-d26bf8a9e935 · outbound

This paper cites Nguyen, M.

An Example Safety Case for Safeguards Against Misuse Nguyen, M

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T14:41:40.441905Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:41:40.441905Z digest=sha256:2248307163042a660a68f6712bcdc20bd002c9c3730d5aeae7bb99857daba434

Observation ae23b508-70ec-4995-b4ad-943ba75a67a0 · outbound

This paper cites Deep research system card.

An Example Safety Case for Safeguards Against Misuse Deep research system card

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:41:42.005787Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:41:40.516495Z digest=sha256:c3a6f0b62cec6bc5696da6f6750e94fc2f28b0156811a39bba6bdea6191ac74f

Observation e91d6733-19f4-401e-8082-22420fce93d6 · outbound

This paper cites Constitutional Classifiers: Defending against Universal Jailbreaks across Thousands of Hours of Red Teaming.

An Example Safety Case for Safeguards Against Misuse Constitutional Classifiers: Defending against Universal Jailbreaks across Thousands of Hours of Red Teaming

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T14:41:40.653560Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:41:40.653560Z digest=sha256:815ad6e0b2083e7b29f546a6800d84fc4936e12bd0630307c71ec7ac79b99e6f

Observation f18cbf0e-99a7-4a2a-b775-ce5f1fb5cac4 · outbound

This paper cites Structured access: an emerging paradigm for safe AI deployment.

An Example Safety Case for Safeguards Against Misuse Structured access: an emerging paradigm for safe AI deployment

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T14:41:40.747237Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:41:40.747237Z digest=sha256:45f5a739a481492e0b0d7c76b1b91fac59cb5566a26e6d430a3b6850c0ec4b10

Observation 8844ff6e-13f7-4be2-9660-abccba0fba33 · outbound

This paper cites Principles for safeguard evaluation.

An Example Safety Case for Safeguards Against Misuse Principles for safeguard evaluation

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:41:41.809116Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:41:40.824626Z digest=sha256:8b0a167807da13e154d8bdeeca7d74ff65bf87ed1462f8fb8b6420c8f1e14362

Observation dd35fb3a-52d6-4909-8173-4c64e0a58071 · outbound

This paper cites Us aisi and uk aisi joint pre-deployment test: Anthropic's claude 3.5 sonnet (october 2024 release).

An Example Safety Case for Safeguards Against Misuse Us aisi and uk aisi joint pre-deployment test: Anthropic's claude 3.5 sonnet (october 2024 release)

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:41:41.611674Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:41:40.895645Z digest=sha256:15bdcd0599936367f24ad95f310310b56c23ac7fc56af0c11ec5b105eea0e4da

Observation f825ba9c-f8be-4e75-a595-da2a2cb40574 · outbound

This paper cites Wasil, J.

An Example Safety Case for Safeguards Against Misuse Wasil, J

Reference 27

Resolution
verified exact
doi, observed 2026-08-07T14:41:41.318936Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:41:41.001436Z digest=sha256:ed9b819f5ee889e54125a6ebc103ca722046cba2a7485bdd1635960c4474998c

Observation c059c0c2-da9d-472f-886a-598f8679fc0d · outbound

This paper cites Jailbroken: How Does LLM Safety Training Fail?.

An Example Safety Case for Safeguards Against Misuse Jailbroken: How Does LLM Safety Training Fail?

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T14:41:41.097514Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:41:41.097514Z digest=sha256:4cf1a47a3305c51d2ad814211468ae6b79da160f7c46175e9a954bc56aa5d829

Pith citing papers

No inbound Pith citation observations are available.