Pith. sign in

Paper Citation Record · LEDGER

An Example Safety Case for Safeguards Against Misuse

As of 14 August 2026, this Paper Citation Record lists 28 of 28 outbound references and 0 inbound Pith citation observations for arXiv:2505.18003.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.18003 v1

Coverage vector

measured 28 of 28 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:41:41.097514Z

measured 28 of 28 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

28 of 28 outbound references displayed

  • verified exact1
  • verified fuzzy11
  • unresolved16
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation d66ee5d5-2095-4980-be2e-7cb74b5fcf68 · outbound

This paper cites Bowman, Ethan Perez, Roger Baker Grosse, and David Duvenaud.

An Example Safety Case for Safeguards Against Misuse Bowman, Ethan Perez, Roger Baker Grosse, and David Duvenaud

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:41:43.363295Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-07T14:41:38.016972Z digest=sha256:1cc7f403d176bf185386d6d161716b357afe75df86ae05c5661064695ca41e47

Observation cdd72cf9-9fa9-4a73-b8cb-21c7647da3b2 · outbound

This paper cites Responsible scaling policy evaluations report - claude 3 opus.

An Example Safety Case for Safeguards Against Misuse Responsible scaling policy evaluations report - claude 3 opus

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:41:43.223956Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-07T14:41:38.111176Z digest=sha256:8863628d9c0358557c1dfafaf0efbe2a643f39a160f3c5f6ca26488d45fc0879

Observation c7dc25e9-e5b4-47ec-93c1-ac00f790029b · outbound

This paper cites Claude 3.7 sonnet system card.

An Example Safety Case for Safeguards Against Misuse Claude 3.7 sonnet system card

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:41:43.063608Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-07T14:41:38.214842Z digest=sha256:714375a2ffef123270ce7c8ed1d4bda99cf60c6d4d949a4f58b2313e3e297d66

Observation 0752dca4-f42e-44c0-9618-9c5525a20ede · outbound

This paper cites Activating ai safety level 3 protections.

An Example Safety Case for Safeguards Against Misuse Activating ai safety level 3 protections

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:41:42.906642Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-07T14:41:38.322428Z digest=sha256:5c98e2e94c6bc6678368fb9855fa562872871b235bfabb8e82a4904349ba6786

Observation a5066116-86d9-4b43-92dc-c417c28d272a · outbound

This paper cites Constitutional AI: Harmlessness from AI Feedback.

An Example Safety Case for Safeguards Against Misuse Constitutional AI: Harmlessness from AI Feedback

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T14:41:38.423224Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:41:38.423224Z digest=sha256:41aa0264f2a0f4ad21152e77383ce07aeb49cadb7705ca80e4d8fbd6a4f1ab7e

Observation c1a6242e-c1da-4a20-b242-d659f65ada4c · outbound

This paper cites Balesni, A.

An Example Safety Case for Safeguards Against Misuse Balesni, A

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:41:42.741164Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-07T14:41:38.578597Z digest=sha256:647e0f8150c5192879344218ef9b0e5f79ec2b66b894c96b084872699307b0ae

Observation 8b4d89ca-9110-4f62-bb7b-0ad8574a1154 · outbound

This paper cites Societal Adaptation to Advanced AI.

An Example Safety Case for Safeguards Against Misuse Societal Adaptation to Advanced AI

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T14:41:38.693339Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:41:38.693339Z digest=sha256:b4b7452b80eb34901b57a82e57259d3a88de196b679aa0418b5f4310b2e07a18

Observation c1ba3a80-6061-4e84-9798-e27db1fdf7f0 · outbound

This paper cites Safety cases for frontier AI.

An Example Safety Case for Safeguards Against Misuse Safety cases for frontier AI

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T14:41:38.796591Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:41:38.796591Z digest=sha256:2b8b642f034f4490d206be32a9c3b536cf2a8d4738d0c0108ddc5f86580bf03e

Observation ccda1e91-a6f2-46b1-b0f3-c64170a11e95 · outbound

This paper cites How can safety cases be used to help with frontier AI safety? Technical report, AI Security Institute, February 2025.

An Example Safety Case for Safeguards Against Misuse How can safety cases be used to help with frontier AI safety? Technical report, AI Security Institute, February 2025

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:41:42.549388Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-07T14:41:38.909037Z digest=sha256:e01adc6625cc5fabf8e5f8f082236226ed989ec00d46e4aeb92b2aa383099eff

Observation 6790c829-70d6-4fd0-8576-ab17b5b14dc2 · outbound

This paper cites Deep reinforcement learning from human preferences.

An Example Safety Case for Safeguards Against Misuse Deep reinforcement learning from human preferences

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T14:41:39.026169Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:41:39.026169Z digest=sha256:272d553918c284f932f039aa15319cd6663558d1ca3a1864741c22b61c158a36

Observation b1d411bd-9eb1-47bc-ae27-0de0a641b9aa · outbound

This paper cites Safety Cases: How to Justify the Safety of Advanced AI Systems.

An Example Safety Case for Safeguards Against Misuse Safety Cases: How to Justify the Safety of Advanced AI Systems

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T14:41:39.162667Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:41:39.162667Z digest=sha256:d04825873e1043b14807d1933376ee70b3746059c95275aa50fab20eb7e39ae5

Observation 50d78cfe-22c5-4e03-8230-578b8e45e0af · outbound

This paper cites Extending control evaluations to non-scheming threats, jan 2025.

An Example Safety Case for Safeguards Against Misuse Extending control evaluations to non-scheming threats, jan 2025

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:41:42.378003Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-07T14:41:39.278205Z digest=sha256:7ca3dbb0a8b14ea9740ceb9acf83fc54509847fad69ee11ee876e8a79f207b6d

Observation 8f30d400-60e3-4ff9-a5a5-c0fcfe46ef9b · outbound

This paper cites Does Unlearning Truly Unlearn? A Black Box Evaluation of LLM Unlearning Methods.

An Example Safety Case for Safeguards Against Misuse Does Unlearning Truly Unlearn? A Black Box Evaluation of LLM Unlearning Methods

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T14:41:39.409032Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:41:39.409032Z digest=sha256:a5cdc9c858fb90b288685f7cb9641d4aa7a2d957ecb1f144c0df23f380bef23f

Observation 71992eba-19a2-4c33-8066-9922db239e57 · outbound

This paper cites Safety case template for frontier AI: A cyber inability argument.

An Example Safety Case for Safeguards Against Misuse Safety case template for frontier AI: A cyber inability argument

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T14:41:39.558791Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:41:39.558791Z digest=sha256:e4a95789ede2b6d9101a32e32a0e35e8bf3e45af6d2fed812001d38b1d9ed2a3

Observation 80032a35-0929-414f-a755-7f8afe57a355 · outbound

This paper cites Gray swan arena, 2025.

An Example Safety Case for Safeguards Against Misuse Gray swan arena, 2025

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:41:42.211978Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-07T14:41:39.677876Z digest=sha256:42df35db4f4046d2df5b9a02f5d3843b7b2fd1c635185675e33c092709ddd853

Observation 001444e2-1725-41ad-ac4d-8119978ad9cc · outbound

This paper cites Alignment faking in large language models.

An Example Safety Case for Safeguards Against Misuse Alignment faking in large language models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T14:41:39.815422Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:41:39.815422Z digest=sha256:a887d5eaff5603f8e045ab9651d03973dcddec2358ea0b9aff1b7d85fc61aee7

Observation 32f2bbe1-000b-499e-944a-72cae53d7fe7 · outbound

This paper cites Best-of-N Jailbreaking.

An Example Safety Case for Safeguards Against Misuse Best-of-N Jailbreaking

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T14:41:39.930783Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:41:39.930783Z digest=sha256:75d076e563c09e61eadf7b015bedc4fef5e1cb7ee9ec3abeec0e2e049c1ee7fb

Observation 82cf8a93-0d39-4176-8680-8947d77416c1 · outbound

This paper cites A sketch of an AI control safety case.

An Example Safety Case for Safeguards Against Misuse A sketch of an AI control safety case

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T14:41:40.068055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:41:40.068055Z digest=sha256:ada959a202322c87c71cf4bd3fe701d74b47d869663163a7d259cb572ff42053

Observation fbb84434-3f2e-492b-b6df-b7c628dee74b · outbound

This paper cites The WMDP Benchmark: Measuring and Reducing Malicious Use With Unlearning.

An Example Safety Case for Safeguards Against Misuse The WMDP Benchmark: Measuring and Reducing Malicious Use With Unlearning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T14:41:40.218546Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:41:40.218546Z digest=sha256:5b076552561249cc76e343fbec8a351e7790025148c453259f0a07c200580300

Observation 0279391a-c922-4b9c-8d55-3403b2325c8d · outbound

This paper cites Tree of Attacks: Jailbreaking Black-Box LLMs Automatically.

An Example Safety Case for Safeguards Against Misuse Tree of Attacks: Jailbreaking Black-Box LLMs Automatically

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T14:41:40.333893Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:41:40.333893Z digest=sha256:7f1a29a124462429d11716ba1df7cce027dcdada7c4fa842292e0b6f1166155e

Observation 898cbfea-2683-4725-9463-d26bf8a9e935 · outbound

This paper cites Nguyen, M.

An Example Safety Case for Safeguards Against Misuse Nguyen, M

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T14:41:40.441905Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:41:40.441905Z digest=sha256:6d8eb236889801bad136a4cb65e4c5969ee87634b3fd26defffe3ba759604def

Observation ae23b508-70ec-4995-b4ad-943ba75a67a0 · outbound

This paper cites Deep research system card.

An Example Safety Case for Safeguards Against Misuse Deep research system card

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:41:42.005787Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-07T14:41:40.516495Z digest=sha256:532afa7efcb6cf1e34ae497a48dd1968714c51ae45a433dda2c33b5589740f28

Observation e91d6733-19f4-401e-8082-22420fce93d6 · outbound

This paper cites Constitutional Classifiers: Defending against Universal Jailbreaks across Thousands of Hours of Red Teaming.

An Example Safety Case for Safeguards Against Misuse Constitutional Classifiers: Defending against Universal Jailbreaks across Thousands of Hours of Red Teaming

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T14:41:40.653560Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:41:40.653560Z digest=sha256:f2082c1be9596590dd51def2601c9cd2200ab4a1adc428f2198a0eecc1e8b409

Observation f18cbf0e-99a7-4a2a-b775-ce5f1fb5cac4 · outbound

This paper cites Structured access: an emerging paradigm for safe AI deployment.

An Example Safety Case for Safeguards Against Misuse Structured access: an emerging paradigm for safe AI deployment

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T14:41:40.747237Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:41:40.747237Z digest=sha256:ab529b802fa3842cbc28ef3223475851e01c1a9e21fa652004b8f64df7eeb63d

Observation 8844ff6e-13f7-4be2-9660-abccba0fba33 · outbound

This paper cites Principles for safeguard evaluation.

An Example Safety Case for Safeguards Against Misuse Principles for safeguard evaluation

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:41:41.809116Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-07T14:41:40.824626Z digest=sha256:f2e959a3f5e7f3349898279503750bb35a5243456c9eaa3e71948a49c4fb4f22

Observation dd35fb3a-52d6-4909-8173-4c64e0a58071 · outbound

This paper cites Us aisi and uk aisi joint pre-deployment test: Anthropic's claude 3.5 sonnet (october 2024 release).

An Example Safety Case for Safeguards Against Misuse Us aisi and uk aisi joint pre-deployment test: Anthropic's claude 3.5 sonnet (october 2024 release)

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:41:41.611674Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-07T14:41:40.895645Z digest=sha256:b4f4665d99dead0515e6685c59e75df14ecef7374f4462f91a6b70d1a6122187

Observation f825ba9c-f8be-4e75-a595-da2a2cb40574 · outbound

This paper cites Wasil, J.

An Example Safety Case for Safeguards Against Misuse Wasil, J

Reference 27

Resolution
verified exact
doi, observed 2026-08-07T14:41:41.318936Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-07T14:41:41.001436Z digest=sha256:d58063e090474bd29d4a9368cfbf5c2b0d591a4218675a152a97668efff2b02f

Observation c059c0c2-da9d-472f-886a-598f8679fc0d · outbound

This paper cites Jailbroken: How Does LLM Safety Training Fail?.

An Example Safety Case for Safeguards Against Misuse Jailbroken: How Does LLM Safety Training Fail?

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T14:41:41.097514Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:41:41.097514Z digest=sha256:f3b1bd1071639f5ab19a781f8bd3dbb58352171505f8e8e984da9810a192aac2

Pith citing papers

No inbound Pith citation observations are available.