Pith. sign in

Paper Citation Record · LEDGER

Tamper-Resistant Safeguards for Open-Weight LLMs

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 23 inbound Pith citation observations for arXiv:2408.00761.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2408.00761 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 23 of 23 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 23 of 23 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:05:55.125060Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T17:58:47.266895Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 84378219-be70-47eb-b045-e8be90f1fa93 · inbound

Harmful Fine-tuning Attacks and Defenses for Large Language Models: A Survey cites this paper.

Harmful Fine-tuning Attacks and Defenses for Large Language Models: A Survey Tamper-Resistant Safeguards for Open-Weight LLMs

Reference 144

Resolution
verified exact
arxiv_id, observed 2026-05-23T20:58:25.958886Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-23T20:58:16.237327Z digest=sha256:76a32d5aef229d6ba9e9b2430ae31a0641e0809defbaa9e7258589605d74c663

Observation 99ed0a68-db84-4abd-bb7c-0eabd1b42ff8 · inbound

Secure LLM Fine-Tuning via Safety-Aware Probing cites this paper.

Secure LLM Fine-Tuning via Safety-Aware Probing Tamper-Resistant Safeguards for Open-Weight LLMs

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-22T13:11:35.792350Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-22T13:07:09.402763Z digest=sha256:d2a602839f38844f495a7fff86c31a0ebd68ad844c3eee729e5e654916d500e2

Observation cedf324f-8abb-4e10-95c4-f2a870cba0f3 · inbound

Existing Large Language Model Unlearning Evaluations Are Inconclusive cites this paper.

Existing Large Language Model Unlearning Evaluations Are Inconclusive Tamper-Resistant Safeguards for Open-Weight LLMs

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T12:05:55.125060Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:05:55.125060Z digest=sha256:67091481d3c0ff49819e5afc463c42c3dddb4d444e9f5708cceffad64fe3876c

Observation 72cbd017-c19c-46bb-ae8f-632a1a95f874 · inbound

Prompt Attacks Reveal Superficial Knowledge Removal in Unlearning Methods cites this paper.

Prompt Attacks Reveal Superficial Knowledge Removal in Unlearning Methods Tamper-Resistant Safeguards for Open-Weight LLMs

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T04:37:07.881113Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:37:07.881113Z digest=sha256:7a767a8f6e317d0f60a5d31869d344b5545061e6b3b89a981a255b5bef6693be

Observation cb9ce222-09c4-4978-ae18-8bf30d6a1af2 · inbound

FORTRESS: Frontier Risk Evaluation for National Security and Public Safety cites this paper.

FORTRESS: Frontier Risk Evaluation for National Security and Public Safety Tamper-Resistant Safeguards for Open-Weight LLMs

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:04.114372Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:04.114372Z digest=sha256:3fda965a9deea6238131ac417519af6a542edd378ad022bc8a7a48fe44ad95dc

Observation a075f13e-8f9b-48b1-990d-36327b68a518 · inbound

Step-by-Step Reasoning Attack: Revealing 'Erased' Knowledge in Large Language Models cites this paper.

Step-by-Step Reasoning Attack: Revealing 'Erased' Knowledge in Large Language Models Tamper-Resistant Safeguards for Open-Weight LLMs

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T00:56:31.605239Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:56:31.605239Z digest=sha256:8174606ac1ba86831bc96de134abf25f17f07ba657f6ba70ad30fca8056e679c

Observation 6a385aea-d936-4f8c-aa82-dbb279c9b69a · inbound

Technical Requirements for Halting Dangerous AI Activities cites this paper.

Technical Requirements for Halting Dangerous AI Activities Tamper-Resistant Safeguards for Open-Weight LLMs

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T17:50:30.826374Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:50:30.826374Z digest=sha256:d2e93d2d3f8a64a799fdb4ed6dc5ce1a10562bf4cacf29562efd79545c084470

Observation f87b89ab-3b43-4f80-b269-6f3f7885f2de · inbound

The Safety Gap Toolkit: Evaluating Hidden Dangers of Open-Source Models cites this paper.

The Safety Gap Toolkit: Evaluating Hidden Dangers of Open-Source Models Tamper-Resistant Safeguards for Open-Weight LLMs

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-06T19:08:26.524698Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:08:26.524698Z digest=sha256:d3e4b65e178a5e15a249537b7c5132c3f848b2f489d156deab6c37eac8c4f346

Observation 79694803-20ba-4bc8-a100-c1970821ae51 · inbound

Towards Integrated Alignment cites this paper.

Towards Integrated Alignment Tamper-Resistant Safeguards for Open-Weight LLMs

Reference 2025

Resolution
malformed identifier
no resolver link, observed 2026-08-05T22:56:28.187316Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T22:56:28.187316Z digest=sha256:edca360db0295c8ee0f3e55d39d28f200a38c28f06362fc80bc62b089d6f2ded

Observation 7b786040-f6e5-476a-bf06-72165ed3e913 · inbound

Downgrade to Upgrade: Optimizer Simplification Enhances Robustness in LLM Unlearning cites this paper.

Downgrade to Upgrade: Optimizer Simplification Enhances Robustness in LLM Unlearning Tamper-Resistant Safeguards for Open-Weight LLMs

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-18T10:46:17.116528Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-18T10:44:53.516653Z digest=sha256:332d3478bc4525cc7d8fe5129305392c35aabb11ced9fc3207412ef58d278e46

Observation b82adb25-374f-4802-bf7b-ceffd2bb1671 · inbound

CacheTrap: Unveiling a Stealthier Gray-Box Trojan against LLMs cites this paper.

CacheTrap: Unveiling a Stealthier Gray-Box Trojan against LLMs Tamper-Resistant Safeguards for Open-Weight LLMs

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-17T04:14:00.155118Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-17T04:13:13.231293Z digest=sha256:88c198239398ebe326d8723da420a3fd484a2fa3d9d2df1e7c0004f1354086f7

Observation 5409273a-ae7e-4ee8-beaf-5d42683a99a9 · inbound

RippleBench: Capturing Ripple Effects Using Existing Knowledge Repositories cites this paper.

RippleBench: Capturing Ripple Effects Using Existing Knowledge Repositories Tamper-Resistant Safeguards for Open-Weight LLMs

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-03T18:41:34.212799Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:41:34.212799Z digest=sha256:76271a755127af85d422735ec4bbb59f975a0a38755b2db0ab21eb680b403f6d

Observation 3cb9b195-6a58-47f2-9ecf-447dd540ddd0 · inbound

Robust Policy Optimization to Prevent Catastrophic Forgetting cites this paper.

Robust Policy Optimization to Prevent Catastrophic Forgetting Tamper-Resistant Safeguards for Open-Weight LLMs

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-05-16T05:37:24.175635Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-16T05:33:42.965249Z digest=sha256:354caa11fbd884d906d1f4cc13b818ac50dc33be832b9f5686e6e94be6f8efd1

Observation 318d1ff5-2362-4a97-8bd0-5f0406b44834 · inbound

Safety Drift After Fine-Tuning: Evidence from High-Stakes Domains cites this paper.

Safety Drift After Fine-Tuning: Evidence from High-Stakes Domains Tamper-Resistant Safeguards for Open-Weight LLMs

Reference 56

Resolution
verified exact
arxiv_id, observed 2026-05-11T23:16:14.104378Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-07T17:53:57.169962Z digest=sha256:bbf035956746aaef8daabb446f81433afa02d4d135175f0d972718b77e33b05e

Observation 98e7d853-8b19-4e6a-b457-09320dd21d99 · inbound

Robust LLM Unlearning Against Relearning Attacks: The Minor Components in Representations Matter cites this paper.

Robust LLM Unlearning Against Relearning Attacks: The Minor Components in Representations Matter Tamper-Resistant Safeguards for Open-Weight LLMs

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-13T01:06:59.987995Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-13T01:06:44.126131Z digest=sha256:c9e6905b50f6c27bd245526f2f832af56c0a24a031a6b365b3f9b0778981e230

Observation c736ebea-b22e-40f4-9696-7b30414548fc · inbound

One Step to the Side: Why Defenses Against Malicious Finetuning Fail Under Adaptive Adversaries cites this paper.

One Step to the Side: Why Defenses Against Malicious Finetuning Fail Under Adaptive Adversaries Tamper-Resistant Safeguards for Open-Weight LLMs

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-06-30T21:05:04.108905Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-30T21:01:25.549340Z digest=sha256:f422f276c4c85ccc76788e49769f2fb617b066dee8abc2570177a8b6e4e2bf62

Observation 943069a1-ce44-4af4-aa9e-9cc7adefb9c1 · inbound

RepSelect: Robust LLM Unlearning via Representation Selectivity cites this paper.

RepSelect: Robust LLM Unlearning via Representation Selectivity Tamper-Resistant Safeguards for Open-Weight LLMs

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-07-03T17:58:47.268760Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-27T03:34:32.388152Z digest=sha256:eef0af1b63cd4ab40c0edc32d0f87b9d69c1e78a5634d8fbfa96ba67e9351226

Observation 0685c087-0354-46ee-85c4-07b18f4a1766 · inbound

FlipGuard: Defending Large Language Models Against Quantization-Conditioned Backdoor Attacks cites this paper.

FlipGuard: Defending Large Language Models Against Quantization-Conditioned Backdoor Attacks Tamper-Resistant Safeguards for Open-Weight LLMs

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-06-30T09:44:37.661358Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-30T09:39:35.673341Z digest=sha256:7bc9b42437758b1f25eb81a2709545f01ff34e9a549827248564ef4fbc24d5ee

Observation b67c46e8-f480-457c-985b-e1b14276b069 · inbound

Breaking the Rounding Trap: Securing LLMs against Quantization-Conditioned Backdoors cites this paper.

Breaking the Rounding Trap: Securing LLMs against Quantization-Conditioned Backdoors Tamper-Resistant Safeguards for Open-Weight LLMs

Reference 53

Resolution
verified exact
arxiv_id, observed 2026-06-30T08:04:28.742903Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-30T07:47:18.350953Z digest=sha256:b286ac81ffaaf35facd6dccb92498ee041d6d72fa6867e88260c0fdcc586757f

Observation 38d2f7f6-5c06-442c-8c8b-ac8bcc26e93d · inbound

Breaking the Rounding Trap: Securing LLMs against Quantization-Conditioned Backdoors cites this paper.

Breaking the Rounding Trap: Securing LLMs against Quantization-Conditioned Backdoors Tamper-Resistant Safeguards for Open-Weight LLMs

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-04T04:39:06.903773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T04:39:06.903773Z digest=sha256:8bf18dbcd24735130979494b6be08c244ff3528ccf2c3436c3bc0230a3fb74c9

Observation 3173a2fd-05d9-4853-8bcc-199d3f156a70 · inbound

LLM Unlearning for Cyber Defense: A Survey on Methods, Challenges, and Emerging Threats cites this paper.

LLM Unlearning for Cyber Defense: A Survey on Methods, Challenges, and Emerging Threats Tamper-Resistant Safeguards for Open-Weight LLMs

Reference 121

Resolution
unresolved
no resolver link, observed 2026-08-02T10:25:19.922744Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T10:25:19.922744Z digest=sha256:fe7acdb949079773689435de4473676325cc1fa70da4ddae2b3bfb2302c3d728

Observation f1f6f2b5-8b9c-41e2-a1b3-d9be574f659c · inbound

Engineering Trustworthy Agentic AI for Critical Systems cites this paper.

Engineering Trustworthy Agentic AI for Critical Systems Tamper-Resistant Safeguards for Open-Weight LLMs

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-01T15:07:27.834694Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:07:27.834694Z digest=sha256:938914e3896f4599562d6eebd1e5660ff1fb51ed6f11ea644938799d28fdaa96

Observation 94205cbb-2c55-4207-bd4c-53627cf2b5ca · inbound

Emergent Misalignment Recruits a Pre-existing Persona Subspace cites this paper.

Emergent Misalignment Recruits a Pre-existing Persona Subspace Tamper-Resistant Safeguards for Open-Weight LLMs

Reference 204

Resolution
unresolved
no resolver link, observed 2026-08-01T07:46:22.166900Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T07:46:22.166900Z digest=sha256:a8fb134e7437a1de659e602f218fb126ff32b633a9f1f5610d348fcaa7270d6f