Pith. sign in

Paper Citation Record · LEDGER

Antidote: Post-fine-tuning Safety Alignment for Large Language Models against Harmful Fine-tuning

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 12 inbound Pith citation observations for arXiv:2408.09600.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2408.09600 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 12 of 12 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 12 of 12 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:15:18.856097Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T20:47:22.918490Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation eeb9954b-7eba-4d31-904c-cca1ddc6ab21 · inbound

Harmful Fine-tuning Attacks and Defenses for Large Language Models: A Survey cites this paper.

Harmful Fine-tuning Attacks and Defenses for Large Language Models: A Survey Antidote: Post-fine-tuning Safety Alignment for Large Language Models against Harmful Fine-tuning

Reference 64

Resolution
verified exact
arxiv_id, observed 2026-05-23T20:58:25.832290Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-23T20:58:16.237327Z digest=sha256:34e016ed8fb37dfc62f794058d35e065d5915750409a6c63f4905138225deb14

Observation 36e46723-4c7c-4503-b2a8-e02bd740e5ec · inbound

Safety at Scale: A Comprehensive Survey of Large Model and Agent Safety cites this paper.

Safety at Scale: A Comprehensive Survey of Large Model and Agent Safety Antidote: Post-fine-tuning Safety Alignment for Large Language Models against Harmful Fine-tuning

Reference 123

Resolution
verified exact
arxiv_id, observed 2026-05-23T04:42:34.050386Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-23T04:39:04.591722Z digest=sha256:6840aacf95b3bb2c98c75c1a86a3c4aec9f13c252891d9e8a7135c852dd3a88e

Observation 5ed09cab-fdb5-4a0a-a16f-8cb48dd54f84 · inbound

Reshaping Representation Space to Balance the Safety and Over-rejection in Large Audio Language Models cites this paper.

Reshaping Representation Space to Balance the Safety and Over-rejection in Large Audio Language Models Antidote: Post-fine-tuning Safety Alignment for Large Language Models against Harmful Fine-tuning

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T14:15:18.856097Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:15:18.856097Z digest=sha256:bc4b2f32fdd58df6bcf88815f3134a9163ee4a916227f91e3a37e06ed754eda6

Observation 5ce7a588-a395-4725-802b-23fa97cb6001 · inbound

Vulnerability-Aware Alignment: Mitigating Uneven Forgetting in Harmful Fine-Tuning cites this paper.

Vulnerability-Aware Alignment: Mitigating Uneven Forgetting in Harmful Fine-Tuning Antidote: Post-fine-tuning Safety Alignment for Large Language Models against Harmful Fine-tuning

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T10:58:29.556438Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:58:29.556438Z digest=sha256:c076e1445ba331adad90e7e66c7fd470e3cc42af9c8dd99671fa0f284460851e

Observation fa4eff3d-7ade-413d-a974-1b00d9d22bc5 · inbound

LoX: Low-Rank Extrapolation Robustifies LLM Safety Against Fine-tuning cites this paper.

LoX: Low-Rank Extrapolation Robustifies LLM Safety Against Fine-tuning Antidote: Post-fine-tuning Safety Alignment for Large Language Models against Harmful Fine-tuning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T23:59:56.106685Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:59:56.106685Z digest=sha256:1db8cc50653cb6942ef26ade7fce7ee40921b76b197a400680c80eae8d47cb68

Observation 497dbac8-a173-476f-b9d7-893127262d88 · inbound

Guardians and Offenders: A Survey on Harmful Content Generation and Safety Mitigation of LLM cites this paper.

Guardians and Offenders: A Survey on Harmful Content Generation and Safety Mitigation of LLM Antidote: Post-fine-tuning Safety Alignment for Large Language Models against Harmful Fine-tuning

Reference 195

Resolution
unresolved
no resolver link, observed 2026-08-05T23:13:05.149386Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:13:05.149386Z digest=sha256:6f45cf65ba975f4e56a35be929239454abb1842261e1c312cb466aef8ba112b2

Observation 27f32768-a236-43d8-87c2-5e2116809be0 · inbound

TamperBench: Systematically Stress-Testing LLM Safety Under Fine-Tuning and Tampering cites this paper.

TamperBench: Systematically Stress-Testing LLM Safety Under Fine-Tuning and Tampering Antidote: Post-fine-tuning Safety Alignment for Large Language Models against Harmful Fine-tuning

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-03T03:46:14.818615Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:46:14.818615Z digest=sha256:789d7dfc88bba59ef27cf006cc2ed03322964a605cead2860cef36ba5b5ceb03

Observation 0c3566bd-0122-4ae0-8728-8856981c14fa · inbound

Preventing Safety Drift in Large Language Models via Coupled Weight and Activation Constraints cites this paper.

Preventing Safety Drift in Large Language Models via Coupled Weight and Activation Constraints Antidote: Post-fine-tuning Safety Alignment for Large Language Models against Harmful Fine-tuning

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-11T09:21:01.762843Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-10T16:04:25.851592Z digest=sha256:3759ba9a322af9d40342f414b072760dc9e92b4144194aed05b640a0246fd31f

Observation c2d39d3c-2be1-4d64-98b2-42a633e1f123 · inbound

Generating Place-Based Compromises Between Two Points of View cites this paper.

Generating Place-Based Compromises Between Two Points of View Antidote: Post-fine-tuning Safety Alignment for Large Language Models against Harmful Fine-tuning

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-05-11T22:01:12.267420Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-08T03:36:31.695964Z digest=sha256:acb7caf8cc1ec2682aff1cf0b514a95cbbae91bb510acd480ccb750a925ea653

Observation 1c0d34fb-72c6-4a6e-8509-cc129c2e3f80 · inbound

GradShield: Alignment Preserving Finetuning cites this paper.

GradShield: Alignment Preserving Finetuning Antidote: Post-fine-tuning Safety Alignment for Large Language Models against Harmful Fine-tuning

Reference 24

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T04:45:00.959333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-15T04:44:06.614390Z digest=sha256:5b0239a24a853710594c458706ae77364c48f4ea8b67d82095920ed3a28df975

Observation 4eaea4bb-852b-4bed-a106-65420c0fad1f · inbound

SafeGene: Reusable Adapters for Transferable Safety Alignment cites this paper.

SafeGene: Reusable Adapters for Transferable Safety Alignment Antidote: Post-fine-tuning Safety Alignment for Large Language Models against Harmful Fine-tuning

Reference 6

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T03:36:29.093869Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-28T09:58:16.956281Z digest=sha256:735a7606dce447d8f075924647a1be878c56055b4891a55d9224ab475fcf7484

Observation a746fa85-63be-4e64-b540-0ad1733a8c34 · inbound

Defending Against Malicious Finetuning by Scaling Train-time Adversarial Attacks cites this paper.

Defending Against Malicious Finetuning by Scaling Train-time Adversarial Attacks Antidote: Post-fine-tuning Safety Alignment for Large Language Models against Harmful Fine-tuning

Reference 32

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T20:47:22.920004Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-27T20:10:25.375603Z digest=sha256:424aede41071ff514216e7125507fa24daad5ad0bbddc7d641a942b8b055f897