Pith. sign in

Paper Citation Record · LEDGER

Representation Noising: A Defence Mechanism Against Harmful Finetuning

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 12 inbound Pith citation observations for arXiv:2405.14577.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2405.14577 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 12 of 12 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 12 of 12 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-09T14:47:15.278636Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T10:59:47.035398Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation f4b00d41-4523-4329-ba55-459af03b4950 · inbound

Harmful Fine-tuning Attacks and Defenses for Large Language Models: A Survey cites this paper.

Harmful Fine-tuning Attacks and Defenses for Large Language Models: A Survey Representation Noising: A Defence Mechanism Against Harmful Finetuning

Reference 128

Resolution
verified exact
arxiv_id, observed 2026-05-23T20:58:26.090038Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-23T20:58:16.237327Z digest=sha256:2d2c8f2241c4e85f43bc5d39d91c063b973c6221ce60d4244e2f5ed2fe310018

Observation d1c44e37-5f68-4494-8618-560ad26d944d · inbound

Model Tampering Attacks Enable More Rigorous Evaluations of LLM Capabilities cites this paper.

Model Tampering Attacks Enable More Rigorous Evaluations of LLM Capabilities Representation Noising: A Defence Mechanism Against Harmful Finetuning

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-09T14:47:15.278636Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T14:47:15.278636Z digest=sha256:358c00e7ab3e9f5e9fcb47c13b2a242bac0552fb4e9b9297b8e9e7a467912a38

Observation 8c0bdbcb-74c2-45ee-80d8-7e72561f0d4f · inbound

Beyond External Monitors: Enhancing Transparency of Large Language Models for Easier Monitoring cites this paper.

Beyond External Monitors: Enhancing Transparency of Large Language Models for Easier Monitoring Representation Noising: A Defence Mechanism Against Harmful Finetuning

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-08T21:06:56.026593Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T21:06:56.026593Z digest=sha256:c5d33e8fdf955d81f986a383f9f624144faa387b87253a7444d3cd5d821741a6

Observation a6dcba4f-cdc1-476c-90c6-c5f3de018471 · inbound

Jailbreaking to Jailbreak cites this paper.

Jailbreaking to Jailbreak Representation Noising: A Defence Mechanism Against Harmful Finetuning

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-08T17:02:47.412426Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:02:47.412426Z digest=sha256:e9e3f436cc77ba92a10b612367d16c244d3eae0247b860e86653b5b474023e5d

Observation 6bfb566f-c754-4fbc-86e7-f52ceb0ed5e2 · inbound

Reshaping Representation Space to Balance the Safety and Over-rejection in Large Audio Language Models cites this paper.

Reshaping Representation Space to Balance the Safety and Over-rejection in Large Audio Language Models Representation Noising: A Defence Mechanism Against Harmful Finetuning

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T14:15:20.532920Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:15:20.532920Z digest=sha256:f146b3b231dceefb6df8ff9b364e921562ae95ede084eafc7175735b5fd69b98

Observation e024dd50-b677-4397-9582-d93faf791fef · inbound

Vulnerability-Aware Alignment: Mitigating Uneven Forgetting in Harmful Fine-Tuning cites this paper.

Vulnerability-Aware Alignment: Mitigating Uneven Forgetting in Harmful Fine-Tuning Representation Noising: A Defence Mechanism Against Harmful Finetuning

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T10:58:31.656919Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:58:31.656919Z digest=sha256:c41a672f81d90742101383ceda98bd637febeb77e1cd3dca419ffebd7b34e9c6

Observation f4a944b9-7dd5-4bb7-9f03-2070e7a327f3 · inbound

FORTRESS: Frontier Risk Evaluation for National Security and Public Safety cites this paper.

FORTRESS: Frontier Risk Evaluation for National Security and Public Safety Representation Noising: A Defence Mechanism Against Harmful Finetuning

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:03.070633Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:03.070633Z digest=sha256:91f779508b47acfbe23fb2e7acdc935ffdb187335d00c6be9f6fe71a757b7252

Observation a6b8f4de-7b43-4ce2-b7a4-f699d257bc52 · inbound

Learning to Stay Safe: Adaptive Regularization Against Safety Degradation during Fine-Tuning cites this paper.

Learning to Stay Safe: Adaptive Regularization Against Safety Degradation during Fine-Tuning Representation Noising: A Defence Mechanism Against Harmful Finetuning

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-15T20:56:37.290144Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-15T20:51:57.399260Z digest=sha256:8b01619547b7c873255c86fde5d1d4905c280fdd32845f35840035fd1736b6c1

Observation ef32365d-7848-40d2-85c3-344861074321 · inbound

Immunizing 3D Gaussian Generative Models Against Unauthorized Fine-Tuning via Attribute-Space Traps cites this paper.

Immunizing 3D Gaussian Generative Models Against Unauthorized Fine-Tuning via Attribute-Space Traps Representation Noising: A Defence Mechanism Against Harmful Finetuning

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-10T22:55:48.920053Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T19:30:06.396482Z digest=sha256:2e137db1a265c3fa21f25576c77a7160acc95fffbe9e7ef949e79aad1b83bb31

Observation d3769af6-9555-47b9-a9e3-0d98f9e839fd · inbound

Continual Safety Alignment via Gradient-Based Sample Selection cites this paper.

Continual Safety Alignment via Gradient-Based Sample Selection Representation Noising: A Defence Mechanism Against Harmful Finetuning

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T07:16:54.733024Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T07:16:53.472918Z digest=sha256:e0d50110ea5fb0edd177a4417dd2452e21aded5d35eec28fc75e87d23f56e476

Observation cff556fb-b075-47fc-a9bc-31fa87ab32f3 · inbound

Safety in Self-Evolving LLM Agent Systems: Threats, Amplification, and Case Studies cites this paper.

Safety in Self-Evolving LLM Agent Systems: Threats, Amplification, and Case Studies Representation Noising: A Defence Mechanism Against Harmful Finetuning

Reference 61

Resolution
verified exact
arxiv_id, observed 2026-07-04T10:59:47.037148Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-26T08:11:30.091859Z digest=sha256:948097705c74f01c5ebcf99ef13e0c527efb18b5a367e2701d3ae7fff72a3fe7

Observation 54a12f0b-55c1-45d7-a659-426c261aec60 · inbound

Gradient Immunity: Null-Space Resistance to Malicious Fine-Tuning cites this paper.

Gradient Immunity: Null-Space Resistance to Malicious Fine-Tuning Representation Noising: A Defence Mechanism Against Harmful Finetuning

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-06T10:46:10.792708Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:46:10.792708Z digest=sha256:8e4e13b442f566f4e04482833b9cd567f6b03a9d358577bc2209ff5d6f4b3427