Pith. sign in

Paper Citation Record · LEDGER

Toward Secure Tuning: Mitigating Security Risks from Instruction Fine-Tuning

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 7 inbound Pith citation observations for arXiv:2410.04524.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2410.04524 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 7 of 7 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 7 of 7 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:02:33.690065Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 58f9adb2-0685-432f-b6b5-f01d0a117206 · inbound

Harmful Fine-tuning Attacks and Defenses for Large Language Models: A Survey cites this paper.

Harmful Fine-tuning Attacks and Defenses for Large Language Models: A Survey Toward Secure Tuning: Mitigating Security Risks from Instruction Fine-Tuning

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-23T20:58:26.039828Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-23T20:58:16.237327Z digest=sha256:f18d899ef4c0c1862d799a32dc5a9e1645982a13664153512143958ed9180622

Observation 945327e5-42c1-4dcd-9ef0-14fb7d36d7b5 · inbound

CTRAP: Embedding Collapse Trap to Safeguard Large Language Models from Harmful Fine-Tuning cites this paper.

CTRAP: Embedding Collapse Trap to Safeguard Large Language Models from Harmful Fine-Tuning Toward Secure Tuning: Mitigating Security Risks from Instruction Fine-Tuning

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-07T15:02:33.690065Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:02:33.690065Z digest=sha256:9f92ab387ada293cb01ce0874f8051fc12910466dd2627ffa5e80de0b44d8e58

Observation 4c0a44ed-ae13-4cf9-9eda-91c3020efaa5 · inbound

Reshaping Representation Space to Balance the Safety and Over-rejection in Large Audio Language Models cites this paper.

Reshaping Representation Space to Balance the Safety and Over-rejection in Large Audio Language Models Toward Secure Tuning: Mitigating Security Risks from Instruction Fine-Tuning

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T14:15:18.441128Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:15:18.441128Z digest=sha256:af9143209bc29866e29252bd1708afe49d6d611435ee2d768529557ae4de78bd

Observation 0a575274-5666-4764-9913-0df741e67ef7 · inbound

MoGU V2: Toward a Higher Pareto Frontier Between Model Usability and Security cites this paper.

MoGU V2: Toward a Higher Pareto Frontier Between Model Usability and Security Toward Secure Tuning: Mitigating Security Risks from Instruction Fine-Tuning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-04T23:09:41.463474Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T23:09:41.463474Z digest=sha256:fb265907d2f93dedcb81c67abcd53106cf2ef2049baf036828c9088fbf8ed208

Observation 5924b097-09e9-462d-8eb5-a4d38184c206 · inbound

Locate, Steer, and Improve: A Practical Survey of Actionable Mechanistic Interpretability in Large Language Models cites this paper.

Locate, Steer, and Improve: A Practical Survey of Actionable Mechanistic Interpretability in Large Language Models Toward Secure Tuning: Mitigating Security Risks from Instruction Fine-Tuning

Reference 62

Resolution
verified exact
arxiv_id, observed 2026-05-16T12:40:54.418934Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-16T12:39:57.398423Z digest=sha256:e315c4ca89f05b59980f81843520390a5591edbd346d238f68a85b0afc44b3dc

Observation cff9a125-00ba-468c-82d0-26bad94a0713 · inbound

TamperBench: Systematically Stress-Testing LLM Safety Under Fine-Tuning and Tampering cites this paper.

TamperBench: Systematically Stress-Testing LLM Safety Under Fine-Tuning and Tampering Toward Secure Tuning: Mitigating Security Risks from Instruction Fine-Tuning

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-03T03:46:14.790420Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:46:14.790420Z digest=sha256:79e4c52d65e4d21565fe8591509e9de38846fade6e7300dd5973989823032b09

Observation 81ef9047-4a2d-4c48-b365-795bf1d7e40b · inbound

SafeSteer: Localized On-Policy Distillation for Efficient Safety Alignment cites this paper.

SafeSteer: Localized On-Policy Distillation for Efficient Safety Alignment Toward Secure Tuning: Mitigating Security Risks from Instruction Fine-Tuning

Reference 32

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T23:06:20.908557Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-28T14:39:11.178976Z digest=sha256:0ca0bb9dc447cd4cbfd9263c8fd0fd1da68e51b2fee0539c1a67b3c4559adeae