Pith. sign in

Paper Citation Record · LEDGER

Toward Secure Tuning: Mitigating Security Risks from Instruction Fine-Tuning

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 7 inbound Pith citation observations for arXiv:2410.04524.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2410.04524 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 7 of 7 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 7 of 7 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:02:33.690065Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 58f9adb2-0685-432f-b6b5-f01d0a117206 · inbound

Harmful Fine-tuning Attacks and Defenses for Large Language Models: A Survey cites this paper.

Harmful Fine-tuning Attacks and Defenses for Large Language Models: A Survey Toward Secure Tuning: Mitigating Security Risks from Instruction Fine-Tuning

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-23T20:58:26.039828Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-23T20:58:16.237327Z digest=sha256:091c1fb6ef531566fa37c60c8108ff1125ea3dae021cd776291b0c34cc58d6f0

Observation 945327e5-42c1-4dcd-9ef0-14fb7d36d7b5 · inbound

CTRAP: Embedding Collapse Trap to Safeguard Large Language Models from Harmful Fine-Tuning cites this paper.

CTRAP: Embedding Collapse Trap to Safeguard Large Language Models from Harmful Fine-Tuning Toward Secure Tuning: Mitigating Security Risks from Instruction Fine-Tuning

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-07T15:02:33.690065Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:02:33.690065Z digest=sha256:9f92ab387ada293cb01ce0874f8051fc12910466dd2627ffa5e80de0b44d8e58

Observation 4c0a44ed-ae13-4cf9-9eda-91c3020efaa5 · inbound

Reshaping Representation Space to Balance the Safety and Over-rejection in Large Audio Language Models cites this paper.

Reshaping Representation Space to Balance the Safety and Over-rejection in Large Audio Language Models Toward Secure Tuning: Mitigating Security Risks from Instruction Fine-Tuning

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T14:15:18.441128Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:15:18.441128Z digest=sha256:b756195d6f6fdf5de23296dde45624bf49006ed6a0eb3da0f8a2ec2db325a39e

Observation 0a575274-5666-4764-9913-0df741e67ef7 · inbound

MoGU V2: Toward a Higher Pareto Frontier Between Model Usability and Security cites this paper.

MoGU V2: Toward a Higher Pareto Frontier Between Model Usability and Security Toward Secure Tuning: Mitigating Security Risks from Instruction Fine-Tuning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-04T23:09:41.463474Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T23:09:41.463474Z digest=sha256:432b5a7e015d7d4aaa89afdb33920b64d37c3db3602c03143e3a4ad795e17e68

Observation 5924b097-09e9-462d-8eb5-a4d38184c206 · inbound

Locate, Steer, and Improve: A Practical Survey of Actionable Mechanistic Interpretability in Large Language Models cites this paper.

Locate, Steer, and Improve: A Practical Survey of Actionable Mechanistic Interpretability in Large Language Models Toward Secure Tuning: Mitigating Security Risks from Instruction Fine-Tuning

Reference 62

Resolution
verified exact
arxiv_id, observed 2026-05-16T12:40:54.418934Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-16T12:39:57.398423Z digest=sha256:475aaf0dfeb7f3e6dcea53c2cc8cf34963ad8107b8f0016dc6452b1dbb156ce2

Observation cff9a125-00ba-468c-82d0-26bad94a0713 · inbound

TamperBench: Systematically Stress-Testing LLM Safety Under Fine-Tuning and Tampering cites this paper.

TamperBench: Systematically Stress-Testing LLM Safety Under Fine-Tuning and Tampering Toward Secure Tuning: Mitigating Security Risks from Instruction Fine-Tuning

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-03T03:46:14.790420Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:46:14.790420Z digest=sha256:2e15de87927df7535c7ab51edcf1e1105511f46e2f37fe086633657bdbfe163e

Observation 81ef9047-4a2d-4c48-b365-795bf1d7e40b · inbound

SafeSteer: Localized On-Policy Distillation for Efficient Safety Alignment cites this paper.

SafeSteer: Localized On-Policy Distillation for Efficient Safety Alignment Toward Secure Tuning: Mitigating Security Risks from Instruction Fine-Tuning

Reference 32

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T23:06:20.908557Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-28T14:39:11.178976Z digest=sha256:4426296769137deb431d632eb53a100daeb7fd80f089715ac99e89d9db8d5bf6