Pith. sign in

Paper Citation Record · LEDGER

Refuse Whenever You Feel Unsafe: Improving Safety in LLMs via Decoupled Refusal Training

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 11 inbound Pith citation observations for arXiv:2407.09121.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2407.09121 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 11 of 11 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 11 of 11 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-09T16:18:40.974536Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-22T14:41:41.501726Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 84397eef-1dc0-4725-8c09-8c0ba4f837e6 · inbound

On Almost Surely Safe Alignment of Large Language Models at Inference-Time cites this paper.

On Almost Surely Safe Alignment of Large Language Models at Inference-Time Refuse Whenever You Feel Unsafe: Improving Safety in LLMs via Decoupled Refusal Training

Reference 102

Resolution
unresolved
no resolver link, observed 2026-08-09T16:18:40.974536Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T16:18:40.974536Z digest=sha256:9c662c71390fb977d2ca25a55e788c7b812abebf0d2e091751cc17f87ed24596

Observation ea402bf6-fec3-45fc-9897-95df0e627f81 · inbound

Safety Reasoning with Guidelines cites this paper.

Safety Reasoning with Guidelines Refuse Whenever You Feel Unsafe: Improving Safety in LLMs via Decoupled Refusal Training

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-08T23:50:36.272301Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T23:50:36.272301Z digest=sha256:d77978b4b4b56d5223b8219d11fd1f547aea5c7e42f7f888a0235857e94e9735

Observation aefc68dd-f5a7-4218-8fe3-77f7c6f34395 · inbound

Model Tampering Attacks Enable More Rigorous Evaluations of LLM Capabilities cites this paper.

Model Tampering Attacks Enable More Rigorous Evaluations of LLM Capabilities Refuse Whenever You Feel Unsafe: Improving Safety in LLMs via Decoupled Refusal Training

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-09T14:47:15.397309Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T14:47:15.397309Z digest=sha256:6d91927fddb31d5750127b50a136388e31996f2035b2a197934189bd746d8c9e

Observation ca9cbe33-01aa-45cb-b919-82b450a60b18 · inbound

Jailbreaking to Jailbreak cites this paper.

Jailbreaking to Jailbreak Refuse Whenever You Feel Unsafe: Improving Safety in LLMs via Decoupled Refusal Training

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-08T17:02:47.514685Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:02:47.514685Z digest=sha256:b61496778e761f5fd00656c28e19100beae40359d39e9369e9a08997f4bbe137

Observation 642ec696-35a0-4a90-9224-1ed5881a722e · inbound

Trustworthy AI: Safety, Bias, and Privacy -- A Survey cites this paper.

Trustworthy AI: Safety, Bias, and Privacy -- A Survey Refuse Whenever You Feel Unsafe: Improving Safety in LLMs via Decoupled Refusal Training

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-08T11:26:27.914775Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T11:26:27.914775Z digest=sha256:a99c924b8adfd32a95c8843442a22b4b0edaa04dc6727ed4a2241b9fd6b2b785

Observation 22d5e7c4-e50c-456b-982f-8600d93c2a86 · inbound

Phonetic Perturbations Reveal Tokenizer-Rooted Safety Gaps in LLMs cites this paper.

Phonetic Perturbations Reveal Tokenizer-Rooted Safety Gaps in LLMs Refuse Whenever You Feel Unsafe: Improving Safety in LLMs via Decoupled Refusal Training

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-22T14:41:41.504894Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-22T14:40:58.506345Z digest=sha256:8a5988d4d772953fdfac78ff6f6e34bc52004f460e7592a2aa56ae826742734b

Observation e5ba28bf-bc90-443c-b2b9-76d055c25e1f · inbound

Lifelong Safety Alignment for Language Models cites this paper.

Lifelong Safety Alignment for Language Models Refuse Whenever You Feel Unsafe: Improving Safety in LLMs via Decoupled Refusal Training

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-07T14:00:09.569239Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:00:09.569239Z digest=sha256:6e5315c3cc7ece00b88315999b52794c8db00f26b33fba95cf9c79c2d74bc996

Observation 24f406f7-dac3-4aae-af34-04176296cb90 · inbound

Align is not Enough: Multimodal Universal Jailbreak Attack against Multimodal Large Language Models cites this paper.

Align is not Enough: Multimodal Universal Jailbreak Attack against Multimodal Large Language Models Refuse Whenever You Feel Unsafe: Improving Safety in LLMs via Decoupled Refusal Training

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T11:51:33.126948Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:51:33.126948Z digest=sha256:a1ad3eefa0f8646ecafdaf1db3c92f676a9392e093bd85056cd6e4099a0f66aa

Observation c8c4074d-5f65-4a4b-b144-c47671818d02 · inbound

A Red Teaming Roadmap Towards System-Level Safety cites this paper.

A Red Teaming Roadmap Towards System-Level Safety Refuse Whenever You Feel Unsafe: Improving Safety in LLMs via Decoupled Refusal Training

Reference 98

Resolution
unresolved
no resolver link, observed 2026-08-07T12:11:25.746161Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:11:25.746161Z digest=sha256:cc082925c55b826907880a8e27ccab9c90c6a88907ca75d2f71d4cb83c7682f2

Observation 168523f8-e758-40be-8b91-fce8b89932d7 · inbound

FORTRESS: Frontier Risk Evaluation for National Security and Public Safety cites this paper.

FORTRESS: Frontier Risk Evaluation for National Security and Public Safety Refuse Whenever You Feel Unsafe: Improving Safety in LLMs via Decoupled Refusal Training

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:05.291677Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:05.291677Z digest=sha256:84e5b0ec58023fd44ffad20d12fa2905ee048402bdc79a63e5bd487300ae6f09

Observation 82d4339c-c82e-4811-b377-3721eb6781f0 · inbound

FreakOut-LLM: The Effect of Emotional Stimuli on Safety Alignment cites this paper.

FreakOut-LLM: The Effect of Emotional Stimuli on Safety Alignment Refuse Whenever You Feel Unsafe: Improving Safety in LLMs via Decoupled Refusal Training

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T16:52:59.995770Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-13T16:51:46.603418Z digest=sha256:9f880d0aae21e7dd83e39c21c3c06cf9c535e57859407a4c269c4caac30a5621