Pith. sign in

Paper Citation Record · LEDGER

SafeDecoding: Defending against Jailbreak Attacks via Safety-Aware Decoding

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 20 inbound Pith citation observations for arXiv:2402.08983.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2402.08983 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 20 of 20 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 20 of 20 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T12:25:30.688823Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T03:59:33.707048Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation cf1f22e1-580c-4e5a-ba21-6dc19222373c · inbound

Jailbreak Attacks and Defenses Against Large Language Models: A Survey cites this paper.

Jailbreak Attacks and Defenses Against Large Language Models: A Survey SafeDecoding: Defending against Jailbreak Attacks via Safety-Aware Decoding

Reference 102

Resolution
verified exact
arxiv_id, observed 2026-05-15T02:20:44.438218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T02:20:44.368219Z digest=sha256:24b93bdf48966052c5deb20658d71a2accf47972bf371f95c96eb0f4b144e595

Observation f0ea2f5a-db58-414d-b1fd-6642c58c4119 · inbound

JBShield: Defending Large Language Models from Jailbreak Attacks through Activated Concept Analysis and Manipulation cites this paper.

JBShield: Defending Large Language Models from Jailbreak Attacks through Activated Concept Analysis and Manipulation SafeDecoding: Defending against Jailbreak Attacks via Safety-Aware Decoding

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-08T12:25:30.688823Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T12:25:30.688823Z digest=sha256:0d542e002ae62df8be6d949a14b580318e83114aa74f5d8183a10c58a5aa326f

Observation 434dbf62-dba0-407d-a3db-c61502453e29 · inbound

The Hidden Dimensions of LLM Alignment: A Multi-Dimensional Analysis of Orthogonal Safety Directions cites this paper.

The Hidden Dimensions of LLM Alignment: A Multi-Dimensional Analysis of Orthogonal Safety Directions SafeDecoding: Defending against Jailbreak Attacks via Safety-Aware Decoding

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T23:08:02.387653Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T23:08:02.387653Z digest=sha256:a6f5579a732b1ab327499dccb29042de9750bde1beaab41b61befb21619dfd3e

Observation 81774b8d-8151-4458-9f5b-abf3727b8984 · inbound

GloSS over Toxicity: Understanding and Mitigating Toxicity in LLMs via Global Toxic Subspace cites this paper.

GloSS over Toxicity: Understanding and Mitigating Toxicity in LLMs via Global Toxic Subspace SafeDecoding: Defending against Jailbreak Attacks via Safety-Aware Decoding

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T15:42:33.939169Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:42:33.939169Z digest=sha256:ad02d690a6c1e95fae1d88a82912c389eea2532f6b0103439db3638dda8ee9ee

Observation 6063ade9-b576-467b-acb4-3cb2196d7f7e · inbound

Just Enough Shifts: Mitigating Over-Refusal in Aligned Language Models with Targeted Representation Fine-Tuning cites this paper.

Just Enough Shifts: Mitigating Over-Refusal in Aligned Language Models with Targeted Representation Fine-Tuning SafeDecoding: Defending against Jailbreak Attacks via Safety-Aware Decoding

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T19:57:44.540195Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:57:44.540195Z digest=sha256:729e517e8e8c76e0eed4733ce9fc041047c9bccf65bf7569b7f1f3fa9ccb23c0

Observation 639dec63-bc50-4249-b2cf-74d0d7b6919f · inbound

PUZZLED: Jailbreaking LLMs through Word-Based Puzzles cites this paper.

PUZZLED: Jailbreaking LLMs through Word-Based Puzzles SafeDecoding: Defending against Jailbreak Attacks via Safety-Aware Decoding

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T05:43:39.465663Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T05:43:39.465663Z digest=sha256:6f99aa784e6a130c953415d2266b4f0caa4560f5bc7791b8842b453b50e0b359

Observation 210db1dc-0fcf-4a74-b547-da981641cae6 · inbound

Beyond Surface-Level Detection: Towards Cognitive-Driven Defense Against Jailbreak Attacks via Meta-Operations Reasoning cites this paper.

Beyond Surface-Level Detection: Towards Cognitive-Driven Defense Against Jailbreak Attacks via Meta-Operations Reasoning SafeDecoding: Defending against Jailbreak Attacks via Safety-Aware Decoding

Reference 187

Resolution
unresolved
no resolver link, observed 2026-08-06T04:47:25.646303Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T04:47:25.646303Z digest=sha256:4a064ce5d151610c653925714aac1f5f74c35808179fc60a62266a5ebd865b98

Observation fa55204e-3bbd-4601-95a4-cfe519237c62 · inbound

ReasoningGuard: Safeguarding Large Reasoning Models with Inference-time Safety Aha Moments cites this paper.

ReasoningGuard: Safeguarding Large Reasoning Models with Inference-time Safety Aha Moments SafeDecoding: Defending against Jailbreak Attacks via Safety-Aware Decoding

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-19T01:02:54.727920Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-19T01:02:07.088724Z digest=sha256:d65a127b45b2531d2ae27aaca62ee21af9a275b0e78dc1e3dd97788a2abd9a86

Observation 0b0fb6d5-9e85-465a-88d3-191905b818b9 · inbound

A Survey on Training-free Alignment of Large Language Models cites this paper.

A Survey on Training-free Alignment of Large Language Models SafeDecoding: Defending against Jailbreak Attacks via Safety-Aware Decoding

Reference 96

Resolution
unresolved
no resolver link, observed 2026-08-05T21:18:49.555908Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:18:49.555908Z digest=sha256:62504c1c722d981a766141222269495db7fe6bbba4eaababe9deea865edbfad3

Observation 7619697b-7203-48f8-8ed1-fb63a1dada03 · inbound

A Comprehensive Survey on Trustworthiness in Reasoning with Large Language Models cites this paper.

A Comprehensive Survey on Trustworthiness in Reasoning with Large Language Models SafeDecoding: Defending against Jailbreak Attacks via Safety-Aware Decoding

Reference 267

Resolution
unresolved
no resolver link, observed 2026-08-05T10:39:08.221533Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:39:08.221533Z digest=sha256:eee22f8e2b1fd17e8f60aff18a15fabfb5764a4a25b9ee273270b88007c12b68

Observation 0986a88f-38f8-4ea1-9def-a33b45ba276e · inbound

MoGU V2: Toward a Higher Pareto Frontier Between Model Usability and Security cites this paper.

MoGU V2: Toward a Higher Pareto Frontier Between Model Usability and Security SafeDecoding: Defending against Jailbreak Attacks via Safety-Aware Decoding

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-04T23:09:41.478401Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T23:09:41.478401Z digest=sha256:6574c158d5c447e996e4af67725f1ff574d463bb08dfb1f95b9180a2832bb053

Observation a5d493be-74b1-4ce1-bc43-1ed95bf91700 · inbound

TrajGuard: Streaming Hidden-state Trajectory Detection for Decoding-time Jailbreak Defense cites this paper.

TrajGuard: Streaming Hidden-state Trajectory Detection for Decoding-time Jailbreak Defense SafeDecoding: Defending against Jailbreak Attacks via Safety-Aware Decoding

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-05-11T00:35:50.904105Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-10T18:26:19.922383Z digest=sha256:defcc2d59ba1f8cc30a6e8b7c756616c4b09ab394b616e74ad5077f1cb1b0319

Observation c920d2a6-be03-4670-8fdc-06c68e8b3dfd · inbound

Jailbreaking the Matrix: Nullspace Steering for Controlled Model Subversion cites this paper.

Jailbreaking the Matrix: Nullspace Steering for Controlled Model Subversion SafeDecoding: Defending against Jailbreak Attacks via Safety-Aware Decoding

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:46:04.241295Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-10T15:19:46.920899Z digest=sha256:ea27b7a088a9c9ce2d8a267f961a1c4e20c180932ce4a1070a6e06e54f4e0f52

Observation 4a9abf7e-acab-4d20-9778-4bb2a44b8bd7 · inbound

Jailbreaking Frontier Foundation Models Through Intention Deception cites this paper.

Jailbreaking Frontier Foundation Models Through Intention Deception SafeDecoding: Defending against Jailbreak Attacks via Safety-Aware Decoding

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-11T22:11:14.065967Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-08T03:17:51.039062Z digest=sha256:d7ba350411c506525dda64c4b66c71350dad16a8453d5c4b5d8dd8bd03653441

Observation 50551200-c250-4ecb-ab89-7297a7cf4dcc · inbound

Minimal, Local, Causal Explanations for Jailbreak Success in Large Language Models cites this paper.

Minimal, Local, Causal Explanations for Jailbreak Success in Large Language Models SafeDecoding: Defending against Jailbreak Attacks via Safety-Aware Decoding

Reference 16

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T15:01:09.635553Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-09T20:39:39.225898Z digest=sha256:5130aca8cd9fa4dd9809d6102b81bff13df4ec48f6b9901b7f1136755f8b54ce

Observation 3efa7aa7-9924-440f-9f25-3af5375d4347 · inbound

SafeSteer: A Decoding-level Defense Mechanism for Multimodal Large Language Models cites this paper.

SafeSteer: A Decoding-level Defense Mechanism for Multimodal Large Language Models SafeDecoding: Defending against Jailbreak Attacks via Safety-Aware Decoding

Reference 11

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T06:57:27.549492Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-13T06:56:10.053418Z digest=sha256:95454abd8d95311d82f6da300fbef2a0846eb073bbcfa0bc076d18b0fbd743da

Observation bd96da19-f6e6-4fbf-8aad-ec088501005c · inbound

NeuroArmor: Safe-Variant-Guided Representation Consistency for Selective Re-Anchoring in Jailbreak Defense cites this paper.

NeuroArmor: Safe-Variant-Guided Representation Consistency for Selective Re-Anchoring in Jailbreak Defense SafeDecoding: Defending against Jailbreak Attacks via Safety-Aware Decoding

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T03:46:32.739091Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-28T09:42:58.615078Z digest=sha256:55b1da87ac5cdb5ff717e9ec2fc73a0f7fbbc857d2e1caf81d82977159b97466

Observation e7372632-2dd5-4540-8213-2e1e0ac76ba5 · inbound

SlotGCG: Exploiting the Positional Vulnerability in LLMs for Jailbreak Attacks cites this paper.

SlotGCG: Exploiting the Positional Vulnerability in LLMs for Jailbreak Attacks SafeDecoding: Defending against Jailbreak Attacks via Safety-Aware Decoding

Reference 62

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T13:26:59.257417Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-28T01:16:07.252429Z digest=sha256:35e994ea491305c688dbc5ccc6ca9dd46dea49628701ea5de04acb9d271de52e

Observation a385f6c9-16d1-429e-ae3e-d9e873aa672b · inbound

SafeSpec: Fast and Safe LLM via Dynamic Reflective Sampling cites this paper.

SafeSpec: Fast and Safe LLM via Dynamic Reflective Sampling SafeDecoding: Defending against Jailbreak Attacks via Safety-Aware Decoding

Reference 13

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T03:59:33.709134Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-26T17:23:06.382761Z digest=sha256:ed08fe00dd47af2041f5fd86a7bdc7d81927ab7d559dde5230ba967c56d08db3

Observation 1596d35c-477a-491b-9715-f47be97690bd · inbound

Robust Harmful Features Under Jailbreak Attacks: Mechanistic Evidence from Attention Head Specialization in Large Language Models cites this paper.

Robust Harmful Features Under Jailbreak Attacks: Mechanistic Evidence from Attention Head Specialization in Large Language Models SafeDecoding: Defending against Jailbreak Attacks via Safety-Aware Decoding

Reference 41

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T17:35:51.363638Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-29T03:35:34.594617Z digest=sha256:e427844dadf32ad2d78514082fedbbe6a074711e62314287de0c4e7de00b946b