Pith. sign in

Paper Citation Record · LEDGER

SafeDecoding: Defending against Jailbreak Attacks via Safety-Aware Decoding

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 20 inbound Pith citation observations for arXiv:2402.08983.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2402.08983 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 20 of 20 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 20 of 20 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T12:25:30.688823Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T03:59:33.707048Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation cf1f22e1-580c-4e5a-ba21-6dc19222373c · inbound

Jailbreak Attacks and Defenses Against Large Language Models: A Survey cites this paper.

Jailbreak Attacks and Defenses Against Large Language Models: A Survey SafeDecoding: Defending against Jailbreak Attacks via Safety-Aware Decoding

Reference 102

Resolution
verified exact
arxiv_id, observed 2026-05-15T02:20:44.438218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T02:20:44.368219Z digest=sha256:2ed732dc1a3b91d81ac8bfdfafd741a803205e592be570b7bf4b5d5ef142c8e5

Observation f0ea2f5a-db58-414d-b1fd-6642c58c4119 · inbound

JBShield: Defending Large Language Models from Jailbreak Attacks through Activated Concept Analysis and Manipulation cites this paper.

JBShield: Defending Large Language Models from Jailbreak Attacks through Activated Concept Analysis and Manipulation SafeDecoding: Defending against Jailbreak Attacks via Safety-Aware Decoding

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-08T12:25:30.688823Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T12:25:30.688823Z digest=sha256:70d9a90c6f7bc93c7ad8330abf7ca47a570d8b2fac9733cb872396857241cc52

Observation 434dbf62-dba0-407d-a3db-c61502453e29 · inbound

The Hidden Dimensions of LLM Alignment: A Multi-Dimensional Analysis of Orthogonal Safety Directions cites this paper.

The Hidden Dimensions of LLM Alignment: A Multi-Dimensional Analysis of Orthogonal Safety Directions SafeDecoding: Defending against Jailbreak Attacks via Safety-Aware Decoding

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T23:08:02.387653Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T23:08:02.387653Z digest=sha256:7fd3ac08f195646c2b9fe6910d961172d1ccd115b73019797cd0dfcd82e20e67

Observation 81774b8d-8151-4458-9f5b-abf3727b8984 · inbound

GloSS over Toxicity: Understanding and Mitigating Toxicity in LLMs via Global Toxic Subspace cites this paper.

GloSS over Toxicity: Understanding and Mitigating Toxicity in LLMs via Global Toxic Subspace SafeDecoding: Defending against Jailbreak Attacks via Safety-Aware Decoding

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T15:42:33.939169Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:42:33.939169Z digest=sha256:ad02d690a6c1e95fae1d88a82912c389eea2532f6b0103439db3638dda8ee9ee

Observation 6063ade9-b576-467b-acb4-3cb2196d7f7e · inbound

Just Enough Shifts: Mitigating Over-Refusal in Aligned Language Models with Targeted Representation Fine-Tuning cites this paper.

Just Enough Shifts: Mitigating Over-Refusal in Aligned Language Models with Targeted Representation Fine-Tuning SafeDecoding: Defending against Jailbreak Attacks via Safety-Aware Decoding

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T19:57:44.540195Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:57:44.540195Z digest=sha256:630b5d07eab520fcff720a29220c5b04a1a5282e270acbb3707f8f4d42afe32d

Observation 639dec63-bc50-4249-b2cf-74d0d7b6919f · inbound

PUZZLED: Jailbreaking LLMs through Word-Based Puzzles cites this paper.

PUZZLED: Jailbreaking LLMs through Word-Based Puzzles SafeDecoding: Defending against Jailbreak Attacks via Safety-Aware Decoding

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T05:43:39.465663Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T05:43:39.465663Z digest=sha256:61a6979ecbe7816946c9113e8ff68cb9dd01c866bfb1f2a779770380bead9ad3

Observation 210db1dc-0fcf-4a74-b547-da981641cae6 · inbound

Beyond Surface-Level Detection: Towards Cognitive-Driven Defense Against Jailbreak Attacks via Meta-Operations Reasoning cites this paper.

Beyond Surface-Level Detection: Towards Cognitive-Driven Defense Against Jailbreak Attacks via Meta-Operations Reasoning SafeDecoding: Defending against Jailbreak Attacks via Safety-Aware Decoding

Reference 187

Resolution
unresolved
no resolver link, observed 2026-08-06T04:47:25.646303Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T04:47:25.646303Z digest=sha256:4a064ce5d151610c653925714aac1f5f74c35808179fc60a62266a5ebd865b98

Observation fa55204e-3bbd-4601-95a4-cfe519237c62 · inbound

ReasoningGuard: Safeguarding Large Reasoning Models with Inference-time Safety Aha Moments cites this paper.

ReasoningGuard: Safeguarding Large Reasoning Models with Inference-time Safety Aha Moments SafeDecoding: Defending against Jailbreak Attacks via Safety-Aware Decoding

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-19T01:02:54.727920Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-19T01:02:07.088724Z digest=sha256:a25cd6c80cf5248e575bfd17485c027d9a8a33e5a3ec6063a292cd0ec7039af3

Observation 0b0fb6d5-9e85-465a-88d3-191905b818b9 · inbound

A Survey on Training-free Alignment of Large Language Models cites this paper.

A Survey on Training-free Alignment of Large Language Models SafeDecoding: Defending against Jailbreak Attacks via Safety-Aware Decoding

Reference 96

Resolution
unresolved
no resolver link, observed 2026-08-05T21:18:49.555908Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:18:49.555908Z digest=sha256:209df377a64108265e0b82e7a3400752c211987ff4d487168855d6610f7d1a60

Observation 7619697b-7203-48f8-8ed1-fb63a1dada03 · inbound

A Comprehensive Survey on Trustworthiness in Reasoning with Large Language Models cites this paper.

A Comprehensive Survey on Trustworthiness in Reasoning with Large Language Models SafeDecoding: Defending against Jailbreak Attacks via Safety-Aware Decoding

Reference 267

Resolution
unresolved
no resolver link, observed 2026-08-05T10:39:08.221533Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:39:08.221533Z digest=sha256:eee22f8e2b1fd17e8f60aff18a15fabfb5764a4a25b9ee273270b88007c12b68

Observation 0986a88f-38f8-4ea1-9def-a33b45ba276e · inbound

MoGU V2: Toward a Higher Pareto Frontier Between Model Usability and Security cites this paper.

MoGU V2: Toward a Higher Pareto Frontier Between Model Usability and Security SafeDecoding: Defending against Jailbreak Attacks via Safety-Aware Decoding

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-04T23:09:41.478401Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T23:09:41.478401Z digest=sha256:6574c158d5c447e996e4af67725f1ff574d463bb08dfb1f95b9180a2832bb053

Observation a5d493be-74b1-4ce1-bc43-1ed95bf91700 · inbound

TrajGuard: Streaming Hidden-state Trajectory Detection for Decoding-time Jailbreak Defense cites this paper.

TrajGuard: Streaming Hidden-state Trajectory Detection for Decoding-time Jailbreak Defense SafeDecoding: Defending against Jailbreak Attacks via Safety-Aware Decoding

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-05-11T00:35:50.904105Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-10T18:26:19.922383Z digest=sha256:788f16b1ccfcc24b514d5e58ee8a81d8f33af355660313404e35a322887af111

Observation c920d2a6-be03-4670-8fdc-06c68e8b3dfd · inbound

Jailbreaking the Matrix: Nullspace Steering for Controlled Model Subversion cites this paper.

Jailbreaking the Matrix: Nullspace Steering for Controlled Model Subversion SafeDecoding: Defending against Jailbreak Attacks via Safety-Aware Decoding

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:46:04.241295Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-10T15:19:46.920899Z digest=sha256:29a6932c8230a644f5910abd56ba9e1c455ee4e0d7a6795a40d6a09ab50ffec7

Observation 4a9abf7e-acab-4d20-9778-4bb2a44b8bd7 · inbound

Jailbreaking Frontier Foundation Models Through Intention Deception cites this paper.

Jailbreaking Frontier Foundation Models Through Intention Deception SafeDecoding: Defending against Jailbreak Attacks via Safety-Aware Decoding

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-11T22:11:14.065967Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-08T03:17:51.039062Z digest=sha256:b413d0546e3ec1405977c12b580ebcbda0b194793f4633290bb5bbb354de68fe

Observation 50551200-c250-4ecb-ab89-7297a7cf4dcc · inbound

Minimal, Local, Causal Explanations for Jailbreak Success in Large Language Models cites this paper.

Minimal, Local, Causal Explanations for Jailbreak Success in Large Language Models SafeDecoding: Defending against Jailbreak Attacks via Safety-Aware Decoding

Reference 16

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T15:01:09.635553Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-09T20:39:39.225898Z digest=sha256:e8799dbce5498b80866ad85251d1a1644a9c972a159b1f1980f391662c2bcb17

Observation 3efa7aa7-9924-440f-9f25-3af5375d4347 · inbound

SafeSteer: A Decoding-level Defense Mechanism for Multimodal Large Language Models cites this paper.

SafeSteer: A Decoding-level Defense Mechanism for Multimodal Large Language Models SafeDecoding: Defending against Jailbreak Attacks via Safety-Aware Decoding

Reference 11

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T06:57:27.549492Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-13T06:56:10.053418Z digest=sha256:669c05bdae583c54da78d166dc08710ca9b21efedbba47e9181e871ffc4720d1

Observation bd96da19-f6e6-4fbf-8aad-ec088501005c · inbound

NeuroArmor: Safe-Variant-Guided Representation Consistency for Selective Re-Anchoring in Jailbreak Defense cites this paper.

NeuroArmor: Safe-Variant-Guided Representation Consistency for Selective Re-Anchoring in Jailbreak Defense SafeDecoding: Defending against Jailbreak Attacks via Safety-Aware Decoding

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T03:46:32.739091Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-28T09:42:58.615078Z digest=sha256:9eb750f0f048a9692a03f2478cca2a8eca3c27ae5c4991f0ae532f6785200946

Observation e7372632-2dd5-4540-8213-2e1e0ac76ba5 · inbound

SlotGCG: Exploiting the Positional Vulnerability in LLMs for Jailbreak Attacks cites this paper.

SlotGCG: Exploiting the Positional Vulnerability in LLMs for Jailbreak Attacks SafeDecoding: Defending against Jailbreak Attacks via Safety-Aware Decoding

Reference 62

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T13:26:59.257417Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-28T01:16:07.252429Z digest=sha256:356d311093d7a417373bfb131c02b244ba8fa9fc80664caba90b7f1ace8d3faf

Observation a385f6c9-16d1-429e-ae3e-d9e873aa672b · inbound

SafeSpec: Fast and Safe LLM via Dynamic Reflective Sampling cites this paper.

SafeSpec: Fast and Safe LLM via Dynamic Reflective Sampling SafeDecoding: Defending against Jailbreak Attacks via Safety-Aware Decoding

Reference 13

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T03:59:33.709134Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-26T17:23:06.382761Z digest=sha256:6eaac77f260812e6254e36f7b7d699bc49a07fa72f13ec04ffbf63566edc31cf

Observation 1596d35c-477a-491b-9715-f47be97690bd · inbound

Robust Harmful Features Under Jailbreak Attacks: Mechanistic Evidence from Attention Head Specialization in Large Language Models cites this paper.

Robust Harmful Features Under Jailbreak Attacks: Mechanistic Evidence from Attention Head Specialization in Large Language Models SafeDecoding: Defending against Jailbreak Attacks via Safety-Aware Decoding

Reference 41

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T17:35:51.363638Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-29T03:35:34.594617Z digest=sha256:9b468e4f01d8c5243d7e8d42872648bb5a49a5a54f2a11d0f9e13bbd114410fe