Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 20 inbound Pith citation observations for arXiv:2402.08983.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-08T12:25:30.688823Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T03:59:33.707048Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation cf1f22e1-580c-4e5a-ba21-6dc19222373c · inbound
Jailbreak Attacks and Defenses Against Large Language Models: A Survey SafeDecoding: Defending against Jailbreak Attacks via Safety-Aware Decoding
Reference 102
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f0ea2f5a-db58-414d-b1fd-6642c58c4119 · inbound
JBShield: Defending Large Language Models from Jailbreak Attacks through Activated Concept Analysis and Manipulation SafeDecoding: Defending against Jailbreak Attacks via Safety-Aware Decoding
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 434dbf62-dba0-407d-a3db-c61502453e29 · inbound
The Hidden Dimensions of LLM Alignment: A Multi-Dimensional Analysis of Orthogonal Safety Directions SafeDecoding: Defending against Jailbreak Attacks via Safety-Aware Decoding
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 81774b8d-8151-4458-9f5b-abf3727b8984 · inbound
GloSS over Toxicity: Understanding and Mitigating Toxicity in LLMs via Global Toxic Subspace SafeDecoding: Defending against Jailbreak Attacks via Safety-Aware Decoding
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6063ade9-b576-467b-acb4-3cb2196d7f7e · inbound
Just Enough Shifts: Mitigating Over-Refusal in Aligned Language Models with Targeted Representation Fine-Tuning SafeDecoding: Defending against Jailbreak Attacks via Safety-Aware Decoding
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 639dec63-bc50-4249-b2cf-74d0d7b6919f · inbound
PUZZLED: Jailbreaking LLMs through Word-Based Puzzles SafeDecoding: Defending against Jailbreak Attacks via Safety-Aware Decoding
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 210db1dc-0fcf-4a74-b547-da981641cae6 · inbound
Beyond Surface-Level Detection: Towards Cognitive-Driven Defense Against Jailbreak Attacks via Meta-Operations Reasoning SafeDecoding: Defending against Jailbreak Attacks via Safety-Aware Decoding
Reference 187
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fa55204e-3bbd-4601-95a4-cfe519237c62 · inbound
ReasoningGuard: Safeguarding Large Reasoning Models with Inference-time Safety Aha Moments SafeDecoding: Defending against Jailbreak Attacks via Safety-Aware Decoding
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 0b0fb6d5-9e85-465a-88d3-191905b818b9 · inbound
A Survey on Training-free Alignment of Large Language Models SafeDecoding: Defending against Jailbreak Attacks via Safety-Aware Decoding
Reference 96
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7619697b-7203-48f8-8ed1-fb63a1dada03 · inbound
A Comprehensive Survey on Trustworthiness in Reasoning with Large Language Models SafeDecoding: Defending against Jailbreak Attacks via Safety-Aware Decoding
Reference 267
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0986a88f-38f8-4ea1-9def-a33b45ba276e · inbound
MoGU V2: Toward a Higher Pareto Frontier Between Model Usability and Security SafeDecoding: Defending against Jailbreak Attacks via Safety-Aware Decoding
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a5d493be-74b1-4ce1-bc43-1ed95bf91700 · inbound
TrajGuard: Streaming Hidden-state Trajectory Detection for Decoding-time Jailbreak Defense SafeDecoding: Defending against Jailbreak Attacks via Safety-Aware Decoding
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c920d2a6-be03-4670-8fdc-06c68e8b3dfd · inbound
Jailbreaking the Matrix: Nullspace Steering for Controlled Model Subversion SafeDecoding: Defending against Jailbreak Attacks via Safety-Aware Decoding
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 4a9abf7e-acab-4d20-9778-4bb2a44b8bd7 · inbound
Jailbreaking Frontier Foundation Models Through Intention Deception SafeDecoding: Defending against Jailbreak Attacks via Safety-Aware Decoding
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 50551200-c250-4ecb-ab89-7297a7cf4dcc · inbound
Minimal, Local, Causal Explanations for Jailbreak Success in Large Language Models SafeDecoding: Defending against Jailbreak Attacks via Safety-Aware Decoding
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 3efa7aa7-9924-440f-9f25-3af5375d4347 · inbound
SafeSteer: A Decoding-level Defense Mechanism for Multimodal Large Language Models SafeDecoding: Defending against Jailbreak Attacks via Safety-Aware Decoding
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation bd96da19-f6e6-4fbf-8aad-ec088501005c · inbound
NeuroArmor: Safe-Variant-Guided Representation Consistency for Selective Re-Anchoring in Jailbreak Defense SafeDecoding: Defending against Jailbreak Attacks via Safety-Aware Decoding
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e7372632-2dd5-4540-8213-2e1e0ac76ba5 · inbound
SlotGCG: Exploiting the Positional Vulnerability in LLMs for Jailbreak Attacks SafeDecoding: Defending against Jailbreak Attacks via Safety-Aware Decoding
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a385f6c9-16d1-429e-ae3e-d9e873aa672b · inbound
SafeSpec: Fast and Safe LLM via Dynamic Reflective Sampling SafeDecoding: Defending against Jailbreak Attacks via Safety-Aware Decoding
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1596d35c-477a-491b-9715-f47be97690bd · inbound
Robust Harmful Features Under Jailbreak Attacks: Mechanistic Evidence from Attention Head Specialization in Large Language Models SafeDecoding: Defending against Jailbreak Attacks via Safety-Aware Decoding
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.