Pith. sign in

Paper Citation Record · LEDGER

GradSafe: Detecting Jailbreak Prompts for LLMs via Safety-Critical Gradient Analysis

As of 14 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 9 inbound Pith citation observations for arXiv:2402.13494.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2402.13494 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 9 of 9 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 9 of 9 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-12T12:56:34.413670Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T08:29:42.339750Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation bfc5b8be-fe6f-42bb-9849-bbbbf7bab971 · inbound

Jailbreak Attacks and Defenses Against Large Language Models: A Survey cites this paper.

Jailbreak Attacks and Defenses Against Large Language Models: A Survey GradSafe: Detecting Jailbreak Prompts for LLMs via Safety-Critical Gradient Analysis

Reference 101

Resolution
verified exact
arxiv_id, observed 2026-05-15T02:20:44.432029Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-15T02:20:44.368219Z digest=sha256:07e1b7e069f74dc09b56920ede066406207ab5ebf08de26d1dfc177e8a9a7e25

Observation c628ec6d-4b6e-4b2d-821f-4fe4f4b1549c · inbound

Preventing Jailbreak Prompts as Malicious Tools for Cybercriminals: A Cyber Defense Perspective cites this paper.

Preventing Jailbreak Prompts as Malicious Tools for Cybercriminals: A Cyber Defense Perspective GradSafe: Detecting Jailbreak Prompts for LLMs via Safety-Critical Gradient Analysis

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-12T12:56:34.413670Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T12:56:34.413670Z digest=sha256:230983cf3708075843a1938d02f02abb6eb285b6e74056cabaccec63451e88b9

Observation b165e047-38ca-4c02-b6e1-6c0447bc211c · inbound

JBShield: Defending Large Language Models from Jailbreak Attacks through Activated Concept Analysis and Manipulation cites this paper.

JBShield: Defending Large Language Models from Jailbreak Attacks through Activated Concept Analysis and Manipulation GradSafe: Detecting Jailbreak Prompts for LLMs via Safety-Critical Gradient Analysis

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-08T12:25:30.681565Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T12:25:30.681565Z digest=sha256:bfe657b382251b3dec0795218cfdaad4ce350cc17530e2bcce89aec0dc2fb32f

Observation d81d0741-e5fd-4645-8705-a26bd19019b9 · inbound

The First Differentiable Transfer-Based Algorithm for Discrete MicroLED Repair cites this paper.

The First Differentiable Transfer-Based Algorithm for Discrete MicroLED Repair GradSafe: Detecting Jailbreak Prompts for LLMs via Safety-Critical Gradient Analysis

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-05T22:21:02.842889Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T22:21:02.842889Z digest=sha256:f8acd4e0ac826f6f3a58390b294f908414825ca0bc489657e25e3e513628fc9b

Observation f04e810c-0149-4d36-a99b-ea39639eca8f · inbound

SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses cites this paper.

SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses GradSafe: Detecting Jailbreak Prompts for LLMs via Safety-Critical Gradient Analysis

Reference 202

Resolution
unresolved
no resolver link, observed 2026-08-04T09:25:58.003099Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:25:58.003099Z digest=sha256:c5a0aa9cd6dec0542b43aed82bab40ffa9cc82537d1d5f53db5885b38a298930

Observation fdb5ed4b-ee3e-45cb-ad40-d7363ff7f1aa · inbound

Distributionally Robust Token Optimization in RLHF cites this paper.

Distributionally Robust Token Optimization in RLHF GradSafe: Detecting Jailbreak Prompts for LLMs via Safety-Critical Gradient Analysis

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-14T23:23:16.015683Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-14T23:19:47.586351Z digest=sha256:0b6a7adca5238ff1f18a7ca2f5d898ac32f8ac1a51050530b033e24f0f3b76cd

Observation dffa1a2a-ab29-4773-ad4c-e83dab82cdf2 · inbound

SafeAgent: A Runtime Protection Architecture for Agentic Systems cites this paper.

SafeAgent: A Runtime Protection Architecture for Agentic Systems GradSafe: Detecting Jailbreak Prompts for LLMs via Safety-Critical Gradient Analysis

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-10T06:11:20.543560Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-10T06:06:49.717091Z digest=sha256:ad63810641a1c87ab8810c4f7f63f68557f2255aa739374786078faed1c2d6f1

Observation cd7b3706-ae23-4913-9dce-efa232d47631 · inbound

Defending Jailbreak Attacks on Large Language Models via Manifold Trajectory Kinetics cites this paper.

Defending Jailbreak Attacks on Large Language Models via Manifold Trajectory Kinetics GradSafe: Detecting Jailbreak Prompts for LLMs via Safety-Critical Gradient Analysis

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-07-02T17:37:14.900803Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-27T21:55:48.561400Z digest=sha256:d0b1da0634da66ea2bee452601b6979a62f52b18ef080a90f4a44b287f8e52ce

Observation 7c183d37-1553-43ec-8b63-7a205ef9d94b · inbound

Investigating The Security of Modern AI and Cloud Infrastructure cites this paper.

Investigating The Security of Modern AI and Cloud Infrastructure GradSafe: Detecting Jailbreak Prompts for LLMs via Safety-Critical Gradient Analysis

Reference 142

Resolution
verified exact
arxiv_id, observed 2026-07-04T08:29:42.341232Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-26T11:31:39.910784Z digest=sha256:e4a1d2f78947f34bd76c68d84b0fe3e88bb4e8889a4a6203ec8479143d05c829