Pith. sign in

Paper Citation Record · LEDGER

JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 28 inbound Pith citation observations for arXiv:2402.05668.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2402.05668 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 28 of 28 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 28 of 28 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T10:23:06.057612Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T19:36:08.555578Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation a9b03fa3-e25e-4aab-91d9-3a40921e7934 · inbound

Refusal in Language Models Is Mediated by a Single Direction cites this paper.

Refusal in Language Models Is Mediated by a Single Direction JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs

Reference 126

Resolution
verified exact
arxiv_id, observed 2026-05-13T10:47:56.068686Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-13T10:47:55.934081Z digest=sha256:d754b94b9cb6b86263ba3e5ea4bc80c77513327c94b66ce8aba295cab21de656

Observation 73c472e5-a747-4521-b49c-e63f8ad69b10 · inbound

Jailbreak Attacks and Defenses Against Large Language Models: A Survey cites this paper.

Jailbreak Attacks and Defenses Against Large Language Models: A Survey JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-15T02:20:44.591202Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T02:20:44.368219Z digest=sha256:ff5430661b32d2e751b4100060cd95b749a91dd803cd0e56b602ed6784faa887

Observation 4efbf67c-a5e7-4ffe-9e72-3869d0b34bee · inbound

Harmful Fine-tuning Attacks and Defenses for Large Language Models: A Survey cites this paper.

Harmful Fine-tuning Attacks and Defenses for Large Language Models: A Survey JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-23T20:58:26.398289Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-23T20:58:16.237327Z digest=sha256:9d4932bddd98f587b925ebf1d467e647e60c53f6160b218c8ad50484618bd27e

Observation 537db251-5dc0-4874-b124-f378e4754d77 · inbound

Bridging the Safety Gap: A Guardrail Pipeline for Trustworthy LLM Inferences cites this paper.

Bridging the Safety Gap: A Guardrail Pipeline for Trustworthy LLM Inferences JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-08T10:23:06.057612Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T10:23:06.057612Z digest=sha256:e09853b0311f7f9c0a3ba3d66fafdc652ebcd3a75034fec21d051e8dc45fe2c1

Observation 9220165c-0f83-4a54-8e6b-180be994e0c6 · inbound

Towards medical AI misalignment: a preliminary study cites this paper.

Towards medical AI misalignment: a preliminary study JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T14:52:15.472625Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:52:15.472625Z digest=sha256:2a618c89271f23de876a03d130ea6c4d81790595f1adabc09e05317d33d66114

Observation 704992cc-917f-4293-987e-29d345098244 · inbound

Audio Jailbreak Attacks: Exposing Vulnerabilities in SpeechGPT in a White-Box Framework cites this paper.

Audio Jailbreak Attacks: Exposing Vulnerabilities in SpeechGPT in a White-Box Framework JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T14:27:05.591822Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:27:05.591822Z digest=sha256:7cf847fb5c0997c90e7be4b9c5b0dcff268b7389db6e40f0ac66193c5d44703d

Observation b2eaf344-2622-4f66-b217-4850c0eb5904 · inbound

Beyond Jailbreaks: Revealing Stealthier and Broader LLM Security Risks Stemming from Alignment Failures cites this paper.

Beyond Jailbreaks: Revealing Stealthier and Broader LLM Security Risks Stemming from Alignment Failures JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T05:40:20.651700Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:40:20.651700Z digest=sha256:eeed345ba11c61ce9d130829b53aa8e31a53d1444fd1f96fe70e653d4ee31fcb

Observation 266333ba-7c41-48f7-85d0-e27a04f77d93 · inbound

The Scales of Justitia: A Comprehensive Survey on Safety Evaluation of LLMs cites this paper.

The Scales of Justitia: A Comprehensive Survey on Safety Evaluation of LLMs JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-07T10:17:26.788943Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:17:26.788943Z digest=sha256:fc1d42d173207dc012f2f67aa8dcbf592bf447ea21d6cfcd2f96db94d7bdb9c3

Observation f4547e9d-2bed-49f3-a872-041039dcc0f2 · inbound

Investigating Vulnerabilities and Defenses Against Audio-Visual Attacks: A Comprehensive Survey Emphasizing Multimodal Models cites this paper.

Investigating Vulnerabilities and Defenses Against Audio-Visual Attacks: A Comprehensive Survey Emphasizing Multimodal Models JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T04:08:39.731690Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:08:39.731690Z digest=sha256:83465b755b9f822841f7d211851c6204651f156439fe07dadbd0bccf3283e3ce

Observation 92da5a23-91f1-46d2-80c5-233ed4293158 · inbound

InfoFlood: Jailbreaking Large Language Models with Information Overload cites this paper.

InfoFlood: Jailbreaking Large Language Models with Information Overload JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T01:02:28.442955Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T01:02:28.442955Z digest=sha256:356d8120ce905554b826fad1e856591bbce2d0dd1b53a6e0987ce057f36ef808

Observation 53031d22-f6a6-49d8-938f-7cc3cf6f8c47 · inbound

Q-resafe: Assessing Safety Risks and Quantization-aware Safety Patching for Quantized Large Language Models cites this paper.

Q-resafe: Assessing Safety Risks and Quantization-aware Safety Patching for Quantized Large Language Models JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs

Reference 2017

Resolution
unresolved
no resolver link, observed 2026-08-06T23:00:19.513275Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:00:19.513275Z digest=sha256:37d332e51e365fa48c36cffbe0a0af439495c5ced4306261f15c6d1ee412f44a

Observation 483123a3-0bfe-4cc5-9c74-2c4268e7b18d · inbound

Understanding How University Guidelines Address Privacy and Security Issues of Generative AI in Academic Settings cites this paper.

Understanding How University Guidelines Address Privacy and Security Issues of Generative AI in Academic Settings JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T22:49:58.020442Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:49:58.020442Z digest=sha256:335d04a802c43abf9d8c8559ed34980b536ded041d4304899be19ff2062add61

Observation 01705d48-dbcf-4a46-bc10-2a5b4228527b · inbound

Linearly Decoding Refused Knowledge in Aligned Language Models cites this paper.

Linearly Decoding Refused Knowledge in Aligned Language Models JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T21:27:34.892626Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:27:34.892626Z digest=sha256:63dc3a74204918fc866f4e7c72c54c5126ef7fe372a56221478d92ccedd83f0b

Observation 155f482f-eb0c-4b09-930b-c5236f1f9e56 · inbound

CAVGAN: Unifying Jailbreak and Defense of LLMs via Generative Adversarial Attacks on their Internal Representations cites this paper.

CAVGAN: Unifying Jailbreak and Defense of LLMs via Generative Adversarial Attacks on their Internal Representations JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T19:18:42.941390Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:18:42.941390Z digest=sha256:bf6d80efab6dc2c7330e8682a60ea40434eea4df315df35dd9c212c97fb53666

Observation 833a0af3-8de8-4956-aa95-189fb01a3a44 · inbound

GuardVal: Dynamic Large Language Model Jailbreak Evaluation for Comprehensive Safety Testing cites this paper.

GuardVal: Dynamic Large Language Model Jailbreak Evaluation for Comprehensive Safety Testing JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-06T18:40:52.130030Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:40:52.130030Z digest=sha256:ef2fe5834c4f76ddc6f7e95b51d414be771bf9850dffdc7672c50b656d8e5e9a

Observation a1426feb-59eb-4b50-8b38-4e88560801cd · inbound

An Audit and Analysis of LLM-Assisted Health Misinformation Jailbreaks Against LLMs cites this paper.

An Audit and Analysis of LLM-Assisted Health Misinformation Jailbreaks Against LLMs JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T01:02:13.588479Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T01:02:13.588479Z digest=sha256:f94a791f7ae24a7ec25ef9841f90e5ce19baea2caf65aa099459616caf2d9ddf

Observation 8d1dba16-b206-4dc7-b45c-769d6d1a2b67 · inbound

Layer-Wise Perturbations via Sparse Autoencoders for Adversarial Text Generation cites this paper.

Layer-Wise Perturbations via Sparse Autoencoders for Adversarial Text Generation JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs

Reference 135

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:45.234154Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:45.234154Z digest=sha256:0a936abc9e9bf99b58d051903286a5cf10d59188bbfd167ae4bf30e1dda0b61a

Observation 3c94b144-c84e-4a7f-92de-00f9a3c5adf4 · inbound

ORFuzz: Fuzzing the "Other Side" of LLM Safety -- Testing Over-Refusal cites this paper.

ORFuzz: Fuzzing the "Other Side" of LLM Safety -- Testing Over-Refusal JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-18T23:31:54.568455Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-18T23:27:52.438709Z digest=sha256:159ad5a37a62ff94adb1e4e865f1b86b3d11ca6ec236a8da825a4ffdd996ab8e

Observation b71221d5-a147-49b2-8228-46fa8024f7d3 · inbound

SALMAN: Stability Analysis of Language Models Through the Maps Between Graph-based Manifolds cites this paper.

SALMAN: Stability Analysis of Language Models Through the Maps Between Graph-based Manifolds JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-05T17:14:37.440851Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:14:37.440851Z digest=sha256:e5393e1aafabd3b9fc4ec8defcfdd42f0f2fe3d0b3aa5ef56c2e6f78222e2849

Observation 4cd21b06-c7ea-41e7-b016-ad9b2b1c5695 · inbound

Turning the Spell Around: Lightweight Alignment Amplification via Rank-One Safety Injection cites this paper.

Turning the Spell Around: Lightweight Alignment Amplification via Rank-One Safety Injection JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-05T14:57:12.791829Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T14:57:12.791829Z digest=sha256:cc43c84b68a5b33ef41354a8f41eff73cc7ccbf77323380b669fcf1aa878bfc6

Observation 914ffc20-a3ee-46f8-9239-f750fb86e931 · inbound

JADES: A Universal Framework for Jailbreak Assessment via Decompositional Scoring cites this paper.

JADES: A Universal Framework for Jailbreak Assessment via Decompositional Scoring JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-05T14:51:03.645921Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T14:51:03.645921Z digest=sha256:25f1198a84ddb4458db830993eaf6d312c8e812b95210a568e2ebab63961aacd

Observation e95fd325-7c11-47a5-ae6b-5c7da99a67f0 · inbound

LLM in the Middle: A Systematic Review of Threats and Mitigations to Real-World LLM-based Systems cites this paper.

LLM in the Middle: A Systematic Review of Threats and Mitigations to Real-World LLM-based Systems JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-04T17:46:15.475308Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:46:15.475308Z digest=sha256:fe752663bceaf3cf9708aa607a75c901bd1055db8a9b81d95b96666f25feb2e7

Observation 4efc1b01-efce-4b62-be4b-0c1e0775fafb · inbound

How Well Do AI Systems Solve AP Physics? A Comparative Evaluation of Large Language Models on Algebra-Based Free Response Questions cites this paper.

How Well Do AI Systems Solve AP Physics? A Comparative Evaluation of Large Language Models on Algebra-Based Free Response Questions JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs

Reference 6

Resolution
unresolved
no resolver link, observed 2026-07-15T13:15:31.225992Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-15T13:15:31.225992Z digest=sha256:342136c6cdea22d2a42b6ccd1c399be221019deb5802b42a2543b82e74f13386

Observation 83b253df-03b6-4e6f-9d82-15dc863af4a2 · inbound

The Art of (Mis)alignment: How Fine-Tuning Methods Effectively Misalign and Realign LLMs in Post-Training cites this paper.

The Art of (Mis)alignment: How Fine-Tuning Methods Effectively Misalign and Realign LLMs in Post-Training JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-11T00:45:50.609141Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T18:18:56.476698Z digest=sha256:cc1ebcb8be56301577162026a26e8ee9b4dc3ee1895bd6e2227eeea31fb8fc09

Observation bb8ba95f-ef39-43f5-b0d1-985a43a95377 · inbound

SoK: Robustness in Large Language Models against Jailbreak Attacks cites this paper.

SoK: Robustness in Large Language Models against Jailbreak Attacks JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-11T18:01:08.866143Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-08T16:42:41.137808Z digest=sha256:733493d00c645b32591239955d8c7867de12197134fc0a7898a170afb35db693

Observation 356aab86-984e-425c-80a4-1165029df8ba · inbound

Persona Attack: Incremental Memory Injection Jailbreak Attack against Large Language Models cites this paper.

Persona Attack: Incremental Memory Injection Jailbreak Attack against Large Language Models JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:36:08.557072Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-28T22:23:00.894059Z digest=sha256:0e3463246ef754b037dfa0a91cd1edf3122e9fdd65a08260b6f20ba970221cbf

Observation 6453eee7-1ef5-458a-9d47-4ea67e510ce6 · inbound

SCARCE: Scalable Cascade Analysis for Rare-event Characterisation via Embeddings cites this paper.

SCARCE: Scalable Cascade Analysis for Rare-event Characterisation via Embeddings JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-06-30T07:04:20.789297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T07:03:36.414202Z digest=sha256:448e2abd91e49c9509c4dd2490a7dc230d46886a56d4a8cf184fdf4df6c22746

Observation 30413506-c871-4d98-ae6d-1776f7b318c1 · inbound

AIR-BENCH Live: An Evolving Safety Benchmark for Foundation Models cites this paper.

AIR-BENCH Live: An Evolving Safety Benchmark for Foundation Models JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-02T08:30:02.436170Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:30:02.436170Z digest=sha256:dbad0365c7fe5deb47b2bbbe75b2ec94658daca5c0dd0699360903bf359c4fa0