Pith. sign in

Paper Citation Record · LEDGER

Bypassing LLM Guardrails: An Empirical Analysis of Evasion Attacks against Prompt Injection and Jailbreak Detection Systems

As of 19 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 12 inbound Pith citation observations for arXiv:2504.11168.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2504.11168 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 12 of 12 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 12 of 12 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T23:23:38.880495Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation f296bbb6-81ee-408c-a4db-8080aae50f7d · inbound

A Byzantine Fault Tolerance Approach towards AI Safety cites this paper.

A Byzantine Fault Tolerance Approach towards AI Safety Bypassing LLM Guardrails: An Empirical Analysis of Evasion Attacks against Prompt Injection and Jailbreak Detection Systems

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-22T19:16:58.783644Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-22T19:15:09.861706Z digest=sha256:b13a4573c51dedfdc881f4c71ad95f6f65ae0252f15d2872e3ac682e36fa46cc

Observation 9e36fdde-7b4d-4e3c-a81f-dfbe9af0e352 · inbound

Red Teaming the Mind of the Machine: A Systematic Evaluation of Prompt Injection and Jailbreak Vulnerabilities in LLMs cites this paper.

Red Teaming the Mind of the Machine: A Systematic Evaluation of Prompt Injection and Jailbreak Vulnerabilities in LLMs Bypassing LLM Guardrails: An Empirical Analysis of Evasion Attacks against Prompt Injection and Jailbreak Detection Systems

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-15T23:23:38.880495Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:23:38.880495Z digest=sha256:86277896b28065845ad14cedb25f963f48a8b498ceb8ebd427eb67a15e9fa322

Observation af66bd55-7b18-4682-93de-aead3f4b420e · inbound

Beyond Semantics: The Unreasonable Effectiveness of Reasonless Intermediate Tokens cites this paper.

Beyond Semantics: The Unreasonable Effectiveness of Reasonless Intermediate Tokens Bypassing LLM Guardrails: An Empirical Analysis of Evasion Attacks against Prompt Injection and Jailbreak Detection Systems

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-15T20:15:17.903838Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:15:17.903838Z digest=sha256:98e7dfe59f76bd409dc61a092ff289032f862bc3acff38aab43007edbec9f496

Observation 656ad9a6-f126-46b6-be17-5d45e6ad8bb2 · inbound

EVADE-Bench: Multimodal Benchmark for Evaluating and Enhancing Evasive Content Detection cites this paper.

EVADE-Bench: Multimodal Benchmark for Evaluating and Enhancing Evasive Content Detection Bypassing LLM Guardrails: An Empirical Analysis of Evasion Attacks against Prompt Injection and Jailbreak Detection Systems

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:03.097933Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:03.097933Z digest=sha256:e2b1bfe774943da44829801dd527cf7ccbc98dfc62f13445f5d27641d2349ecf

Observation a94fe550-8ff2-4ed4-b795-cb290d71e09d · inbound

When Your Reviewer is an LLM: Biases, Divergence, and Prompt Injection Risks in Peer Review cites this paper.

When Your Reviewer is an LLM: Biases, Divergence, and Prompt Injection Risks in Peer Review Bypassing LLM Guardrails: An Empirical Analysis of Evasion Attacks against Prompt Injection and Jailbreak Detection Systems

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-04T18:34:49.857445Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:34:49.857445Z digest=sha256:97ce95196617204d8273e9468748a139eb28e75ff43aed20b00a3a5e56d6958c

Observation 76bcc266-7f5b-4bf9-ae9a-6f331c26adf1 · inbound

Exploiting Web Search Tools of AI Agents for Data Exfiltration cites this paper.

Exploiting Web Search Tools of AI Agents for Data Exfiltration Bypassing LLM Guardrails: An Empirical Analysis of Evasion Attacks against Prompt Injection and Jailbreak Detection Systems

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-18T08:36:07.217273Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-18T08:36:04.262528Z digest=sha256:43f0f782a90f476fc35d72616e6d73735ce99e2ffe10975d053def0a3ad42002

Observation cbe39f0e-1efa-492f-9af0-78502964a57d · inbound

Beyond Pattern Matching: Seven Cross-Domain Techniques for Prompt Injection Detection cites this paper.

Beyond Pattern Matching: Seven Cross-Domain Techniques for Prompt Injection Detection Bypassing LLM Guardrails: An Empirical Analysis of Evasion Attacks against Prompt Injection and Jailbreak Detection Systems

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-02T15:56:48.257415Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T15:56:48.257415Z digest=sha256:3bcc1e78407a4379a20610fa3c287f22799497b468f5add6930decb985d7aaf0

Observation 5b6e66dc-3014-4f87-bc5f-81cb25b6775c · inbound

PsychoPass: Geometric Profiling of Multi-Turn Adversarial LLM Conversations cites this paper.

PsychoPass: Geometric Profiling of Multi-Turn Adversarial LLM Conversations Bypassing LLM Guardrails: An Empirical Analysis of Evasion Attacks against Prompt Injection and Jailbreak Detection Systems

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-07-02T03:36:29.442477Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T09:55:56.227335Z digest=sha256:f2faebc4816d0925a6e1edd09bf20448c8c75b46c64dfdeb7ada195ca9be87a4

Observation 9c7a0633-192b-4021-b033-0c2f1c7bfaa4 · inbound

Short paper: Models in the dark -- Rectification and erasure under GDPR in ML supply chains cites this paper.

Short paper: Models in the dark -- Rectification and erasure under GDPR in ML supply chains Bypassing LLM Guardrails: An Empirical Analysis of Evasion Attacks against Prompt Injection and Jailbreak Detection Systems

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-06-28T03:31:30.366296Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T03:24:45.955589Z digest=sha256:fd17812b7266880ae8247c48a5e6214ca07e9bea3bdd7a766749cafa9f426478

Observation adde285b-0aec-4e56-bb24-13ed97d70d5a · inbound

From Shield to Target: Denial-of-Service Attacks on LLM-Based Agent Guardrails cites this paper.

From Shield to Target: Denial-of-Service Attacks on LLM-Based Agent Guardrails Bypassing LLM Guardrails: An Empirical Analysis of Evasion Attacks against Prompt Injection and Jailbreak Detection Systems

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-07-03T16:58:43.511470Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-27T04:42:05.886984Z digest=sha256:0f77b30f029ffa3a9a5009471613c4243c083c75af0d25f5dcc222183d8f39d7

Observation c42ea2ac-1377-4141-8751-65258d51ca10 · inbound

Investigating The Security of Modern AI and Cloud Infrastructure cites this paper.

Investigating The Security of Modern AI and Cloud Infrastructure Bypassing LLM Guardrails: An Empirical Analysis of Evasion Attacks against Prompt Injection and Jailbreak Detection Systems

Reference 58

Resolution
verified exact
arxiv_id, observed 2026-07-04T08:29:42.302084Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-26T11:31:39.910784Z digest=sha256:74203feda37093378c98cee923fa12d905fac42d5cdf2471493391471ce5b320

Observation 3d867835-2020-452d-a5c9-1100f787cedb · inbound

Toward Self-Evolution-Ready Workflow Harnesses: A Reversible Migration Path and Convertibility Taxonomy for Expert LLM Pipelines cites this paper.

Toward Self-Evolution-Ready Workflow Harnesses: A Reversible Migration Path and Convertibility Taxonomy for Expert LLM Pipelines Bypassing LLM Guardrails: An Empirical Analysis of Evasion Attacks against Prompt Injection and Jailbreak Detection Systems

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-07-03T17:48:46.404918Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-27T03:39:24.426862Z digest=sha256:69fe1504081ed2f2f591ecd65be356d96095014fac66d867ea25ad29b3de8f1d