Pith. sign in

Paper Citation Record · LEDGER

Jailbreaking Large Language Models Against Moderation Guardrails via Cipher Characters

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2405.20413.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2405.20413 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:34:34.241608Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-18T21:36:52.437767Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 33ea019e-6694-4774-9866-fa52758af9d8 · inbound

JailbreakBench: An Open Robustness Benchmark for Jailbreaking Large Language Models cites this paper.

JailbreakBench: An Open Robustness Benchmark for Jailbreaking Large Language Models Jailbreaking Large Language Models Against Moderation Guardrails via Cipher Characters

Reference 24

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T06:08:05.627014Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-15T06:08:05.386345Z digest=sha256:5183fd4d85936bbe3e89dfb1fddea9b116b643b2dde15fa5bb8fa7965e7b1148

Observation 79c19ef7-639f-4dca-99db-c8d3a743b0e5 · inbound

From Hallucinations to Jailbreaks: Rethinking the Vulnerability of Large Foundation Models cites this paper.

From Hallucinations to Jailbreaks: Rethinking the Vulnerability of Large Foundation Models Jailbreaking Large Language Models Against Moderation Guardrails via Cipher Characters

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T12:34:34.241608Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:34:34.241608Z digest=sha256:d9e3fefb438eabc24e1d33867601cc784074e6e5bf6f33d50d8e91c6ade678a4

Observation 0aaacc26-e9ae-46e3-b432-fb1b3be88750 · inbound

InfoFlood: Jailbreaking Large Language Models with Information Overload cites this paper.

InfoFlood: Jailbreaking Large Language Models with Information Overload Jailbreaking Large Language Models Against Moderation Guardrails via Cipher Characters

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T01:02:28.929877Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T01:02:28.929877Z digest=sha256:32301665d748f5f08c892d589bb21491e2fd38b3f98a6bcf393b9463e3e65e6c

Observation e1f59553-784b-44f9-986b-a3ad5fa7f9e7 · inbound

"I Cannot Write This Because It Violates Our Content Policy": Understanding Content Moderation Policies and User Experiences in Generative AI Products cites this paper.

"I Cannot Write This Because It Violates Our Content Policy": Understanding Content Moderation Policies and User Experiences in Generative AI Products Jailbreaking Large Language Models Against Moderation Guardrails via Cipher Characters

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T00:27:19.635348Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:27:19.635348Z digest=sha256:249755a087aaad493b5a85af24ee92a7b09435c9561a075ac9ae42e36432ff8f

Observation 44863539-cf70-48cd-8278-ea533d6fc824 · inbound

GUARD: Guideline Upholding Test through Adaptive Role-play and Jailbreak Diagnostics for LLMs cites this paper.

GUARD: Guideline Upholding Test through Adaptive Role-play and Jailbreak Diagnostics for LLMs Jailbreaking Large Language Models Against Moderation Guardrails via Cipher Characters

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-05-18T21:36:52.440502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-18T21:34:51.665401Z digest=sha256:c44ac7c6e3c768e54ac4232e0acf6866baec925b60af1e5a29d079a2e1148f40

Observation b755d1cb-d319-4967-811b-25bc96a16eef · inbound

ASTRA: An Automated Framework for Strategy Discovery, Retrieval, and Evolution for Jailbreaking LLMs cites this paper.

ASTRA: An Automated Framework for Strategy Discovery, Retrieval, and Evolution for Jailbreaking LLMs Jailbreaking Large Language Models Against Moderation Guardrails via Cipher Characters

Reference 23

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T01:55:37.951931Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-18T01:54:22.995178Z digest=sha256:4b09bee4f8d6b0d998001b4453002d47c29ae141c977b96f796b7eb1a7837ac1