Pith. sign in

Paper Citation Record · LEDGER

Don't Say No: Jailbreaking LLM by Suppressing Refusal

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 12 inbound Pith citation observations for arXiv:2404.16369.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2404.16369 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 12 of 12 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 12 of 12 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T10:28:52.855847Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-24T00:38:39.819593Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 366c2e06-2992-430f-869e-3a8443a2753c · inbound

Uncovering Logit Suppression Vulnerabilities in LLM Safety Alignment cites this paper.

Uncovering Logit Suppression Vulnerabilities in LLM Safety Alignment Don't Say No: Jailbreaking LLM by Suppressing Refusal

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-24T00:38:39.826939Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-24T00:38:36.992597Z digest=sha256:ddf147947224f8c2cacdcc4951267065843c5083b6139f93d1760adb5cf6a6fe

Observation e224c9e7-a992-4dda-95ae-7a0912f68518 · inbound

Jailbreak Attacks and Defenses Against Large Language Models: A Survey cites this paper.

Jailbreak Attacks and Defenses Against Large Language Models: A Survey Don't Say No: Jailbreaking LLM by Suppressing Refusal

Reference 123

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T02:20:44.527408Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-15T02:20:44.368219Z digest=sha256:d3fd3244df8878d5690310f6ff9dd299c5b855c1f0bef0df056b35a8c461fe56

Observation 34896cab-3e45-46f0-b59a-f53c3fcb04d0 · inbound

From Rogue to Safe AI: The Role of Explicit Refusals in Aligning LLMs with International Humanitarian Law cites this paper.

From Rogue to Safe AI: The Role of Explicit Refusals in Aligning LLMs with International Humanitarian Law Don't Say No: Jailbreaking LLM by Suppressing Refusal

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:52.855847Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:28:52.855847Z digest=sha256:71d613d597e3f83f2432992295d31005ac79a2f8b644e7f5bd0e54c15f82ba00

Observation 72ba783a-5b5a-4037-a109-1035b69cdfe8 · inbound

Beyond Jailbreaks: Revealing Stealthier and Broader LLM Security Risks Stemming from Alignment Failures cites this paper.

Beyond Jailbreaks: Revealing Stealthier and Broader LLM Security Risks Stemming from Alignment Failures Don't Say No: Jailbreaking LLM by Suppressing Refusal

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T05:40:20.621417Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:40:20.621417Z digest=sha256:779ca27fd15ebb925f304d2ff0627cc3db0bf9fb1e4b75554dc8e55876b1a096

Observation 42a48662-4628-4735-83ee-9602c01ffcd3 · inbound

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models cites this paper.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Don't Say No: Jailbreaking LLM by Suppressing Refusal

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:57.656839Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:57.656839Z digest=sha256:74e9935e0e81d0ce4e6e0dd9125f4fcc5e1b80bbfe20e16350733c9cca7bda39

Observation 36a7ccd1-900a-40af-92c3-8adda79cb4ea · inbound

Leveraging the Potential of Prompt Engineering for Hate Speech Detection in Low-Resource Languages cites this paper.

Leveraging the Potential of Prompt Engineering for Hate Speech Detection in Low-Resource Languages Don't Say No: Jailbreaking LLM by Suppressing Refusal

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-06T21:34:41.418171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:34:41.418171Z digest=sha256:f622edb0d1355c51e5d8e759e646d0f1871422537cff5f401b95dc159ff664a4

Observation 17b10907-dc4b-451b-b535-e0841e28f91e · inbound

Layer-Wise Perturbations via Sparse Autoencoders for Adversarial Text Generation cites this paper.

Layer-Wise Perturbations via Sparse Autoencoders for Adversarial Text Generation Don't Say No: Jailbreaking LLM by Suppressing Refusal

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:33.756873Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:33.756873Z digest=sha256:5da7969cbac4fff2ffa2ecb0d63955d73cdb0d2954f69b216cea8fb22c440916

Observation 0acbcdc5-64f0-4ef2-937b-b660adecc5c2 · inbound

JADES: A Universal Framework for Jailbreak Assessment via Decompositional Scoring cites this paper.

JADES: A Universal Framework for Jailbreak Assessment via Decompositional Scoring Don't Say No: Jailbreaking LLM by Suppressing Refusal

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-05T14:51:03.812481Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T14:51:03.812481Z digest=sha256:f0eee666e5e7de9af903325552f2e371dc7a30362d4149d73224f7a66cbf09f6

Observation 53dee416-6ad5-4d83-a501-a5a3c659e3de · inbound

SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses cites this paper.

SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Don't Say No: Jailbreaking LLM by Suppressing Refusal

Reference 246

Resolution
unresolved
no resolver link, observed 2026-08-04T09:25:58.741705Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:25:58.741705Z digest=sha256:d9f1ec02d037821aa600b78b84de9be9303c5a36528ded8c808e3576d4bae342

Observation 85d5ee7b-4e79-40c2-a9b7-8a97940035a2 · inbound

Break Me If You Can: Self-Jailbreaking of Aligned LLMs via Lexical Insertion Prompting cites this paper.

Break Me If You Can: Self-Jailbreaking of Aligned LLMs via Lexical Insertion Prompting Don't Say No: Jailbreaking LLM by Suppressing Refusal

Reference 9

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T18:08:13.063281Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-16T18:04:44.543311Z digest=sha256:56a64364652cacaf3e33de630ae2f75de37c29dab5c50acd5532baa3ad94f184

Observation a46b9f70-7727-424c-9626-25a7dfb622f3 · inbound

The Salami Slicing Threat: Exploiting Cumulative Risks in LLM Systems cites this paper.

The Salami Slicing Threat: Exploiting Cumulative Risks in LLM Systems Don't Say No: Jailbreaking LLM by Suppressing Refusal

Reference 35

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T09:16:03.956567Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T16:07:31.602378Z digest=sha256:ede84d1958bd00839fb5b709083896c28c9ee5003a7db0aab5af94ff9c0b7f8c

Observation ed39c6ea-b358-4627-9245-5009824ddd61 · inbound

Adversarial Reframing: A Framework for Targeted Generation in Language Models cites this paper.

Adversarial Reframing: A Framework for Targeted Generation in Language Models Don't Say No: Jailbreaking LLM by Suppressing Refusal

Reference 64

Resolution
verified exact
arxiv_id, observed 2026-05-22T09:36:20.999209Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-22T09:35:51.862736Z digest=sha256:d97820ccad1c7cb3be2909c0cfd886a53bd581e927438aa779319a94515b2ee4