Pith. sign in

Paper Citation Record · LEDGER

Exploiting Programmatic Behavior of LLMs: Dual-Use Through Standard Security Attacks

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 24 inbound Pith citation observations for arXiv:2302.05733.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2302.05733 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 24 of 24 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 24 of 24 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T19:26:13.657815Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

27
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 40842c4f-588b-44d0-85a7-b46df7e64d03 · inbound

Jailbroken: How Does LLM Safety Training Fail? cites this paper.

Jailbroken: How Does LLM Safety Training Fail? Exploiting Programmatic Behavior of LLMs: Dual-Use Through Standard Security Attacks

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-14T18:17:42.855510Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-14T18:17:42.752997Z digest=sha256:d5d7bd292a087d985cf1ad3154246b2971cbb346e083ead7c0480e3eb0a9af0c

Observation 890a477b-2c06-42aa-b0c7-c603ae315350 · inbound

"Do Anything Now": Characterizing and Evaluating In-The-Wild Jailbreak Prompts on Large Language Models cites this paper.

"Do Anything Now": Characterizing and Evaluating In-The-Wild Jailbreak Prompts on Large Language Models Exploiting Programmatic Behavior of LLMs: Dual-Use Through Standard Security Attacks

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-05-17T08:39:28.174978Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-17T08:39:28.047394Z digest=sha256:7e03d8c4fcd1da93ead856b7b9c3ef1064408c09ed1d21123e20f590b43d072e

Observation 72287e84-52b2-4c64-83c7-763383e4bd24 · inbound

Baseline Defenses for Adversarial Attacks Against Aligned Language Models cites this paper.

Baseline Defenses for Adversarial Attacks Against Aligned Language Models Exploiting Programmatic Behavior of LLMs: Dual-Use Through Standard Security Attacks

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-13T23:24:40.023866Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-13T23:24:39.835347Z digest=sha256:9e7494708afcfb0e4296a14bf002c984bb9fcc5aefb99a2cf4a108381c83e0c3

Observation dec75bb5-f715-4e46-a240-8010246d6287 · inbound

AutoDAN: Generating Stealthy Jailbreak Prompts on Aligned Large Language Models cites this paper.

AutoDAN: Generating Stealthy Jailbreak Prompts on Aligned Large Language Models Exploiting Programmatic Behavior of LLMs: Dual-Use Through Standard Security Attacks

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-12T16:28:04.062859Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-12T16:28:03.996446Z digest=sha256:4fa408edaf3bc5262b528f92e33fa38f3a073ed0a496dcb1530c8c2e655be2b3

Observation 5ca57504-058d-4863-8fff-5f7e9942d0d9 · inbound

Catastrophic Jailbreak of Open-source LLMs via Exploiting Generation cites this paper.

Catastrophic Jailbreak of Open-source LLMs via Exploiting Generation Exploiting Programmatic Behavior of LLMs: Dual-Use Through Standard Security Attacks

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-16T22:00:51.556822Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-16T22:00:51.487120Z digest=sha256:29ee7487661ddb1d987911f1d54bf2980fe1390ca78ddb2a1cb55fef5f99bfa6

Observation 8f19794a-fcc9-4122-ac96-88a49465957f · inbound

Whispers in the Machine: Confidentiality in Agentic Systems cites this paper.

Whispers in the Machine: Confidentiality in Agentic Systems Exploiting Programmatic Behavior of LLMs: Dual-Use Through Standard Security Attacks

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-24T04:03:53.849912Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-24T03:59:03.972043Z digest=sha256:65330eb31379ec71f778f98d39eecdd73cb74cd786ba9883b9ea85a7b97b1185

Observation a1aa8d47-9e09-4692-b3f3-498a0ada33ba · inbound

A StrongREJECT for Empty Jailbreaks cites this paper.

A StrongREJECT for Empty Jailbreaks Exploiting Programmatic Behavior of LLMs: Dual-Use Through Standard Security Attacks

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-16T21:28:02.825041Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-16T21:28:02.745230Z digest=sha256:a5bd52e71c0ce5bcf1a54118d686e510e591ab26823475ac56dbe46692e9ad0b

Observation 83ea2ea0-dc41-4608-8feb-36b8746ef217 · inbound

LLM Agents can Autonomously Exploit One-day Vulnerabilities cites this paper.

LLM Agents can Autonomously Exploit One-day Vulnerabilities Exploiting Programmatic Behavior of LLMs: Dual-Use Through Standard Security Attacks

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-18T04:18:27.709259Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-18T04:18:27.597704Z digest=sha256:b50cbcc7b1b41eea4b8ff8c4164b36ee94c077e80057bbb90e306a00ce7936af

Observation db4aa327-8f7b-4cc0-938c-63201fa0ad10 · inbound

Jailbreak Attacks and Defenses Against Large Language Models: A Survey cites this paper.

Jailbreak Attacks and Defenses Against Large Language Models: A Survey Exploiting Programmatic Behavior of LLMs: Dual-Use Through Standard Security Attacks

Reference 43

Resolution
verified exact
arxiv_id, observed 2026-05-15T02:20:44.691880Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T02:20:44.368219Z digest=sha256:3cb67def6adb041ffcbb99955425b3e85ac7e93bcad45c0e006e5e89b9636c13

Observation d826b6f6-c434-4f2f-9f3c-ebcb4d31cbff · inbound

Prompt Infection: LLM-to-LLM Prompt Injection within Multi-Agent Systems cites this paper.

Prompt Infection: LLM-to-LLM Prompt Injection within Multi-Agent Systems Exploiting Programmatic Behavior of LLMs: Dual-Use Through Standard Security Attacks

Reference 64

Resolution
verified exact
arxiv_id, observed 2026-05-15T19:32:19.734280Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-15T19:32:19.405615Z digest=sha256:42342f814c5a9fa9396fdbfdda3174cb3cff16114f7b4e1d2b4167601f9db962

Observation 87dea015-967a-41c4-b047-429f37b479c2 · inbound

Faster-GCG: Efficient Discrete Optimization Jailbreak Attacks against Aligned Large Language Models cites this paper.

Faster-GCG: Efficient Discrete Optimization Jailbreak Attacks against Aligned Large Language Models Exploiting Programmatic Behavior of LLMs: Dual-Use Through Standard Security Attacks

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-05-23T19:03:21.401463Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-23T19:01:43.922848Z digest=sha256:5c541bf54ff8f65f3a6d98d451195203f4e984029175d4f790cbe694ad14b7ac

Observation 87c3da8a-2bd1-4af8-98b4-e89550e2d730 · inbound

Circumventing Safety Alignment in Large Language Models Through Embedding Space Toxicity Attenuation cites this paper.

Circumventing Safety Alignment in Large Language Models Through Embedding Space Toxicity Attenuation Exploiting Programmatic Behavior of LLMs: Dual-Use Through Standard Security Attacks

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T19:26:13.657815Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:26:13.657815Z digest=sha256:64fd8e123bcb9cf509cc3fc5f71eeaecd4ef6ac2c899965368fb46d3891e8f32

Observation 503994c0-873c-4d01-bc23-7d6f96f3ef5c · inbound

Layer-Wise Perturbations via Sparse Autoencoders for Adversarial Text Generation cites this paper.

Layer-Wise Perturbations via Sparse Autoencoders for Adversarial Text Generation Exploiting Programmatic Behavior of LLMs: Dual-Use Through Standard Security Attacks

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:41.107700Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:41.107700Z digest=sha256:72951007d680926bb8302c0b2faba1295e72f49fd3a1c306af21e51ebad5f644

Observation a6857368-c41a-444e-9b2e-866c399a234a · inbound

JADES: A Universal Framework for Jailbreak Assessment via Decompositional Scoring cites this paper.

JADES: A Universal Framework for Jailbreak Assessment via Decompositional Scoring Exploiting Programmatic Behavior of LLMs: Dual-Use Through Standard Security Attacks

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-05T14:51:03.693056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T14:51:03.693056Z digest=sha256:bca800c7aa330e80ffd03745c731aff0e4e8e3ab087f36a9a57e266f56b1f614

Observation 7f0c4c94-9fea-4cec-b91e-ed6b9cb57b4e · inbound

Breaking to Build: A Threat Model of Prompt-Based Attacks for Securing LLMs cites this paper.

Breaking to Build: A Threat Model of Prompt-Based Attacks for Securing LLMs Exploiting Programmatic Behavior of LLMs: Dual-Use Through Standard Security Attacks

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-05T05:59:27.941873Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T05:59:27.941873Z digest=sha256:8aaa1b74af1e7b3979605e2870549df13e45de041c3061869d3ec216ce860074

Observation 3540126a-5c94-4b96-a78e-565f1ab7fa6f · inbound

Anchoring Refusal Direction: Mitigating Safety Risks in Tuning via Projection Constraint cites this paper.

Anchoring Refusal Direction: Mitigating Safety Risks in Tuning via Projection Constraint Exploiting Programmatic Behavior of LLMs: Dual-Use Through Standard Security Attacks

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-04T23:10:21.297410Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T23:10:21.297410Z digest=sha256:847eec5b775f19d808652b6060c3361bcf5972247a88a260d863b4b037fc026c

Observation efbff674-42ba-4754-b406-ece032bb95c6 · inbound

MoGU V2: Toward a Higher Pareto Frontier Between Model Usability and Security cites this paper.

MoGU V2: Toward a Higher Pareto Frontier Between Model Usability and Security Exploiting Programmatic Behavior of LLMs: Dual-Use Through Standard Security Attacks

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-04T23:09:41.524212Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T23:09:41.524212Z digest=sha256:98e0d973bde9221e6ed22ecc31999d1b546cd10acc7b913729902cedeb00e255

Observation d508c59e-88d6-4476-ab2e-39b72eba5cac · inbound

Stop Testing Attacks, Start Diagnosing Defenses: The Four-Checkpoint Framework Reveals Where LLM Safety Breaks cites this paper.

Stop Testing Attacks, Start Diagnosing Defenses: The Four-Checkpoint Framework Reveals Where LLM Safety Breaks Exploiting Programmatic Behavior of LLMs: Dual-Use Through Standard Security Attacks

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-03T02:51:56.253042Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:51:56.253042Z digest=sha256:320c22bf2b298e6e4f0712e1abcaeb2a63007e601a724593dcbfe99734cfd240

Observation 1160e7af-8cad-404a-a8de-db28bcebb7d5 · inbound

An Empirical Security Evaluation of LLM-Generated Cryptographic Rust Code cites this paper.

An Empirical Security Evaluation of LLM-Generated Cryptographic Rust Code Exploiting Programmatic Behavior of LLMs: Dual-Use Through Standard Security Attacks

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-12T08:56:26.794119Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-07T13:26:17.100399Z digest=sha256:47587315421ec59983a44565adb11b3eef578e207a9aa4b67f22fa7d1893df44

Observation ed14cc08-8635-43ca-add9-8f1e6044f166 · inbound

Block-wise Codeword Embedding for Reliable Multi-bit Text Watermarking cites this paper.

Block-wise Codeword Embedding for Reliable Multi-bit Text Watermarking Exploiting Programmatic Behavior of LLMs: Dual-Use Through Standard Security Attacks

Reference 58

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T15:31:06.043863Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-09T19:52:21.350072Z digest=sha256:0e3dfdee6947f9cdbc7120e968f91b6a4273333e4f3b5a4e9a0cdf059458be8c

Observation a2877a87-885f-4c60-b65a-24eb260debe9 · inbound

Speculative Decoding at Temperature Zero: A Scoped Safety-Invariance Screen with a 48,072-Sample Expansion cites this paper.

Speculative Decoding at Temperature Zero: A Scoped Safety-Invariance Screen with a 48,072-Sample Expansion Exploiting Programmatic Behavior of LLMs: Dual-Use Through Standard Security Attacks

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-07-04T16:59:58.296057Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-26T00:05:06.647295Z digest=sha256:6c60e0841276b362557be2aaf27293938c5c114877e42e9257606e02abc74d2e

Observation a0791066-8569-4077-912c-6a9276d55ff4 · inbound

Security--Fidelity Tradeoffs: The Hidden Cost of Prompt Injection Defense cites this paper.

Security--Fidelity Tradeoffs: The Hidden Cost of Prompt Injection Defense Exploiting Programmatic Behavior of LLMs: Dual-Use Through Standard Security Attacks

Reference 107

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T12:45:44.893381Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-07-01T01:44:07.700127Z digest=sha256:01c4a275b599f0d91fc111e10a83863cadbf19187b48ee934abd620e1b1170e4

Observation 6b1f1121-dc29-424d-8748-8d7fa42f2719 · inbound

A Lifecycle and Application-Stack Survey of Large Language Model Vulnerabilities: Attacks, Risks, Defenses, and Open Problems cites this paper.

A Lifecycle and Application-Stack Survey of Large Language Model Vulnerabilities: Attacks, Risks, Defenses, and Open Problems Exploiting Programmatic Behavior of LLMs: Dual-Use Through Standard Security Attacks

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-07-01T11:05:42.347271Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-01T04:44:23.543728Z digest=sha256:8ae940a7baa086b71f01627b739d25249bffb18b3e19cbbce700c8ed5e916dd8

Observation 9d2ee81b-6ff7-4404-86db-19c084bdfc09 · inbound

RoguePrompt: Dual-Layer Encoding for Self-Reconstruction to Circumvent LLM Moderation cites this paper.

RoguePrompt: Dual-Layer Encoding for Self-Reconstruction to Circumvent LLM Moderation Exploiting Programmatic Behavior of LLMs: Dual-Use Through Standard Security Attacks

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-01T08:37:33.396857Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T08:37:33.396857Z digest=sha256:98a9fb559d52f777291e488c4da54786635eff80269c7fb0c98d0b06aa6b6fa7