Pith. sign in

Paper Citation Record · LEDGER

A Comprehensive Study of Jailbreak Attack versus Defense for Large Language Models

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 13 inbound Pith citation observations for arXiv:2402.13457.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2402.13457 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 13 of 13 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 13 of 13 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-09T17:58:57.361973Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T19:36:08.540565Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 96816ab0-ea0e-4e03-b2d8-bd1b244f6442 · inbound

Refusal in Language Models Is Mediated by a Single Direction cites this paper.

Refusal in Language Models Is Mediated by a Single Direction A Comprehensive Study of Jailbreak Attack versus Defense for Large Language Models

Reference 201

Resolution
verified exact
arxiv_id, observed 2026-05-13T10:47:56.160539Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-13T10:47:55.934081Z digest=sha256:55b43177a666395646bd7ca414dec6e0b185fd9105af3ee2b238c553023f1bbf

Observation 5028a698-6630-4eff-907b-08d67c8a250f · inbound

`Do as I say not as I do': A Semi-Automated Approach for Jailbreak Prompt Attack against Multimodal LLMs cites this paper.

`Do as I say not as I do': A Semi-Automated Approach for Jailbreak Prompt Attack against Multimodal LLMs A Comprehensive Study of Jailbreak Attack versus Defense for Large Language Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-09T17:58:57.361973Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T17:58:57.361973Z digest=sha256:9d64d4cda6b12a293611a21c94316a7558bd2a1c3453754815a0779657f0dc0b

Observation ba93dbd7-895a-4e70-a74b-d4b06a2bf240 · inbound

Bridging the Safety Gap: A Guardrail Pipeline for Trustworthy LLM Inferences cites this paper.

Bridging the Safety Gap: A Guardrail Pipeline for Trustworthy LLM Inferences A Comprehensive Study of Jailbreak Attack versus Defense for Large Language Models

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-08T10:23:06.284811Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T10:23:06.284811Z digest=sha256:222a5f5c28e9a2df414d08a6d8021ac9cdf74c769be685ece10bcdc491c0624a

Observation 4ec1c18d-206d-45a6-ad1b-233c7e70c5d4 · inbound

Quality-Diversity Red-Teaming: Automated Generation of High-Quality and Diverse Attackers for Large Language Models cites this paper.

Quality-Diversity Red-Teaming: Automated Generation of High-Quality and Diverse Attackers for Large Language Models A Comprehensive Study of Jailbreak Attack versus Defense for Large Language Models

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-07T05:47:06.503637Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:47:06.503637Z digest=sha256:6e277d6ca924fe99c5573d0e615b19d28808de3a4707259f81fb1ce1d49e3c83

Observation c3576dbb-ce6f-4b77-ad91-e23a35fc89f5 · inbound

Understanding the Supply Chain and Risks of Large Language Model Applications cites this paper.

Understanding the Supply Chain and Risks of Large Language Model Applications A Comprehensive Study of Jailbreak Attack versus Defense for Large Language Models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T14:45:42.912963Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:45:42.912963Z digest=sha256:92bc4941deeb2efffeb521330e1e06f396a8a2d9e83f16150da1044d6d5eedc9

Observation 25f5f323-7136-4f16-b93b-e827df8a1a61 · inbound

A Real-Time, Self-Tuning Moderator Framework for Adversarial Prompt Detection cites this paper.

A Real-Time, Self-Tuning Moderator Framework for Adversarial Prompt Detection A Comprehensive Study of Jailbreak Attack versus Defense for Large Language Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-05T22:24:23.361494Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T22:24:23.361494Z digest=sha256:7c57a091bc91f9745eff806a63163deb95194d920abc424edeacf01da26a4984

Observation 9c20c9f8-8f70-4da7-a200-a906231ac9c9 · inbound

Layer-Wise Perturbations via Sparse Autoencoders for Adversarial Text Generation cites this paper.

Layer-Wise Perturbations via Sparse Autoencoders for Adversarial Text Generation A Comprehensive Study of Jailbreak Attack versus Defense for Large Language Models

Reference 137

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:45.398903Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:45.398903Z digest=sha256:3375ce29c156957d05f735faf83a74237dc1a41d0088d8003b39c0e0a7cae09a

Observation 5f48ed78-a3b3-4fbc-8a64-39d31cbf9d26 · inbound

ORFuzz: Fuzzing the "Other Side" of LLM Safety -- Testing Over-Refusal cites this paper.

ORFuzz: Fuzzing the "Other Side" of LLM Safety -- Testing Over-Refusal A Comprehensive Study of Jailbreak Attack versus Defense for Large Language Models

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-18T23:31:54.591943Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-18T23:27:52.438709Z digest=sha256:3a05ec2b79e6b34076a20a4aa4b94dfe45e5f011568458bddefb014bf6a197f1

Observation ead34801-dcd9-434d-b7cf-10580f06e3e0 · inbound

MoGU V2: Toward a Higher Pareto Frontier Between Model Usability and Security cites this paper.

MoGU V2: Toward a Higher Pareto Frontier Between Model Usability and Security A Comprehensive Study of Jailbreak Attack versus Defense for Large Language Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-04T23:09:41.435223Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T23:09:41.435223Z digest=sha256:87df5f16227dc00063f2606709b47657b74079af3b33da7b654f2cf0b78ef979

Observation aa9e6cc6-b697-400f-9c2e-fd2873f19520 · inbound

Paladin: Defending LLM-enabled Phishing Emails with a New Trigger-Tag Paradigm cites this paper.

Paladin: Defending LLM-enabled Phishing Emails with a New Trigger-Tag Paradigm A Comprehensive Study of Jailbreak Attack versus Defense for Large Language Models

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-04T22:33:28.168620Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T22:33:28.168620Z digest=sha256:5d0d3275afa64b26b55b79b575315b8eaf4c2bbed361e6535781ff1840961a3c

Observation f8a6ee22-da7a-4884-9f88-a3c118714fda · inbound

ADAM: A Systematic Data Extraction Attack on Agent Memory via Adaptive Querying cites this paper.

ADAM: A Systematic Data Extraction Attack on Agent Memory via Adaptive Querying A Comprehensive Study of Jailbreak Attack versus Defense for Large Language Models

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-11T05:45:57.609447Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T17:57:59.918483Z digest=sha256:f629db534ac6c58b8b42e0b295f7e7e38514f721993df4c6ab13601c595c9afc

Observation ca2742c6-95f2-45f1-8ee9-319bba9a33bb · inbound

The Salami Slicing Threat: Exploiting Cumulative Risks in LLM Systems cites this paper.

The Salami Slicing Threat: Exploiting Cumulative Risks in LLM Systems A Comprehensive Study of Jailbreak Attack versus Defense for Large Language Models

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-11T09:16:04.374380Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T16:07:31.602378Z digest=sha256:de07f96bc4b65b5bbe22ea4d23670597e88aeb4f58c04471010cd0a999c857fa

Observation e46e7c62-8614-458f-9e27-838cb4dc1f6d · inbound

Persona Attack: Incremental Memory Injection Jailbreak Attack against Large Language Models cites this paper.

Persona Attack: Incremental Memory Injection Jailbreak Attack against Large Language Models A Comprehensive Study of Jailbreak Attack versus Defense for Large Language Models

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:36:08.542160Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-28T22:23:00.894059Z digest=sha256:82c0b41abac4c96327d7d8178c15791ef66e0363b042baa8a2d6149d99ef4772