Pith. sign in

Paper Citation Record · LEDGER

Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers

As of 10 August 2026, this Paper Citation Record lists 43 of 43 outbound references and 0 inbound Pith citation observations for arXiv:2507.13474.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.13474 v1

Coverage vector

measured 43 of 43 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T16:28:54.067424Z

measured 43 of 43 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

43 of 43 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved43
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 3b31598d-7fe3-4c30-ba6c-7772bfb40a2a · outbound

This paper cites online" 'onlinestring :=.

Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers online" 'onlinestring :=

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:48.640057Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:28:48.640057Z digest=sha256:c53e88c28059a5d5b223b9ec486f1053fa6542dd198b9a6f614ee2d4dc5bbb42

Observation c84bf8c5-bd24-48eb-aab1-e0dff72f0958 · outbound

This paper cites write newline.

Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:48.703778Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:28:48.703778Z digest=sha256:c3df642c213e1e6faa8a26b82bffedcb02adf306f6958d77f4c36064a03cdf66

Observation 236e6eb3-81f5-4e6e-978d-4b077f0e229e · outbound

This paper cites GPT-4 Technical Report.

Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers GPT-4 Technical Report

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:48.869873Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:28:48.869873Z digest=sha256:a105f6e8079972cc801b8ad6ab6cd5428d9800081242881fc217de5c948bb180

Observation f97162dd-5eda-44e3-879a-59456895e55f · outbound

This paper cites an unresolved cited work.

Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers Unresolved cited work

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:48.945367Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:28:48.945367Z digest=sha256:ace4ee834bbbcb39b9d53c50469b973ea11f9361218f8283465fb94343dfa48d

Observation 503487ad-c806-4465-9cce-6055d4bc2ff3 · outbound

This paper cites an unresolved cited work.

Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-06T16:28:56.244127Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T16:28:49.077604Z digest=sha256:978b5268d295f55d798428dc1d2a1887d513a183f26afa1f8e0d2f29ac01c059

Observation 347a9bda-f834-4fa3-9990-35be56a60160 · outbound

This paper cites Detecting Language Model Attacks with Perplexity.

Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers Detecting Language Model Attacks with Perplexity

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:49.185819Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:28:49.185819Z digest=sha256:717e89395c88416c4ae68b9b761f19fdfa8cc0742c8274222ba904d55b195c3d

Observation 956382b7-d9de-498f-af1b-1c214190361f · outbound

This paper cites an unresolved cited work.

Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-06T16:28:56.102262Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T16:28:49.271940Z digest=sha256:f44556b4df602b48c07a0695e9c294e3b8eb7820b39eca300f4885537c3c38ea

Observation c265d1b5-3cfa-44a5-8c4d-ca081e1e73a2 · outbound

This paper cites an unresolved cited work.

Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-06T16:28:55.933928Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T16:28:49.402906Z digest=sha256:60664da2e9cfeae9b10e38ab6059ca2321d18aa6f2d2312df554178ed4c5767c

Observation c758b55e-9df9-4e38-8fd8-acaba6266b51 · outbound

This paper cites Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback.

Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:49.574566Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:28:49.574566Z digest=sha256:006c500a211ffc4dfd1bf343a543ef4fc9d139921ffe80ce80eea8ac0021e497

Observation 8c179423-b704-42b2-a5df-bfe92fa6186e · outbound

This paper cites an unresolved cited work.

Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-06T16:28:55.742936Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T16:28:49.675620Z digest=sha256:ab91be5e2742eb008cd9da143cc5951b8be5263e46bad1ec1768955c25bf105a

Observation 4b1863eb-3d29-49fb-a8e3-40e0256476e6 · outbound

This paper cites Defending Against Alignment-Breaking Attacks via Robustly Aligned LLM.

Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers Defending Against Alignment-Breaking Attacks via Robustly Aligned LLM

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:49.811945Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:28:49.811945Z digest=sha256:be0c5aed2397ff434ea3bda4ec524bef1fdca7e7adcaa2fe0365ca13ca400939

Observation 8e1abffa-02be-423c-8057-9d245034e607 · outbound

This paper cites JailbreakBench: An Open Robustness Benchmark for Jailbreaking Large Language Models.

Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers JailbreakBench: An Open Robustness Benchmark for Jailbreaking Large Language Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:49.917562Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:28:49.917562Z digest=sha256:23e6622b9fd719fba74fdbf31b9b025abaccaed98b85783c0c050414f5a77c3b

Observation d41ab251-a115-42fd-a0eb-452ae08eb9bb · outbound

This paper cites Jailbreaking Black Box Large Language Models in Twenty Queries.

Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers Jailbreaking Black Box Large Language Models in Twenty Queries

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:50.068705Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:28:50.068705Z digest=sha256:36579df22616f4ce93b235a582e42c1afc8f984f7a4bb5394bd3e783128e33d6

Observation 91bf5cca-2805-46c2-ab92-d0bbc040feeb · outbound

This paper cites an unresolved cited work.

Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers Unresolved cited work

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:50.163671Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:28:50.163671Z digest=sha256:ae0ddafe8785bc412b64607677ab689edc2cfdbec98adbec5cd9028808f98039

Observation 3c10ce6e-2b23-484f-84b6-923160c54a4a · outbound

This paper cites Build it Break it Fix it for Dialogue Safety: Robustness from Adversarial Human Attack.

Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers Build it Break it Fix it for Dialogue Safety: Robustness from Adversarial Human Attack

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:50.247767Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:28:50.247767Z digest=sha256:ca6c52691791798c86347f3c834e330fce7fb4b52be7da9d56f251c4a2e04310

Observation 794e5e78-c718-48ad-a1c2-d83518a3af0d · outbound

This paper cites MART: Improving LLM Safety with Multi-round Automatic Red-Teaming.

Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers MART: Improving LLM Safety with Multi-round Automatic Red-Teaming

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:50.312168Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:28:50.312168Z digest=sha256:53ceb23f91db7fcccfeb3adcff1a264ba0a6884dc3827fdcac0c65f15b422a33

Observation 7e5b5baf-6437-4902-bdf0-c3a83b46c85d · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:50.420937Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:28:50.420937Z digest=sha256:fe5b2f465fe1842c1621e67aab8b2c866e3b74703dba2e0eeb9fab0126175cf1

Observation 210b3c37-d214-4c21-9fc6-bc16dd1be8f5 · outbound

This paper cites Llama Guard: LLM-based Input-Output Safeguard for Human-AI Conversations.

Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers Llama Guard: LLM-based Input-Output Safeguard for Human-AI Conversations

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:50.565875Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:28:50.565875Z digest=sha256:3d20ae8e6ccf7a1d0c5f3a618b636e227e77e88ebb849ac65a546d19683e5272

Observation b8bb6335-c7c8-4ec4-a734-8546c72d084a · outbound

This paper cites an unresolved cited work.

Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers Unresolved cited work

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:50.707380Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:28:50.707380Z digest=sha256:d631a003d17d6466f18486ef5f171bc3f4c3a85d0750ce704d8e421b6b46c83b

Observation a3a0305a-3e57-4a1d-846e-50db95490588 · outbound

This paper cites ArtPrompt: ASCII Art-based Jailbreak Attacks against Aligned LLMs.

Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers ArtPrompt: ASCII Art-based Jailbreak Attacks against Aligned LLMs

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:50.842462Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:28:50.842462Z digest=sha256:398ef07664f0f10c6ceab0a473ec72f3ab1bf865f3a1441918d3fc502f4a0b6a

Observation c1cbee39-d902-4178-b7e0-cc4dbfe6d414 · outbound

This paper cites an unresolved cited work.

Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-08-06T16:28:55.556499Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T16:28:50.971616Z digest=sha256:f7fa78200965288656a7243bfd772fc4b108ede76b978f8db6f8bdb4926e574a

Observation ed845802-21d3-4881-823e-906235ed37bc · outbound

This paper cites an unresolved cited work.

Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-06T16:28:55.373960Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T16:28:51.101716Z digest=sha256:f0380897edbe7a789a4c16f70b8d1451d403cb44a50455d39774433ec104ea25

Observation 896367b5-f986-4cfa-a302-74d89d9b855b · outbound

This paper cites DeepInception: Hypnotize Large Language Model to Be Jailbreaker.

Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers DeepInception: Hypnotize Large Language Model to Be Jailbreaker

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:51.241644Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:28:51.241644Z digest=sha256:7d467edc69a76fc77408fac756d5cd6e92337cd152ecd19519b836262556b425

Observation db972186-50fe-40df-b387-53a0fc9a0b50 · outbound

This paper cites an unresolved cited work.

Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-06T16:28:55.201958Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T16:28:51.394418Z digest=sha256:e1a3209776fe858e54d180adef1ec9fd41c2cac837b44eac70429d07e93ad8b0

Observation 611b438e-ffb8-4e54-92a8-7169dbd60563 · outbound

This paper cites AutoDAN: Generating Stealthy Jailbreak Prompts on Aligned Large Language Models.

Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers AutoDAN: Generating Stealthy Jailbreak Prompts on Aligned Large Language Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:51.548057Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:28:51.548057Z digest=sha256:f47e940c54643160487b92df01da35c17952c8acb5346bdc0b4954e9c669429f

Observation 1f306486-9e95-46ee-b8b1-049df3a4c5b3 · outbound

This paper cites an unresolved cited work.

Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-06T16:28:55.013080Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T16:28:51.725752Z digest=sha256:e85e830e2debbd252e49411d0aef6de46bdf6a647cf22d773e638978d95ff23a

Observation 4305db1e-0e1f-4968-af51-2eadbb8c14f9 · outbound

This paper cites Large Language Models: A Survey.

Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers Large Language Models: A Survey

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:51.817938Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:28:51.817938Z digest=sha256:f44f8a67204a7d81a74ddee2d82e5ae8a0716dce0653768208b3555ca24d9542

Observation 5721b676-1875-4f95-9a51-a4b439eac072 · outbound

This paper cites an unresolved cited work.

Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-06T16:28:54.822625Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T16:28:51.990606Z digest=sha256:2a808a7f49c41dd5e80d29e27f7878885bd32d1c912bed8a828ac41309d1b1a4

Observation 7e535475-329a-4045-b5de-6caa0787094f · outbound

This paper cites an unresolved cited work.

Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers Unresolved cited work

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:52.144439Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:28:52.144439Z digest=sha256:6d7cfce1beab0c3cf274436591f3fec313903d521ff1310daecf6924c89dc949

Observation ae2ceb50-aa3a-48a1-8e3f-19eb9ac38a2b · outbound

This paper cites an unresolved cited work.

Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-08-06T16:28:54.642109Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T16:28:52.240065Z digest=sha256:e5373fae6074bc28849b7dbaf9cee6bac10c7f3db32c426d249f8aadfc83b421

Observation fbf7b2ac-b8c5-4d43-bd13-dc47513b9fa4 · outbound

This paper cites Fine-tuning Aligned Language Models Compromises Safety, Even When Users Do Not Intend To!.

Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers Fine-tuning Aligned Language Models Compromises Safety, Even When Users Do Not Intend To!

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:52.360605Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:28:52.360605Z digest=sha256:b345d8868ef48091c22595919c4a2cb7733c419ca6fe2e14f3d35969ae1149bf

Observation c75b6dba-86ac-4d4b-bb3e-a676cd000403 · outbound

This paper cites an unresolved cited work.

Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers Unresolved cited work

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:52.493755Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:28:52.493755Z digest=sha256:4f22022bb2f1baf8642a1d95157871092807f29c555798919bf2ce17b348d34c

Observation f382b434-c798-4cb9-b749-214729b7cb5d · outbound

This paper cites SmoothLLM: Defending Large Language Models Against Jailbreaking Attacks.

Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers SmoothLLM: Defending Large Language Models Against Jailbreaking Attacks

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:52.637985Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:28:52.637985Z digest=sha256:cd9c6a9a8906ae504b2b5cde5aa424405ad734e3cdd353488d91e2465a0f9c1d

Observation d2e988a5-2a3d-4bb2-ace0-6d9abfe43f2a · outbound

This paper cites "Do Anything Now": Characterizing and Evaluating In-The-Wild Jailbreak Prompts on Large Language Models.

Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers "Do Anything Now": Characterizing and Evaluating In-The-Wild Jailbreak Prompts on Large Language Models

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:52.739133Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:28:52.739133Z digest=sha256:1f3a0c66e4b94704c2157135cf78e77faf99b1e9bf6abaedce99642e48e66be6

Observation df9224f7-33de-4564-b168-9ec5247baa30 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:52.931076Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:28:52.931076Z digest=sha256:bc08fb78886ae4e08279815ed553c2e1c76d8fb92f64177ccaf83f892460932f

Observation 3e363830-d5e3-45cf-be3e-9474bdea96c6 · outbound

This paper cites an unresolved cited work.

Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers Unresolved cited work

Reference 36

Resolution
unresolved
raw_fallback, observed 2026-08-06T16:28:54.416687Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T16:28:53.032268Z digest=sha256:5a0ad85f689ce63fb0f70589c606c013e73e2f0956df33b6d8b75402c8e95806

Observation e2115a4e-2767-4e58-a3ca-b627164c590e · outbound

This paper cites Ethical and social risks of harm from Language Models.

Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers Ethical and social risks of harm from Language Models

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:53.214889Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:28:53.214889Z digest=sha256:f1e338aaacb22f4a66c415495183641ced9725d3a7818dd9ec1bc742e68e477d

Observation cf851105-cb94-44f7-b922-4c0c5690b70d · outbound

This paper cites Jailbreak Attacks and Defenses Against Large Language Models: A Survey.

Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers Jailbreak Attacks and Defenses Against Large Language Models: A Survey

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:53.364399Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:28:53.364399Z digest=sha256:8c90e1f39844db6ed34b496500252e265242ae0e905dfd4746ef88f01e4b6145

Observation a938959c-9abd-4204-8835-ff1968fc4e5d · outbound

This paper cites GPT-4 Is Too Smart To Be Safe: Stealthy Chat with LLMs via Cipher.

Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers GPT-4 Is Too Smart To Be Safe: Stealthy Chat with LLMs via Cipher

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:53.451033Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:28:53.451033Z digest=sha256:737a2dfc2f598cadfe6d77c341e6b99c6c6881b596f0a911478d83cbc3f92ef4

Observation 4fd72723-3dd6-45be-8ceb-8ac186636391 · outbound

This paper cites How Johnny Can Persuade LLMs to Jailbreak Them: Rethinking Persuasion to Challenge AI Safety by Humanizing LLMs.

Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers How Johnny Can Persuade LLMs to Jailbreak Them: Rethinking Persuasion to Challenge AI Safety by Humanizing LLMs

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:53.616114Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:28:53.616114Z digest=sha256:e81bfa4ee198bd0aa5c9c39460433702c6e5d5556d3ae534ce03f70a6a8a4e93

Observation ff35eece-cd72-4c90-8265-2c4068223693 · outbound

This paper cites JailGuard: A Universal Detection Framework for LLM Prompt-based Attacks.

Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers JailGuard: A Universal Detection Framework for LLM Prompt-based Attacks

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:53.745955Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:28:53.745955Z digest=sha256:1a263c17b13cf1f7e1272eded442741897add1e4adc96d5c4ab98bdecaea216e

Observation 51d00b66-1e49-4927-b73b-35dfe637e7e1 · outbound

This paper cites How Alignment and Jailbreak Work: Explain LLM Safety through Intermediate Hidden States.

Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers How Alignment and Jailbreak Work: Explain LLM Safety through Intermediate Hidden States

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:53.895617Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:28:53.895617Z digest=sha256:6e7c546f5fe9f2b0d4113a53e92b4675cdf4f7799c7fc774094a139656f60d68

Observation 6b688f2f-1418-4268-8b2c-054e0d455c8b · outbound

This paper cites Universal and Transferable Adversarial Attacks on Aligned Language Models.

Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers Universal and Transferable Adversarial Attacks on Aligned Language Models

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:54.067424Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:28:54.067424Z digest=sha256:663e0bb7f7b416bc12eec8a9d716405460ed653dc7dbdf3dcdb3ae586a9036b4

Pith citing papers

No inbound Pith citation observations are available.