Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T16:28:54.067424Z
Paper Citation Record · LEDGER
As of 10 August 2026, this Paper Citation Record lists 43 of 43 outbound references and 0 inbound Pith citation observations for arXiv:2507.13474.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T16:28:54.067424Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
43 of 43 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 3b31598d-7fe3-4c30-ba6c-7772bfb40a2a · outbound
Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers online" 'onlinestring :=
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c84bf8c5-bd24-48eb-aab1-e0dff72f0958 · outbound
Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers write newline
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 236e6eb3-81f5-4e6e-978d-4b077f0e229e · outbound
Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers GPT-4 Technical Report
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f97162dd-5eda-44e3-879a-59456895e55f · outbound
Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers Unresolved cited work
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 503487ad-c806-4465-9cce-6055d4bc2ff3 · outbound
Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers Unresolved cited work
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 347a9bda-f834-4fa3-9990-35be56a60160 · outbound
Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers Detecting Language Model Attacks with Perplexity
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 956382b7-d9de-498f-af1b-1c214190361f · outbound
Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers Unresolved cited work
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation c265d1b5-3cfa-44a5-8c4d-ca081e1e73a2 · outbound
Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers Unresolved cited work
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation c758b55e-9df9-4e38-8fd8-acaba6266b51 · outbound
Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8c179423-b704-42b2-a5df-bfe92fa6186e · outbound
Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers Unresolved cited work
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 4b1863eb-3d29-49fb-a8e3-40e0256476e6 · outbound
Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers Defending Against Alignment-Breaking Attacks via Robustly Aligned LLM
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8e1abffa-02be-423c-8057-9d245034e607 · outbound
Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers JailbreakBench: An Open Robustness Benchmark for Jailbreaking Large Language Models
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d41ab251-a115-42fd-a0eb-452ae08eb9bb · outbound
Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers Jailbreaking Black Box Large Language Models in Twenty Queries
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 91bf5cca-2805-46c2-ab92-d0bbc040feeb · outbound
Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers Unresolved cited work
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3c10ce6e-2b23-484f-84b6-923160c54a4a · outbound
Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers Build it Break it Fix it for Dialogue Safety: Robustness from Adversarial Human Attack
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 794e5e78-c718-48ad-a1c2-d83518a3af0d · outbound
Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers MART: Improving LLM Safety with Multi-round Automatic Red-Teaming
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7e5b5baf-6437-4902-bdf0-c3a83b46c85d · outbound
Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 210b3c37-d214-4c21-9fc6-bc16dd1be8f5 · outbound
Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers Llama Guard: LLM-based Input-Output Safeguard for Human-AI Conversations
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b8bb6335-c7c8-4ec4-a734-8546c72d084a · outbound
Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers Unresolved cited work
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a3a0305a-3e57-4a1d-846e-50db95490588 · outbound
Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers ArtPrompt: ASCII Art-based Jailbreak Attacks against Aligned LLMs
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c1cbee39-d902-4178-b7e0-cc4dbfe6d414 · outbound
Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers Unresolved cited work
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation ed845802-21d3-4881-823e-906235ed37bc · outbound
Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers Unresolved cited work
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 896367b5-f986-4cfa-a302-74d89d9b855b · outbound
Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers DeepInception: Hypnotize Large Language Model to Be Jailbreaker
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation db972186-50fe-40df-b387-53a0fc9a0b50 · outbound
Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers Unresolved cited work
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 611b438e-ffb8-4e54-92a8-7169dbd60563 · outbound
Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers AutoDAN: Generating Stealthy Jailbreak Prompts on Aligned Large Language Models
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1f306486-9e95-46ee-b8b1-049df3a4c5b3 · outbound
Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers Unresolved cited work
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 4305db1e-0e1f-4968-af51-2eadbb8c14f9 · outbound
Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers Large Language Models: A Survey
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5721b676-1875-4f95-9a51-a4b439eac072 · outbound
Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers Unresolved cited work
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 7e535475-329a-4045-b5de-6caa0787094f · outbound
Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers Unresolved cited work
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ae2ceb50-aa3a-48a1-8e3f-19eb9ac38a2b · outbound
Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers Unresolved cited work
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation fbf7b2ac-b8c5-4d43-bd13-dc47513b9fa4 · outbound
Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers Fine-tuning Aligned Language Models Compromises Safety, Even When Users Do Not Intend To!
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c75b6dba-86ac-4d4b-bb3e-a676cd000403 · outbound
Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers Unresolved cited work
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f382b434-c798-4cb9-b749-214729b7cb5d · outbound
Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers SmoothLLM: Defending Large Language Models Against Jailbreaking Attacks
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d2e988a5-2a3d-4bb2-ace0-6d9abfe43f2a · outbound
Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers "Do Anything Now": Characterizing and Evaluating In-The-Wild Jailbreak Prompts on Large Language Models
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation df9224f7-33de-4564-b168-9ec5247baa30 · outbound
Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers Llama 2: Open Foundation and Fine-Tuned Chat Models
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3e363830-d5e3-45cf-be3e-9474bdea96c6 · outbound
Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers Unresolved cited work
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation e2115a4e-2767-4e58-a3ca-b627164c590e · outbound
Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers Ethical and social risks of harm from Language Models
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cf851105-cb94-44f7-b922-4c0c5690b70d · outbound
Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers Jailbreak Attacks and Defenses Against Large Language Models: A Survey
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a938959c-9abd-4204-8835-ff1968fc4e5d · outbound
Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers GPT-4 Is Too Smart To Be Safe: Stealthy Chat with LLMs via Cipher
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4fd72723-3dd6-45be-8ceb-8ac186636391 · outbound
Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers How Johnny Can Persuade LLMs to Jailbreak Them: Rethinking Persuasion to Challenge AI Safety by Humanizing LLMs
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ff35eece-cd72-4c90-8265-2c4068223693 · outbound
Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers JailGuard: A Universal Detection Framework for LLM Prompt-based Attacks
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 51d00b66-1e49-4927-b73b-35dfe637e7e1 · outbound
Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers How Alignment and Jailbreak Work: Explain LLM Safety through Intermediate Hidden States
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6b688f2f-1418-4268-8b2c-054e0d455c8b · outbound
Paper Summary Attack: Jailbreaking LLMs through LLM Safety Papers Universal and Transferable Adversarial Attacks on Aligned Language Models
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.