Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T21:05:18.310292Z
Paper Citation Record · LEDGER
As of 12 August 2026, this Paper Citation Record lists 18 of 18 outbound references and 5 inbound Pith citation observations for arXiv:2411.09125.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T21:05:18.310292Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-06T23:20:57.184523Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
18 of 18 outbound references displayed
External citation measurements
0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
Observation 84f5bb80-e193-458a-b6ad-37248f04cbf0 · outbound
DROJ: A Prompt-Driven Attack against Large Language Models GPT-4 Technical Report
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 499b9844-f69f-4404-bd0a-06424cd433fd · outbound
DROJ: A Prompt-Driven Attack against Large Language Models Foundational Challenges in Assuring Alignment and Safety of Large Language Models
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2b5f6412-8e9b-4feb-88ad-3adbe95040a8 · outbound
DROJ: A Prompt-Driven Attack against Large Language Models Catastrophic Jailbreak of Open-source LLMs via Exploiting Generation
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 97883a66-11dd-4ed0-9fe8-aef432d90b6c · outbound
DROJ: A Prompt-Driven Attack against Large Language Models Baseline Defenses for Adversarial Attacks Against Aligned Language Models
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 55177f3b-2436-43cc-b623-78593b39ce49 · outbound
DROJ: A Prompt-Driven Attack against Large Language Models Mistral 7B
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3748be3e-42c4-4602-8de9-3c06ec7fa0e9 · outbound
DROJ: A Prompt-Driven Attack against Large Language Models Open Sesame! Universal Black Box Jailbreaking of Large Language Models
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6d0b4eaa-7234-4610-8c9b-989234954785 · outbound
DROJ: A Prompt-Driven Attack against Large Language Models AutoDAN: Generating Stealthy Jailbreak Prompts on Aligned Large Language Models
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 76410ec2-e056-4511-b470-3b16c89a2efb · outbound
DROJ: A Prompt-Driven Attack against Large Language Models ASETF: A Novel Method for Jailbreak Attack on LLMs through Translate Suffix Embeddings
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 12eb1332-d419-40c4-bf66-1a17822cbae8 · outbound
DROJ: A Prompt-Driven Attack against Large Language Models Finetuned Language Models Are Zero-Shot Learners
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6381b773-e8c9-48c4-92ef-f99986a5ad55 · outbound
DROJ: A Prompt-Driven Attack against Large Language Models Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 17bb32fe-3f83-446e-9564-6c6cd57c1830 · outbound
DROJ: A Prompt-Driven Attack against Large Language Models GPT-4 Is Too Smart To Be Safe: Stealthy Chat with LLMs via Cipher
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ba0aabf4-8c01-4579-94de-e61e0ca19bfa · outbound
DROJ: A Prompt-Driven Attack against Large Language Models On Prompt-Driven Safeguarding for Large Language Models
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 77a4ff95-2fe2-4112-9023-64d5aff29bb1 · outbound
DROJ: A Prompt-Driven Attack against Large Language Models Universal and Transferable Adversarial Attacks on Aligned Language Models
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1b99365c-887d-447b-9138-53138c6114d3 · outbound
DROJ: A Prompt-Driven Attack against Large Language Models Gradient-based Adversarial Attacks against Text Transformers
Reference 2017
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 243d5db0-569a-4077-b852-58297066fc7d · outbound
DROJ: A Prompt-Driven Attack against Large Language Models LLM Self Defense: By Self Examination, LLMs Know They Are Being Tricked
Reference 2021
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7c429825-233c-474c-bd26-21d3cf617a2e · outbound
DROJ: A Prompt-Driven Attack against Large Language Models Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cfaa11b8-3421-401a-adc5-b3e4c3b2b82f · outbound
DROJ: A Prompt-Driven Attack against Large Language Models Detecting Language Model Attacks with Perplexity
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fca8b031-e28e-4df9-96ad-b4dd1afe537c · outbound
DROJ: A Prompt-Driven Attack against Large Language Models Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4340b1d7-49c0-436b-84dc-94c912fa73fd · inbound
SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models DROJ: A Prompt-Driven Attack against Large Language Models
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a5be9d95-d140-4449-9c6f-9a341cc5624a · inbound
On Surjectivity of Neural Networks: Can you elicit any behavior from your model? DROJ: A Prompt-Driven Attack against Large Language Models
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a308273d-44f2-4fa2-939c-cbd9a7e61673 · inbound
LLM in the Middle: A Systematic Review of Threats and Mitigations to Real-World LLM-based Systems DROJ: A Prompt-Driven Attack against Large Language Models
Reference 126
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dab38a09-c02a-471d-ac78-c674820b964f · inbound
Probing the Difficulty Perception Mechanism of Large Language Models DROJ: A Prompt-Driven Attack against Large Language Models
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5c385afd-db4c-42fe-b7bf-baee5a6e493e · inbound
Prompt Governance? On Governing Technologies Governed by Natural Language DROJ: A Prompt-Driven Attack against Large Language Models
Reference 137
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.