Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T21:07:54.305218Z
Paper Citation Record · LEDGER
As of 21 August 2026, this Paper Citation Record lists 23 of 23 outbound references and 3 inbound Pith citation observations for arXiv:2505.10838.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T21:07:54.305218Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-04T09:40:45.239429Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
23 of 23 outbound references displayed
External citation measurements
0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
Observation 4886c8f0-3c39-4604-b996-ca36dc7c5c2a · outbound
LARGO: Latent Adversarial Reflection through Gradient Optimization for Jailbreaking LLMs Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fc74c2c6-b7e9-48bb-9577-8dd5e944dbde · outbound
LARGO: Latent Adversarial Reflection through Gradient Optimization for Jailbreaking LLMs Jailbreaking Black Box Large Language Models in Twenty Queries
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0558ed9f-d215-4fd7-88f2-9fa78cf01c43 · outbound
LARGO: Latent Adversarial Reflection through Gradient Optimization for Jailbreaking LLMs JailbreakBench: An Open Robustness Benchmark for Jailbreaking Large Language Models
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 12c35f7c-00c6-4303-8b2b-ad5f15a485d5 · outbound
LARGO: Latent Adversarial Reflection through Gradient Optimization for Jailbreaking LLMs h4rm3l: A language for Composable Jailbreak Attack Synthesis
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation afd0c072-8f98-45e5-950e-adae69160f5d · outbound
LARGO: Latent Adversarial Reflection through Gradient Optimization for Jailbreaking LLMs The Ethics of Interaction: Mitigating Security Threats in LLMs
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1d9d3cf5-292e-417d-addf-6855bbb2cd63 · outbound
LARGO: Latent Adversarial Reflection through Gradient Optimization for Jailbreaking LLMs Safety Layers in Aligned Large Language Models: The Key to LLM Security
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b2a8cdaf-56a9-4f85-9777-daeac4b660b5 · outbound
LARGO: Latent Adversarial Reflection through Gradient Optimization for Jailbreaking LLMs do anything now
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 4d4dd1ba-e018-459d-861e-1072918808dd · outbound
LARGO: Latent Adversarial Reflection through Gradient Optimization for Jailbreaking LLMs Latentqa: Teaching llms to decode activations into natural language
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e6972e7f-d34c-480e-8fb5-a6679c73ba4e · outbound
LARGO: Latent Adversarial Reflection through Gradient Optimization for Jailbreaking LLMs AdvPrompter: Fast Adaptive Adversarial Prompting for LLMs
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 13e823af-61a1-4d3a-8fcd-f0d0f6006a17 · outbound
LARGO: Latent Adversarial Reflection through Gradient Optimization for Jailbreaking LLMs Xstest: A test suite for identifying exaggerated safety behaviours in large language models
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 56603d5d-99cb-42b3-bd00-823844663d00 · outbound
LARGO: Latent Adversarial Reflection through Gradient Optimization for Jailbreaking LLMs Code Llama: Open Foundation Models for Code
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 560e6ea2-ec0d-478a-8667-281a510804c1 · outbound
LARGO: Latent Adversarial Reflection through Gradient Optimization for Jailbreaking LLMs Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ada4356b-fe2b-4e22-b20a-127aa1d8ce32 · outbound
LARGO: Latent Adversarial Reflection through Gradient Optimization for Jailbreaking LLMs do anything now
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation df76b44f-9918-48b2-9a4b-4d133ffe3090 · outbound
LARGO: Latent Adversarial Reflection through Gradient Optimization for Jailbreaking LLMs A StrongREJECT for Empty Jailbreaks
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d3d04e62-d5d7-455a-93bf-b027e88c856d · outbound
LARGO: Latent Adversarial Reflection through Gradient Optimization for Jailbreaking LLMs CodeGemma: Open Code Models Based on Gemma
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4c3798b9-88a7-4d5d-860d-032fbfe94e9e · outbound
LARGO: Latent Adversarial Reflection through Gradient Optimization for Jailbreaking LLMs Gemini: A Family of Highly Capable Multimodal Models
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8be58ad4-c78b-41cd-8468-6e7ef2e02092 · outbound
LARGO: Latent Adversarial Reflection through Gradient Optimization for Jailbreaking LLMs Llama 2: Open Foundation and Fine-Tuned Chat Models
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 26789c46-4c3d-493b-bfd0-528e5ef4a00e · outbound
LARGO: Latent Adversarial Reflection through Gradient Optimization for Jailbreaking LLMs GPTFUZZER: Red Teaming Large Language Models with Auto-Generated Jailbreak Prompts
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 79d9aee2-1d1a-4180-a597-2876a0025a83 · outbound
LARGO: Latent Adversarial Reflection through Gradient Optimization for Jailbreaking LLMs Diversity Helps Jailbreak Large Language Models
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3e84eb57-77db-4962-bb35-0430c05f9b81 · outbound
LARGO: Latent Adversarial Reflection through Gradient Optimization for Jailbreaking LLMs Universal and Transferable Adversarial Attacks on Aligned Language Models
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f4ae84b9-2f5c-4a1f-937f-4762ea740237 · outbound
LARGO: Latent Adversarial Reflection through Gradient Optimization for Jailbreaking LLMs Guard: Role-playing to generate natural-language jailbreakings to test guideline adherence of large language models
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 86f0df2b-8e00-4c46-9f38-42f629633b82 · outbound
LARGO: Latent Adversarial Reflection through Gradient Optimization for Jailbreaking LLMs Jailbreaking Leading Safety-Aligned LLMs with Simple Adaptive Attacks
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 70f1671c-982a-4930-a56b-a59c05ffe13e · outbound
LARGO: Latent Adversarial Reflection through Gradient Optimization for Jailbreaking LLMs Detecting Language Model Attacks with Perplexity
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 96604676-e465-415a-8d7b-14ad7da08157 · inbound
Echoes of Human Malice in Agents: Benchmarking LLMs for Multi-Turn Online Harassment Attacks LARGO: Latent Adversarial Reflection through Gradient Optimization for Jailbreaking LLMs
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3ac00727-fbd8-40b5-9a4f-c40f5e92ecdf · inbound
REALISTA: Realistic Latent Adversarial Attacks that Elicit LLM Hallucinations LARGO: Latent Adversarial Reflection through Gradient Optimization for Jailbreaking LLMs
Reference 75
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 3511bb0f-cefa-4b94-808f-cd4bb9419a47 · inbound
LASH: Adaptive Semantic Hybridization for Black-Box Jailbreaking of Large Language Models LARGO: Latent Adversarial Reflection through Gradient Optimization for Jailbreaking LLMs
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.