Pith. sign in

Paper Citation Record · LEDGER

LARGO: Latent Adversarial Reflection through Gradient Optimization for Jailbreaking LLMs

As of 21 August 2026, this Paper Citation Record lists 23 of 23 outbound references and 3 inbound Pith citation observations for arXiv:2505.10838.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.10838 v1

Coverage vector

measured 23 of 23 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T21:07:54.305218Z

measured 26 of 26 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-04T09:40:45.239429Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

23 of 23 outbound references displayed

  • verified exact0
  • verified fuzzy3
  • unresolved20
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation 4886c8f0-3c39-4604-b996-ca36dc7c5c2a · outbound

This paper cites Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone.

LARGO: Latent Adversarial Reflection through Gradient Optimization for Jailbreaking LLMs Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-15T21:07:54.219516Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:07:54.219516Z digest=sha256:936573e07464a83083e8086b51e588b7a4940fbf3e1cade82ab1b8835754b49a

Observation fc74c2c6-b7e9-48bb-9577-8dd5e944dbde · outbound

This paper cites Jailbreaking Black Box Large Language Models in Twenty Queries.

LARGO: Latent Adversarial Reflection through Gradient Optimization for Jailbreaking LLMs Jailbreaking Black Box Large Language Models in Twenty Queries

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-15T21:07:54.232733Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:07:54.232733Z digest=sha256:09c2867b7b7f8098fd12f9933125dd1c056b7d006deebaf6b6b0bdf8ee278906

Observation 0558ed9f-d215-4fd7-88f2-9fa78cf01c43 · outbound

This paper cites JailbreakBench: An Open Robustness Benchmark for Jailbreaking Large Language Models.

LARGO: Latent Adversarial Reflection through Gradient Optimization for Jailbreaking LLMs JailbreakBench: An Open Robustness Benchmark for Jailbreaking Large Language Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-15T21:07:54.237043Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:07:54.237043Z digest=sha256:87d14546bcb9013cdd7969c74c4d651cc95edb05fd127a8508db5a047b2dc277

Observation 12c35f7c-00c6-4303-8b2b-ad5f15a485d5 · outbound

This paper cites h4rm3l: A language for Composable Jailbreak Attack Synthesis.

LARGO: Latent Adversarial Reflection through Gradient Optimization for Jailbreaking LLMs h4rm3l: A language for Composable Jailbreak Attack Synthesis

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-15T21:07:54.241085Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:07:54.241085Z digest=sha256:f6b3b77353611a7ba1a9a917948c50d1e1c0df8626c8d73d8bb03d5e17de948b

Observation afd0c072-8f98-45e5-950e-adae69160f5d · outbound

This paper cites The Ethics of Interaction: Mitigating Security Threats in LLMs.

LARGO: Latent Adversarial Reflection through Gradient Optimization for Jailbreaking LLMs The Ethics of Interaction: Mitigating Security Threats in LLMs

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T21:07:54.249235Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:07:54.249235Z digest=sha256:24b016bf0b96eebf29f81c2715ecce58297feac537e96d50b24ca836c7a15986

Observation 1d9d3cf5-292e-417d-addf-6855bbb2cd63 · outbound

This paper cites Safety Layers in Aligned Large Language Models: The Key to LLM Security.

LARGO: Latent Adversarial Reflection through Gradient Optimization for Jailbreaking LLMs Safety Layers in Aligned Large Language Models: The Key to LLM Security

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T21:07:54.253022Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:07:54.253022Z digest=sha256:e947266792e6b683d3a555b0913ed1e26e3c0c166d9db341ba555b41cd38bde1

Observation b2a8cdaf-56a9-4f85-9777-daeac4b660b5 · outbound

This paper cites do anything now.

LARGO: Latent Adversarial Reflection through Gradient Optimization for Jailbreaking LLMs do anything now

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:07:54.683112Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T21:07:54.256701Z digest=sha256:07b04a347bbd208685b0a018421e6d449ee13e2c4564c515fb8de5e6d8ac826c

Observation 4d4dd1ba-e018-459d-861e-1072918808dd · outbound

This paper cites Latentqa: Teaching llms to decode activations into natural language.

LARGO: Latent Adversarial Reflection through Gradient Optimization for Jailbreaking LLMs Latentqa: Teaching llms to decode activations into natural language

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-15T21:07:54.260409Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:07:54.260409Z digest=sha256:56f0023ecbb97752289ac1403c0380a243f412489bc24f80fbe777c536f2a147

Observation e6972e7f-d34c-480e-8fb5-a6679c73ba4e · outbound

This paper cites AdvPrompter: Fast Adaptive Adversarial Prompting for LLMs.

LARGO: Latent Adversarial Reflection through Gradient Optimization for Jailbreaking LLMs AdvPrompter: Fast Adaptive Adversarial Prompting for LLMs

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-15T21:07:54.263941Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:07:54.263941Z digest=sha256:2f1b147c349fdaf81d452320cd9090c5e898fd619c6f4959e1e27557de363718

Observation 13e823af-61a1-4d3a-8fcd-f0d0f6006a17 · outbound

This paper cites Xstest: A test suite for identifying exaggerated safety behaviours in large language models.

LARGO: Latent Adversarial Reflection through Gradient Optimization for Jailbreaking LLMs Xstest: A test suite for identifying exaggerated safety behaviours in large language models

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:07:54.673175Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T21:07:54.267903Z digest=sha256:c7ce541a28d37f104917891b93b8788a4464260a994b9e9d88f22572518e1e82

Observation 56603d5d-99cb-42b3-bd00-823844663d00 · outbound

This paper cites Code Llama: Open Foundation Models for Code.

LARGO: Latent Adversarial Reflection through Gradient Optimization for Jailbreaking LLMs Code Llama: Open Foundation Models for Code

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-15T21:07:54.271485Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:07:54.271485Z digest=sha256:b4186a56ea1a3fd2b129e7701978474b110d67e4cbc73eff7fc311c07a6d94ec

Observation 560e6ea2-ec0d-478a-8667-281a510804c1 · outbound

This paper cites Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation.

LARGO: Latent Adversarial Reflection through Gradient Optimization for Jailbreaking LLMs Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-15T21:07:54.275044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:07:54.275044Z digest=sha256:161898f51f58a7bf4406b5886973deee6322d854994aef2e7362a8802e33674c

Observation ada4356b-fe2b-4e22-b20a-127aa1d8ce32 · outbound

This paper cites do anything now.

LARGO: Latent Adversarial Reflection through Gradient Optimization for Jailbreaking LLMs do anything now

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:07:54.661845Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T21:07:54.278787Z digest=sha256:10b5234472234b5705ed955f1a9c5b46b9a9414bf249956e3c7a47e18bbaf700

Observation df76b44f-9918-48b2-9a4b-4d133ffe3090 · outbound

This paper cites A StrongREJECT for Empty Jailbreaks.

LARGO: Latent Adversarial Reflection through Gradient Optimization for Jailbreaking LLMs A StrongREJECT for Empty Jailbreaks

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-15T21:07:54.282513Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:07:54.282513Z digest=sha256:871980a757295c042d37f0226aa0f2b147d9af201facbde6741718ca82505f21

Observation d3d04e62-d5d7-455a-93bf-b027e88c856d · outbound

This paper cites CodeGemma: Open Code Models Based on Gemma.

LARGO: Latent Adversarial Reflection through Gradient Optimization for Jailbreaking LLMs CodeGemma: Open Code Models Based on Gemma

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-15T21:07:54.286434Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:07:54.286434Z digest=sha256:efab03f00b7b973b33c1e1df6ff5428c22f48027cd913d319636bad613dab9ba

Observation 4c3798b9-88a7-4d5d-860d-032fbfe94e9e · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

LARGO: Latent Adversarial Reflection through Gradient Optimization for Jailbreaking LLMs Gemini: A Family of Highly Capable Multimodal Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-15T21:07:54.290345Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:07:54.290345Z digest=sha256:cba56e47e36a6762ebfae336f7ec1f4a406e25435c461500850e11df7d152cf2

Observation 8be58ad4-c78b-41cd-8468-6e7ef2e02092 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

LARGO: Latent Adversarial Reflection through Gradient Optimization for Jailbreaking LLMs Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-15T21:07:54.293852Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:07:54.293852Z digest=sha256:413b7f3e156dd5b4ce664a61642650b230e7434d0ca55dfd9fc0cb08df050073

Observation 26789c46-4c3d-493b-bfd0-528e5ef4a00e · outbound

This paper cites GPTFUZZER: Red Teaming Large Language Models with Auto-Generated Jailbreak Prompts.

LARGO: Latent Adversarial Reflection through Gradient Optimization for Jailbreaking LLMs GPTFUZZER: Red Teaming Large Language Models with Auto-Generated Jailbreak Prompts

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-15T21:07:54.297752Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:07:54.297752Z digest=sha256:b58f11e41604390f12a87ac7dda5ee9740930b761a6d3efa59eb92d15791330c

Observation 79d9aee2-1d1a-4180-a597-2876a0025a83 · outbound

This paper cites Diversity Helps Jailbreak Large Language Models.

LARGO: Latent Adversarial Reflection through Gradient Optimization for Jailbreaking LLMs Diversity Helps Jailbreak Large Language Models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-15T21:07:54.301502Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:07:54.301502Z digest=sha256:58403c860969b4d3170a008a77b5ba1df80d287c52f1b809c978b0ac144b58eb

Observation 3e84eb57-77db-4962-bb35-0430c05f9b81 · outbound

This paper cites Universal and Transferable Adversarial Attacks on Aligned Language Models.

LARGO: Latent Adversarial Reflection through Gradient Optimization for Jailbreaking LLMs Universal and Transferable Adversarial Attacks on Aligned Language Models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-15T21:07:54.305218Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:07:54.305218Z digest=sha256:436750ab88df6fa76e9c6959edca40a9c897ad773b1e39afc979e843c7df8819

Observation f4ae84b9-2f5c-4a1f-937f-4762ea740237 · outbound

This paper cites Guard: Role-playing to generate natural-language jailbreakings to test guideline adherence of large language models.

LARGO: Latent Adversarial Reflection through Gradient Optimization for Jailbreaking LLMs Guard: Role-playing to generate natural-language jailbreakings to test guideline adherence of large language models

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-15T21:07:54.245356Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:07:54.245356Z digest=sha256:c322b94c949ff8878733001883c4a6a6044ead175845a0636a5608a6d754a457

Observation 86f0df2b-8e00-4c46-9f38-42f629633b82 · outbound

This paper cites Jailbreaking Leading Safety-Aligned LLMs with Simple Adaptive Attacks.

LARGO: Latent Adversarial Reflection through Gradient Optimization for Jailbreaking LLMs Jailbreaking Leading Safety-Aligned LLMs with Simple Adaptive Attacks

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-15T21:07:54.228190Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:07:54.228190Z digest=sha256:34216c35b476d3eb3d584845b76c74fa90c9f8c6d8ee3549264752cb745c588c

Observation 70f1671c-982a-4930-a56b-a59c05ffe13e · outbound

This paper cites Detecting Language Model Attacks with Perplexity.

LARGO: Latent Adversarial Reflection through Gradient Optimization for Jailbreaking LLMs Detecting Language Model Attacks with Perplexity

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-15T21:07:54.223932Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:07:54.223932Z digest=sha256:662ddafd3ab3a33e4fc3bf68d87f3f38643ab0358e4200cf42bcdce7ee9a4596

Pith citing papers

Observation 96604676-e465-415a-8d7b-14ad7da08157 · inbound

Echoes of Human Malice in Agents: Benchmarking LLMs for Multi-Turn Online Harassment Attacks cites this paper.

Echoes of Human Malice in Agents: Benchmarking LLMs for Multi-Turn Online Harassment Attacks LARGO: Latent Adversarial Reflection through Gradient Optimization for Jailbreaking LLMs

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-04T09:40:45.239429Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:40:45.239429Z digest=sha256:443dcb0875137894064aa739cf1ca9574d191bdbf243e471948b87b8124e4c93

Observation 3ac00727-fbd8-40b5-9a4f-c40f5e92ecdf · inbound

REALISTA: Realistic Latent Adversarial Attacks that Elicit LLM Hallucinations cites this paper.

REALISTA: Realistic Latent Adversarial Attacks that Elicit LLM Hallucinations LARGO: Latent Adversarial Reflection through Gradient Optimization for Jailbreaking LLMs

Reference 75

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T20:17:56.372026Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-05-14T20:13:10.814899Z digest=sha256:44a07d0ee3b9fc0589d43f51c8ddd96b9546be3465b4c7d905a0fa5364011ec6

Observation 3511bb0f-cefa-4b94-808f-cd4bb9419a47 · inbound

LASH: Adaptive Semantic Hybridization for Black-Box Jailbreaking of Large Language Models cites this paper.

LASH: Adaptive Semantic Hybridization for Black-Box Jailbreaking of Large Language Models LARGO: Latent Adversarial Reflection through Gradient Optimization for Jailbreaking LLMs

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-21T04:49:35.512077Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-21T04:48:19.926845Z digest=sha256:9329c45d79135a09d1d05813eb84995f5ac1909e4ba76772d5f93856094bc175