Pith. sign in

Paper Citation Record · LEDGER

Visual Adversarial Examples Jailbreak Aligned Large Language Models

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 25 inbound Pith citation observations for arXiv:2306.13213.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2306.13213 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 25 of 25 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 25 of 25 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T11:17:01.672767Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

13
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 08fdb8bf-d39d-44ee-8847-db4b383435d5 · inbound

Universal Adversarial Attack on Aligned Multimodal LLMs cites this paper.

Universal Adversarial Attack on Aligned Multimodal LLMs Visual Adversarial Examples Jailbreak Aligned Large Language Models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-08T11:17:01.672767Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T11:17:01.672767Z digest=sha256:7d2c68797774a6b3688233b064e2ac0a4b10b23713126069bd171fdbd8a6c92c

Observation c26d4dcb-c1a4-43df-a860-0fe687fcd6f6 · inbound

VLM-Guard: Safeguarding Vision-Language Models via Fulfilling Safety Alignment Gap cites this paper.

VLM-Guard: Safeguarding Vision-Language Models via Fulfilling Safety Alignment Gap Visual Adversarial Examples Jailbreak Aligned Large Language Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T19:46:40.707864Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T19:46:40.707864Z digest=sha256:b3e852e55e12b419387db22b303d889997112736e8d67b031dd7aea43d10cbc2

Observation 65b0250c-a2b7-456a-a846-6ac6b5dc2f96 · inbound

Towards an AI co-scientist cites this paper.

Towards an AI co-scientist Visual Adversarial Examples Jailbreak Aligned Large Language Models

Reference 16

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T13:02:44.356995Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-11T13:02:43.571234Z digest=sha256:b12715ada27918cc4ad6af96e446beff297ef3e983c98f848f150c0ed6f22f1a

Observation b8bad99f-9b0a-4851-beef-152ef6c85924 · inbound

Seeing the Threat: Vulnerabilities in Vision-Language Models to Adversarial Attack cites this paper.

Seeing the Threat: Vulnerabilities in Vision-Language Models to Adversarial Attack Visual Adversarial Examples Jailbreak Aligned Large Language Models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T13:22:02.206932Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:22:02.206932Z digest=sha256:43d011fd696f2cc0c9a48017f51414fd98be0314eb92914155cb3fc196480a56

Observation a4393d1e-3329-4552-94e9-39b145ddd14c · inbound

Bootstrapping LLM Robustness for VLM Safety via Reducing the Pretraining Modality Gap cites this paper.

Bootstrapping LLM Robustness for VLM Safety via Reducing the Pretraining Modality Gap Visual Adversarial Examples Jailbreak Aligned Large Language Models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T12:37:21.846602Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:37:21.846602Z digest=sha256:45c6484d9957525421171c7e3ab460ca5b545d79506102e5fc89fd5e095ce45f

Observation cfa67e39-094b-494c-8967-ee361b543bc3 · inbound

From Hallucinations to Jailbreaks: Rethinking the Vulnerability of Large Foundation Models cites this paper.

From Hallucinations to Jailbreaks: Rethinking the Vulnerability of Large Foundation Models Visual Adversarial Examples Jailbreak Aligned Large Language Models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T12:34:35.624649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:34:35.624649Z digest=sha256:cb73652d96d4c8d63c312048f60515b57d429b1ed64e467271929925a4d8975a

Observation aaec619b-8e1a-4d76-ad0a-f774bf657377 · inbound

AMIA: Automatic Masking and Joint Intention Analysis Makes LVLMs Robust Jailbreak Defenders cites this paper.

AMIA: Automatic Masking and Joint Intention Analysis Makes LVLMs Robust Jailbreak Defenders Visual Adversarial Examples Jailbreak Aligned Large Language Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:27.915330Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:27.915330Z digest=sha256:bceafd82e33617b3767b5ccb822344b0bec0cec47d38239448b3c1c63066f552

Observation 925b90e8-dbdf-43e8-b2ee-94e35ed6c82c · inbound

Pushing the Limits of Safety: A Technical Report on the ATLAS Challenge 2025 cites this paper.

Pushing the Limits of Safety: A Technical Report on the ATLAS Challenge 2025 Visual Adversarial Examples Jailbreak Aligned Large Language Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T00:57:13.759973Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:57:13.759973Z digest=sha256:f29980df785b785875f0b20b5fc2751089804ed7abcc1cd2f3b3388acd62ce9c

Observation b1632b72-81db-4f0f-861a-6e8f9ece354b · inbound

Bridging the Gap in Vision Language Models in Identifying Unsafe Concepts Across Modalities cites this paper.

Bridging the Gap in Vision Language Models in Identifying Unsafe Concepts Across Modalities Visual Adversarial Examples Jailbreak Aligned Large Language Models

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-06T17:21:35.290281Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:21:35.290281Z digest=sha256:b39e36b86708b16d877025530ab2a9fd323721204d2aa8adabf8fc9bca9bde79

Observation 9964697c-2f9a-48fe-8bac-26af82655bea · inbound

Layer-Wise Perturbations via Sparse Autoencoders for Adversarial Text Generation cites this paper.

Layer-Wise Perturbations via Sparse Autoencoders for Adversarial Text Generation Visual Adversarial Examples Jailbreak Aligned Large Language Models

Reference 109

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:43.040164Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:43.040164Z digest=sha256:ae5036b207a7d856aa97b13ca23ea46d24f0a2649e04dfeb7b056f324c282d52

Observation 2a5973cb-a02f-4be9-a475-36a9752362cd · inbound

Activation Steering Meets Preference Optimization: Defense Against Jailbreaks in Vision Language Models cites this paper.

Activation Steering Meets Preference Optimization: Defense Against Jailbreaks in Vision Language Models Visual Adversarial Examples Jailbreak Aligned Large Language Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-05T13:44:56.261344Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T13:44:56.261344Z digest=sha256:39a1ae2d266077b1a1d869e544861902258de91795a14ffe727f900d6692d17b

Observation 9a351771-dc91-40a1-969e-0b6c8e39baf7 · inbound

VISOR++: Universal Visual Inputs based Steering for Large Vision Language Models cites this paper.

VISOR++: Universal Visual Inputs based Steering for Large Vision Language Models Visual Adversarial Examples Jailbreak Aligned Large Language Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-04T13:45:27.025943Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T13:45:27.025943Z digest=sha256:89a167ce274f6df06b8795dbee60b7ecb244db967552e5a70519f4dc9d62d9f2

Observation 531ffd92-d8cd-4423-bc72-0320ad466a2e · inbound

SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses cites this paper.

SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Visual Adversarial Examples Jailbreak Aligned Large Language Models

Reference 144

Resolution
unresolved
no resolver link, observed 2026-08-04T09:25:51.343016Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:25:51.343016Z digest=sha256:f900851d33baffbc97bcd9397477726fb3bdb1b771c9cbf398e0615a5cb10698

Observation f107f72a-e7bd-4c96-9a39-fd1ce2e05c5d · inbound

Toward a Safe Internet of Agents cites this paper.

Toward a Safe Internet of Agents Visual Adversarial Examples Jailbreak Aligned Large Language Models

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-17T03:18:56.863655Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-17T03:16:40.622915Z digest=sha256:e693d584c6b0eca280b920f1fa3ac20a1e4f4efe4d414decd93e3cc03faf3485

Observation 1f9debe2-c222-4d8d-bac2-b0d82b3df792 · inbound

SALLIE: Safeguarding Against Latent Language & Image Exploits cites this paper.

SALLIE: Safeguarding Against Latent Language & Image Exploits Visual Adversarial Examples Jailbreak Aligned Large Language Models

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:30:51.168905Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T19:04:46.426969Z digest=sha256:229ea50fe6e76edeaef5a195352edcefd95cb1f15aae7a132e8a5d6de1635175

Observation a4b2afe4-6b81-4278-9a47-a7b21d37858c · inbound

Gaslight, Gatekeep, V1-V3: Early Visual Cortex Alignment Shields Vision-Language Models from Sycophantic Manipulation cites this paper.

Gaslight, Gatekeep, V1-V3: Early Visual Cortex Alignment Shields Vision-Language Models from Sycophantic Manipulation Visual Adversarial Examples Jailbreak Aligned Large Language Models

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-10T13:10:26.230242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-10T13:09:35.407790Z digest=sha256:00e7bca34e65dcf8b10322997bffe4e5de33c859a86346ba3d94ad06a594b0bc

Observation ec89f14c-30b9-4c5e-aa7e-c98dd706d83b · inbound

VisInject: Disruption != Injection -- A Dual-Dimension Evaluation of Universal Adversarial Attacks on Vision-Language Models cites this paper.

VisInject: Disruption != Injection -- A Dual-Dimension Evaluation of Universal Adversarial Attacks on Vision-Language Models Visual Adversarial Examples Jailbreak Aligned Large Language Models

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-11T16:56:08.117312Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-09T14:24:48.999632Z digest=sha256:55696374d05095b318c1d5f26fc944e644ab6ab7a1fa573213c6f0c239818f2f

Observation 2f395c81-659c-41af-a702-2abee27ca02f · inbound

Catching the Infection Before It Spreads: Foresight-Guided Defense in Multi-Agent Systems cites this paper.

Catching the Infection Before It Spreads: Foresight-Guided Defense in Multi-Agent Systems Visual Adversarial Examples Jailbreak Aligned Large Language Models

Reference 32

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T07:15:12.011445Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T07:11:16.568353Z digest=sha256:022ed23118ac37aedadc01aca0f93cff2c6cbec9bd50b2beae65190e784a4260

Observation c5dcd8a0-1ecd-4bba-ae02-34177f862ff6 · inbound

Single-Sample Black-Box Membership Inference Attack against Vision-Language Models via Cross-modal Semantic Alignment cites this paper.

Single-Sample Black-Box Membership Inference Attack against Vision-Language Models via Cross-modal Semantic Alignment Visual Adversarial Examples Jailbreak Aligned Large Language Models

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-20T14:08:20.903832Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-20T14:07:56.386933Z digest=sha256:56349238ab8c6cd0b5a164d2a96ed800115a08f926b9b1efe043f856b68ffa8b

Observation 95cd0a8a-75c5-4b9a-82fd-0fa8c46dafdd · inbound

VisualLeakBench: Reproducible Action-Boundary Propagation Failures in Vision-Language Agents cites this paper.

VisualLeakBench: Reproducible Action-Boundary Propagation Failures in Vision-Language Agents Visual Adversarial Examples Jailbreak Aligned Large Language Models

Reference 17

Resolution
metadata mismatch
arxiv_id, observed 2026-06-28T23:02:46.377898Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-28T22:59:43.041678Z digest=sha256:754eba6d98ec94c0b6284dd33207686aa04d4ad3b6042c61cad9ecf507ff1222

Observation 91d8210c-4431-484c-9eeb-1075252bf4a4 · inbound

Unveiling Privacy Risks in Multi-modal Large Language Models: Task-specific Vulnerabilities and Mitigation Challenges cites this paper.

Unveiling Privacy Risks in Multi-modal Large Language Models: Task-specific Vulnerabilities and Mitigation Challenges Visual Adversarial Examples Jailbreak Aligned Large Language Models

Reference 100

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T01:27:30.909100Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-27T16:33:28.848573Z digest=sha256:6410d9df270dc72fa7ae0818b4a940672d6c77518ed91ec0c563aa54113d2113

Observation 4bbe5218-121c-47a8-9c65-c0a83595a84b · inbound

Fine-tuning Multi-modal LLMs with ART: Art-based Reinforcement Training cites this paper.

Fine-tuning Multi-modal LLMs with ART: Art-based Reinforcement Training Visual Adversarial Examples Jailbreak Aligned Large Language Models

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-07-03T09:07:48.157454Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-27T10:30:51.490047Z digest=sha256:52163685a438e84bdb5b025172e6faf07e977481aa173644e48c99d1ba941b10

Observation 5e9a01e2-6a59-421f-91a3-95693b5f4c73 · inbound

MIRAGE: Stealthy Visual Prompt Injection for Vulnerability Detection in Web Agents cites this paper.

MIRAGE: Stealthy Visual Prompt Injection for Vulnerability Detection in Web Agents Visual Adversarial Examples Jailbreak Aligned Large Language Models

Reference 18

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T21:08:58.336406Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-27T00:54:00.944706Z digest=sha256:af61f7a3dca928b1ace4358831f6ec42302df1a045fb99e6f1b7ad25abe0156d

Observation 0d447d9b-423f-4ea3-b4c0-762534a77e4a · inbound

Adversarial Diffusion Across Modalities: A Fusion Survey of Attacks, Defenses, and Evaluation for Text, Vision, and Vision-Language Models cites this paper.

Adversarial Diffusion Across Modalities: A Fusion Survey of Attacks, Defenses, and Evaluation for Text, Vision, and Vision-Language Models Visual Adversarial Examples Jailbreak Aligned Large Language Models

Reference 54

Resolution
verified exact
arxiv_id, observed 2026-06-26T04:38:59.046933Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-26T04:35:51.583460Z digest=sha256:2134f358138ce61f06e63805f465d33779f0579053d0ad9e2aea0fe4d57958f3

Observation ec5622d3-6877-43fe-9108-89fb88938cc3 · inbound

Securing Multimodal AI through Internal Information Decomposition cites this paper.

Securing Multimodal AI through Internal Information Decomposition Visual Adversarial Examples Jailbreak Aligned Large Language Models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-02T15:03:00.459707Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T15:03:00.459707Z digest=sha256:83cebf576258394931ecaac4f833d70257bb870811d6b28263492b9139f208e8