Pith. sign in

Paper Citation Record · LEDGER

Visual-RolePlay: Universal Jailbreak Attack on MultiModal Large Language Models via Role-playing Image Character

As of 12 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 30 inbound Pith citation observations for arXiv:2405.20773.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2405.20773 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 30 of 30 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 30 of 30 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-11T10:32:28.295755Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation e3b24ad5-4c12-46d1-ba32-413f9cc03941 · inbound

Divide and Conquer: A Hybrid Strategy Defeats Multimodal Large Language Models cites this paper.

Divide and Conquer: A Hybrid Strategy Defeats Multimodal Large Language Models Visual-RolePlay: Universal Jailbreak Attack on MultiModal Large Language Models via Role-playing Image Character

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-11T10:32:28.295755Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T10:32:28.295755Z digest=sha256:a66a5049d8e39c42b5c572b449cc3c47939c95df299f42b5e47af18fabce7d8f

Observation d06010ab-dcd7-464f-9472-a6daabaff6c6 · inbound

Mirage in the Eyes: Hallucination Attack on Multi-modal Large Language Models with Only Attention Sink cites this paper.

Mirage in the Eyes: Hallucination Attack on Multi-modal Large Language Models with Only Attention Sink Visual-RolePlay: Universal Jailbreak Attack on MultiModal Large Language Models via Role-playing Image Character

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-10T14:29:28.074482Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:29:28.074482Z digest=sha256:20f6e6d44c06202f6dabbcad9711dd1d7f98e3f75c78195bf950fc53e7bbe9cf

Observation 11d35ee0-45f9-4f58-9bdd-aea793577148 · inbound

Safety at Scale: A Comprehensive Survey of Large Model and Agent Safety cites this paper.

Safety at Scale: A Comprehensive Survey of Large Model and Agent Safety Visual-RolePlay: Universal Jailbreak Attack on MultiModal Large Language Models via Role-playing Image Character

Reference 283

Resolution
verified exact
arxiv_id, observed 2026-05-23T04:42:34.086175Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-23T04:39:04.591722Z digest=sha256:bd6620150a41c8f6f19003623114cf0e735e79da61c23733714ba92aa3e6777b

Observation 631adf9a-90c7-43e1-ae22-9610b9e7c1ea · inbound

When Data Manipulation Meets Attack Goals: An In-depth Survey of Attacks for VLMs cites this paper.

When Data Manipulation Meets Attack Goals: An In-depth Survey of Attacks for VLMs Visual-RolePlay: Universal Jailbreak Attack on MultiModal Large Language Models via Role-playing Image Character

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-08T15:38:17.089320Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T15:38:17.089320Z digest=sha256:0de52d7ecd966cae40c14541d84ca0df80d36c88f583fc1d43dec8fd2dd63638

Observation f7fb9fc0-da95-4b8a-b745-96f5d8022a25 · inbound

Universal Adversarial Attack on Aligned Multimodal LLMs cites this paper.

Universal Adversarial Attack on Aligned Multimodal LLMs Visual-RolePlay: Universal Jailbreak Attack on MultiModal Large Language Models via Role-playing Image Character

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-08T11:17:01.654636Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T11:17:01.654636Z digest=sha256:cbea8936455ecc03a39a8c9fef41601a8b531fef160544200d4fa8b83b79d2e3

Observation 2983a16f-e1b5-4204-a268-c3055cc3eced · inbound

A Survey of Safety on Large Vision-Language Models: Attacks, Defenses and Evaluations cites this paper.

A Survey of Safety on Large Vision-Language Models: Attacks, Defenses and Evaluations Visual-RolePlay: Universal Jailbreak Attack on MultiModal Large Language Models via Role-playing Image Character

Reference 105

Resolution
unresolved
no resolver link, observed 2026-08-07T19:45:19.478474Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T19:45:19.478474Z digest=sha256:a8a49faaae4ecfcc77fc71c8b0bfd00b162923a9300f14e66cfab952574b4bb8

Observation 8e654e81-4eeb-4fc4-9474-7c49d04a16fd · inbound

RedDiffuser: Auditing Multimodal Safety Failures in Vision-Language Models via Reinforced Diffusion cites this paper.

RedDiffuser: Auditing Multimodal Safety Failures in Vision-Language Models via Reinforced Diffusion Visual-RolePlay: Universal Jailbreak Attack on MultiModal Large Language Models via Role-playing Image Character

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-23T00:15:14.824301Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-23T00:13:08.603115Z digest=sha256:46831e0003acba574c4bbc91d147243a3ff8cf4f0e9aae0cd29e9fd1fb55208b

Observation db1cdf8a-433a-4028-a163-02d76706a3b6 · inbound

Backdoor Cleaning without External Guidance in MLLM Fine-tuning cites this paper.

Backdoor Cleaning without External Guidance in MLLM Fine-tuning Visual-RolePlay: Universal Jailbreak Attack on MultiModal Large Language Models via Role-playing Image Character

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T14:56:53.654176Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:56:53.654176Z digest=sha256:0062d38f2f1356aec35bf1a933f5be50f15f6b9636b8a8f2ccb19ea8b1233ee4

Observation 8e8cf2df-b10e-489a-ac96-bc15025f4f03 · inbound

Towards medical AI misalignment: a preliminary study cites this paper.

Towards medical AI misalignment: a preliminary study Visual-RolePlay: Universal Jailbreak Attack on MultiModal Large Language Models via Role-playing Image Character

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T14:52:15.961141Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:52:15.961141Z digest=sha256:1d38bd7ab9db7259a8c983b60d25756a916d6e324ee79d7a4057888191c409cb

Observation 175db9d6-728f-4140-a7eb-cb7830d3e6b1 · inbound

Seeing the Threat: Vulnerabilities in Vision-Language Models to Adversarial Attack cites this paper.

Seeing the Threat: Vulnerabilities in Vision-Language Models to Adversarial Attack Visual-RolePlay: Universal Jailbreak Attack on MultiModal Large Language Models via Role-playing Image Character

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T13:22:01.801888Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:22:01.801888Z digest=sha256:db92315e32710bfe8ca76be8329f6527df0545976684405b81715e4476cbc148

Observation c2d64ec6-2de9-4e5a-b235-bc0f4f94626a · inbound

Pushing the Limits of Safety: A Technical Report on the ATLAS Challenge 2025 cites this paper.

Pushing the Limits of Safety: A Technical Report on the ATLAS Challenge 2025 Visual-RolePlay: Universal Jailbreak Attack on MultiModal Large Language Models via Role-playing Image Character

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T00:57:13.747840Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:57:13.747840Z digest=sha256:21fbd0557dd3f941d1963d6af111108a86e7815b198aa25a4449fdc29538d398

Observation b9874aa5-95cf-4190-b522-76878372fe42 · inbound

Guardians and Offenders: A Survey on Harmful Content Generation and Safety Mitigation of LLM cites this paper.

Guardians and Offenders: A Survey on Harmful Content Generation and Safety Mitigation of LLM Visual-RolePlay: Universal Jailbreak Attack on MultiModal Large Language Models via Role-playing Image Character

Reference 184

Resolution
unresolved
no resolver link, observed 2026-08-05T23:13:05.101293Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:13:05.101293Z digest=sha256:6185cb5cb4071f0cf8f9b79cd6c4da8196b345467c316074fa77adb34b0746c5

Observation 7427c66f-c0ee-49c1-b825-77a9ac21efee · inbound

The First Differentiable Transfer-Based Algorithm for Discrete MicroLED Repair cites this paper.

The First Differentiable Transfer-Based Algorithm for Discrete MicroLED Repair Visual-RolePlay: Universal Jailbreak Attack on MultiModal Large Language Models via Role-playing Image Character

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-05T22:21:02.811865Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T22:21:02.811865Z digest=sha256:e00dea0807fdb74a1f2a4e0e2f551d4baf32b750e6e392f142ae5a2c92959f47

Observation f0fab43f-50a9-42b8-9577-5e536a723db5 · inbound

Blockchain Network Analysis using Quantum Inspired Graph Neural Networks & Ensemble Models cites this paper.

Blockchain Network Analysis using Quantum Inspired Graph Neural Networks & Ensemble Models Visual-RolePlay: Universal Jailbreak Attack on MultiModal Large Language Models via Role-playing Image Character

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-05T21:22:37.649627Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T21:22:37.649627Z digest=sha256:4a2027f450e5dfe970acff90da22d49fceb461d09e5e14092f5c052176fb7bcb

Observation 507b2e82-a73e-4034-8334-7fa17ee9e1e8 · inbound

Empowering Multimodal LLMs with External Tools: A Comprehensive Survey cites this paper.

Empowering Multimodal LLMs with External Tools: A Comprehensive Survey Visual-RolePlay: Universal Jailbreak Attack on MultiModal Large Language Models via Role-playing Image Character

Reference 231

Resolution
unresolved
no resolver link, observed 2026-08-05T20:29:06.482070Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:29:06.482070Z digest=sha256:6e86f2c7e97534765ed93187aec6ff686566c02b14b93d96196be5915b8b5dae

Observation 672e6be8-72ed-474c-b237-579cc9d71b9e · inbound

On Surjectivity of Neural Networks: Can you elicit any behavior from your model? cites this paper.

On Surjectivity of Neural Networks: Can you elicit any behavior from your model? Visual-RolePlay: Universal Jailbreak Attack on MultiModal Large Language Models via Role-playing Image Character

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-05T16:00:48.746871Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:00:48.746871Z digest=sha256:90fd812ad185d0cff34802c0ffa3e3852ba8ecbf63b0989f0917db8209e856d2

Observation d5be7c2a-2861-4dfc-a5c9-102b6a476f91 · inbound

SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses cites this paper.

SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Visual-RolePlay: Universal Jailbreak Attack on MultiModal Large Language Models via Role-playing Image Character

Reference 133

Resolution
unresolved
no resolver link, observed 2026-08-04T09:25:50.007530Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:25:50.007530Z digest=sha256:73c402f5edccc09f3dc25487ab79d17e92f14f178cfbbfab8174214cf05ea1de

Observation 59cd7582-ec3a-4e5a-9246-11362bc47eb0 · inbound

VERA-V: Variational Inference Framework for Jailbreaking Vision-Language Models cites this paper.

VERA-V: Variational Inference Framework for Jailbreaking Vision-Language Models Visual-RolePlay: Universal Jailbreak Attack on MultiModal Large Language Models via Role-playing Image Character

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-04T09:02:38.637172Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T09:02:38.637172Z digest=sha256:d26ee0ea228e3ba20addb754861f53881de47162e3e59b0e5cfd3272ace9e7ab

Observation f036b930-a299-47f6-ba77-7ab84fc7cea0 · inbound

Multi-Turn Adaptive Prompting Attack on Large Vision-Language Models cites this paper.

Multi-Turn Adaptive Prompting Attack on Large Vision-Language Models Visual-RolePlay: Universal Jailbreak Attack on MultiModal Large Language Models via Role-playing Image Character

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-02T23:16:36.824638Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:16:36.824638Z digest=sha256:9f091914d4eb3799072aaf1b3cdc5725cc29fc035de37252c17825053d5240e0

Observation f22ac78a-055d-4254-a4df-a8e26b4706be · inbound

Making MLLMs Blind: Adversarial Smuggling Attacks in MLLM Content Moderation cites this paper.

Making MLLMs Blind: Adversarial Smuggling Attacks in MLLM Content Moderation Visual-RolePlay: Universal Jailbreak Attack on MultiModal Large Language Models via Role-playing Image Character

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T00:45:50.274862Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-10T18:19:58.970952Z digest=sha256:05e5897d9d3e587e1233385f2ddf99a631b743489deccb3276b2285c059dad4c

Observation 92aad731-47a5-4fca-8760-fc4a54da3ee5 · inbound

Dictionary-Aligned Concept Control for Safeguarding Multimodal LLMs cites this paper.

Dictionary-Aligned Concept Control for Safeguarding Multimodal LLMs Visual-RolePlay: Universal Jailbreak Attack on MultiModal Large Language Models via Role-playing Image Character

Reference 59

Resolution
verified exact
arxiv_id, observed 2026-05-11T05:35:57.164743Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-10T18:04:05.157103Z digest=sha256:f877cb01420e354685e03fd498eef40d7beeddcef7dce412d61aa76bcbc8d49d

Observation 83a652be-a6d8-4526-9a4e-2fc782028d26 · inbound

Targeted Interpretable Safety Neuron Enhancement for Multilingual Vision-Language Large Models cites this paper.

Targeted Interpretable Safety Neuron Enhancement for Multilingual Vision-Language Large Models Visual-RolePlay: Universal Jailbreak Attack on MultiModal Large Language Models via Role-playing Image Character

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-11T05:20:58.246993Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-10T18:12:34.909229Z digest=sha256:15ada9652a57e18655d7de0008e53ec94ede32793602ef797fce3349f55f4f66

Observation 5e4b0086-e642-4c06-9d0b-98d122104444 · inbound

Targeted Interpretable Safety Neuron Enhancement for Multilingual Vision-Language Large Models cites this paper.

Targeted Interpretable Safety Neuron Enhancement for Multilingual Vision-Language Large Models Visual-RolePlay: Universal Jailbreak Attack on MultiModal Large Language Models via Role-playing Image Character

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-02T16:34:05.343320Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T16:34:05.343320Z digest=sha256:2da23515e97f54696532baecc84c2f4a6d5b21e68fe195d9f828eb33c2f71658

Observation 0ae55dc8-cbca-449b-a83a-7bd0208d6261 · inbound

Through the Lens of Character: Resolving Modality-Role Interference in Multimodal Role-Playing Agent cites this paper.

Through the Lens of Character: Resolving Modality-Role Interference in Multimodal Role-Playing Agent Visual-RolePlay: Universal Jailbreak Attack on MultiModal Large Language Models via Role-playing Image Character

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-12T05:51:27.691230Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-12T04:49:56.954433Z digest=sha256:986acd5b85a1f23b093ec1c94adb5869c96c84df3902d564efe12917f6e30458

Observation c8fc2ab4-fa70-42b6-b1e0-179ecf95f63e · inbound

Localization then Neutralization: Gradient-guided Token Suppression against Visual Prompt Injection Attack cites this paper.

Localization then Neutralization: Gradient-guided Token Suppression against Visual Prompt Injection Attack Visual-RolePlay: Universal Jailbreak Attack on MultiModal Large Language Models via Role-playing Image Character

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T12:24:39.790379Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-06-30T12:19:12.910048Z digest=sha256:297c0f96b6e36d0bb63ae0fcdf672098b60cc50e1c3f3a65c555af0f1429477c

Observation 7a59e5e3-7b94-4e75-b435-471ac28639ca · inbound

MLingualFC: Evaluating Jailbreak Vulnerabilities in Multilingual Vision-Language Models cites this paper.

MLingualFC: Evaluating Jailbreak Vulnerabilities in Multilingual Vision-Language Models Visual-RolePlay: Universal Jailbreak Attack on MultiModal Large Language Models via Role-playing Image Character

Reference 103

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T18:57:16.800737Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-06-27T21:47:22.295896Z digest=sha256:d86da7aa573786c36a5f0d39694e3f15b1f2b9b0b44631686e5d5dc6e5612b8b

Observation 883b4f59-0aa4-434a-8c3e-0dd1f017de4b · inbound

Adversarial Diffusion Across Modalities: A Fusion Survey of Attacks, Defenses, and Evaluation for Text, Vision, and Vision-Language Models cites this paper.

Adversarial Diffusion Across Modalities: A Fusion Survey of Attacks, Defenses, and Evaluation for Text, Vision, and Vision-Language Models Visual-RolePlay: Universal Jailbreak Attack on MultiModal Large Language Models via Role-playing Image Character

Reference 43

Resolution
verified exact
arxiv_id, observed 2026-06-26T04:38:59.119984Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-06-26T04:35:51.583460Z digest=sha256:87bac877f35421e5a791be57566c1dc664fd4a11a12a4dcc4fda6bc19d540cb2

Observation 8d8a2ab6-d404-41d3-971f-ace533de2ab2 · inbound

V-DEAL: Diagnosing Video Safety De-Calibration as an Understanding-Refusal Coupling Failure cites this paper.

V-DEAL: Diagnosing Video Safety De-Calibration as an Understanding-Refusal Coupling Failure Visual-RolePlay: Universal Jailbreak Attack on MultiModal Large Language Models via Role-playing Image Character

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-01T08:23:25.911168Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T08:23:25.911168Z digest=sha256:0c89d251dbfd08e311ef955d32b1b36981bfdaf7086eb0cc34217b55a8949256

Observation 7e46ab36-d070-4706-8c15-0307e586e9e0 · inbound

The Mirage of LLM Guardrails: A Case Study in AI-Assisted Medical Note Manipulation cites this paper.

The Mirage of LLM Guardrails: A Case Study in AI-Assisted Medical Note Manipulation Visual-RolePlay: Universal Jailbreak Attack on MultiModal Large Language Models via Role-playing Image Character

Reference 6

Resolution
unresolved
no resolver link, observed 2026-07-30T20:22:15.367808Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T20:22:15.367808Z digest=sha256:8f3877d6b96bd1c2c5b57ee1b93d71836762688515bbb11adcc63ad52ed121f2

Observation 07a487c6-9b7a-4608-b51e-29d8e149ba68 · inbound

One Anchor for All: Unified Multilingual and Multimodal Safety Alignment for LVLMs cites this paper.

One Anchor for All: Unified Multilingual and Multimodal Safety Alignment for LVLMs Visual-RolePlay: Universal Jailbreak Attack on MultiModal Large Language Models via Role-playing Image Character

Reference 35

Resolution
unresolved
no resolver link, observed 2026-07-31T23:07:39.941483Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-31T23:07:39.941483Z digest=sha256:bc48a0c7a7d4a31abf0460d26fed65bf13d55ffbdd80aa5997895fcb57d27fda