Pith. sign in

Paper Citation Record · LEDGER

Jailbreaking Large Language Diffusion Models: Revealing Hidden Safety Flaws in Diffusion-Based Text Generation

As of 17 August 2026, this Paper Citation Record lists 51 of 51 outbound references and 12 inbound Pith citation observations for arXiv:2507.19227.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.19227 v1

Coverage vector

measured 51 of 51 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T18:00:26.754277Z

measured 63 of 63 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 12 of 12 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-03T05:45:43.832756Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

51 of 51 outbound references displayed

  • verified exact0
  • verified fuzzy4
  • unresolved47
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation fc0b7c81-b95a-4bbe-bb2d-1b6fa5a66208 · outbound

This paper cites , " * write output.state after.block = add.period write newline.

Jailbreaking Large Language Diffusion Models: Revealing Hidden Safety Flaws in Diffusion-Based Text Generation , " * write output.state after.block = add.period write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-15T18:00:26.542385Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:00:26.542385Z digest=sha256:8380a858766ecf2433ec0196b55ff3ea013da63e671a55d04683b2bde994740f

Observation 45a8c06b-1c6b-4fe5-8761-9b401cf02319 · outbound

This paper cites write newline.

Jailbreaking Large Language Diffusion Models: Revealing Hidden Safety Flaws in Diffusion-Based Text Generation write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-15T18:00:26.547941Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:00:26.547941Z digest=sha256:0671cd0c9b54de910af680123c871fa0136566b4363ff5f54b999cea4a66365f

Observation 71cf21c1-4512-468c-828b-d3d4d0019cca · outbound

This paper cites an unresolved cited work.

Jailbreaking Large Language Diffusion Models: Revealing Hidden Safety Flaws in Diffusion-Based Text Generation Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-15T18:00:27.426993Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T18:00:26.552361Z digest=sha256:554cfe87a643b8a0cf2cf872cd561c4f0f013240d2a3c283a39fbb32b180bb1d

Observation 17de26bc-bf67-4d54-beae-2a91a9560b2e · outbound

This paper cites Structured Denoising Diffusion Models in Discrete State-Spaces.

Jailbreaking Large Language Diffusion Models: Revealing Hidden Safety Flaws in Diffusion-Based Text Generation Structured Denoising Diffusion Models in Discrete State-Spaces

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-15T18:00:26.556286Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:00:26.556286Z digest=sha256:24098b8f8b3b11d063a513b2e9bc83981bfed27ff94f60d0fec7c95521172c30

Observation 485356a7-d5b4-4bde-b6fc-0910fcc7002a · outbound

This paper cites Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback.

Jailbreaking Large Language Diffusion Models: Revealing Hidden Safety Flaws in Diffusion-Based Text Generation Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-15T18:00:26.560593Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:00:26.560593Z digest=sha256:aa65a97db017c4cff5ef2701a59b569feb59d9dde73eac68f9853b5a0de67a2d

Observation fe0c5607-8f46-4185-828e-a6a9db94536c · outbound

This paper cites Safety-Tuned LLaMAs: Lessons From Improving the Safety of Large Language Models that Follow Instructions.

Jailbreaking Large Language Diffusion Models: Revealing Hidden Safety Flaws in Diffusion-Based Text Generation Safety-Tuned LLaMAs: Lessons From Improving the Safety of Large Language Models that Follow Instructions

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-15T18:00:26.564515Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:00:26.564515Z digest=sha256:b79f54a0521ccb97786287443a80ed993a9d11883df61d773d05547c9ade82c3

Observation 7563844b-31a4-4d47-9dfb-e4ae4fffbf3c · outbound

This paper cites Defending Against Alignment-Breaking Attacks via Robustly Aligned LLM.

Jailbreaking Large Language Diffusion Models: Revealing Hidden Safety Flaws in Diffusion-Based Text Generation Defending Against Alignment-Breaking Attacks via Robustly Aligned LLM

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-15T18:00:26.568673Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:00:26.568673Z digest=sha256:0c4bc8b5c03586f9b6cee795341121ffa182cd75768761f94fd9d578ca82e89f

Observation cf71fb20-1bc9-4b55-9f33-1c19a879455f · outbound

This paper cites Jailbreaking Black Box Large Language Models in Twenty Queries.

Jailbreaking Large Language Diffusion Models: Revealing Hidden Safety Flaws in Diffusion-Based Text Generation Jailbreaking Black Box Large Language Models in Twenty Queries

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T18:00:26.572409Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:00:26.572409Z digest=sha256:6f14d208b86e9445a110691b151a88a26be53a06def4934e75c6ef266301ffe7

Observation 721d55f6-cf88-4638-b644-71b3c3ba815b · outbound

This paper cites J.; and Wong, E.

Jailbreaking Large Language Diffusion Models: Revealing Hidden Safety Flaws in Diffusion-Based Text Generation J.; and Wong, E

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:00:27.416031Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T18:00:26.576574Z digest=sha256:b6985fed75a0338fa7dd2d2142f450c0a933075f5ba5b1b4aa6a39009569985b

Observation 8dc1b1fa-79f8-4593-85a8-1bf3dfaee379 · outbound

This paper cites DeepSeek LLM: Scaling Open-Source Language Models with Longtermism.

Jailbreaking Large Language Diffusion Models: Revealing Hidden Safety Flaws in Diffusion-Based Text Generation DeepSeek LLM: Scaling Open-Source Language Models with Longtermism

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-15T18:00:26.580456Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:00:26.580456Z digest=sha256:43e037ea298559822eb052ea965e38d1449d481ea534dbb4011031c78956bcd9

Observation 685a581a-323c-4f0a-8a85-063ef75de1c3 · outbound

This paper cites BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding.

Jailbreaking Large Language Diffusion Models: Revealing Hidden Safety Flaws in Diffusion-Based Text Generation BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-15T18:00:26.584264Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:00:26.584264Z digest=sha256:abd9e39a3937dff843e5aa2e3b64d28c1b36c924d269cc844ebcbecf510a735a

Observation 937eb45a-9584-42df-8789-1bdbe584f99e · outbound

This paper cites an unresolved cited work.

Jailbreaking Large Language Diffusion Models: Revealing Hidden Safety Flaws in Diffusion-Based Text Generation Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-15T18:00:27.404266Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T18:00:26.587605Z digest=sha256:74768d62ecb694678192592a389ef22800b75779a1938c2451e6cf21c2809ce0

Observation 32c91957-d976-4bfd-9b1e-0df7db7243d8 · outbound

This paper cites A Survey on LLM-as-a-Judge.

Jailbreaking Large Language Diffusion Models: Revealing Hidden Safety Flaws in Diffusion-Based Text Generation A Survey on LLM-as-a-Judge

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-15T18:00:26.595001Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:00:26.595001Z digest=sha256:70f66cc64f1a1a40798ee5062d271a035ef3441f05c6382e33567ea95ac2e29e

Observation 8b2da963-8359-40a5-b15a-9d7f40bc1128 · outbound

This paper cites COLD-Attack: Jailbreaking LLMs with Stealthiness and Controllability.

Jailbreaking Large Language Diffusion Models: Revealing Hidden Safety Flaws in Diffusion-Based Text Generation COLD-Attack: Jailbreaking LLMs with Stealthiness and Controllability

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-15T18:00:26.599149Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:00:26.599149Z digest=sha256:0ecfdd0f7469613d7359afe85e91b716f53113ff53d567b0343033030a8a129c

Observation fa39c0d1-d6b7-4671-9e23-2f1a305df492 · outbound

This paper cites DiffusionBERT: Improving Generative Masked Language Models with Diffusion Models.

Jailbreaking Large Language Diffusion Models: Revealing Hidden Safety Flaws in Diffusion-Based Text Generation DiffusionBERT: Improving Generative Masked Language Models with Diffusion Models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-15T18:00:26.603025Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:00:26.603025Z digest=sha256:9f4404b72a6399a49830ce8fde6189f67f3cc3212df6fd7ee7406950e32f0cf7

Observation 302d38fd-29d1-44bc-a0d8-f5fa12c8b764 · outbound

This paper cites Denoising Diffusion Probabilistic Models.

Jailbreaking Large Language Diffusion Models: Revealing Hidden Safety Flaws in Diffusion-Based Text Generation Denoising Diffusion Probabilistic Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-15T18:00:26.606967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:00:26.606967Z digest=sha256:eda6cd959d668ec96f2815ecee2fa5cbaaeae436d1683906de0bb3d8228a3634

Observation 9739a23a-2c81-4c34-91b7-9bad033a98d2 · outbound

This paper cites Llama Guard: LLM-based Input-Output Safeguard for Human-AI Conversations.

Jailbreaking Large Language Diffusion Models: Revealing Hidden Safety Flaws in Diffusion-Based Text Generation Llama Guard: LLM-based Input-Output Safeguard for Human-AI Conversations

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-15T18:00:26.611055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:00:26.611055Z digest=sha256:f7a08c412c9babf60bbbdbc231dec8eaf98a9955c93363a04029325bc3542931

Observation f6af1631-8bf9-4be8-8a56-a07cf5738255 · outbound

This paper cites Baseline Defenses for Adversarial Attacks Against Aligned Language Models.

Jailbreaking Large Language Diffusion Models: Revealing Hidden Safety Flaws in Diffusion-Based Text Generation Baseline Defenses for Adversarial Attacks Against Aligned Language Models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-15T18:00:26.619019Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:00:26.619019Z digest=sha256:a6ae120a0c753e57f046ee0fff66ed86d5d4c945e19d1678d4ca53970ef963cc

Observation 1961d8ca-26d2-449c-aedb-debd7b08be4b · outbound

This paper cites Improved Techniques for Optimization-Based Jailbreaking on Large Language Models.

Jailbreaking Large Language Diffusion Models: Revealing Hidden Safety Flaws in Diffusion-Based Text Generation Improved Techniques for Optimization-Based Jailbreaking on Large Language Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-15T18:00:26.622586Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:00:26.622586Z digest=sha256:f903fef74279f9059ef8431627e091785e040df2b7685f5fc2313aab6cb1c66b

Observation a8d63836-4eeb-4dd1-9287-06270a9b227d · outbound

This paper cites Mistral 7B.

Jailbreaking Large Language Diffusion Models: Revealing Hidden Safety Flaws in Diffusion-Based Text Generation Mistral 7B

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-15T18:00:26.626472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:00:26.626472Z digest=sha256:a38ec30eb535b87a1b050ac736750602398f3b8da5c477c70bff5d88f14e9ab1

Observation 3ffc8441-8614-4078-9641-bf7c7843c96d · outbound

This paper cites Y.; and Poovendran, R.

Jailbreaking Large Language Diffusion Models: Revealing Hidden Safety Flaws in Diffusion-Based Text Generation Y.; and Poovendran, R

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:00:27.389578Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T18:00:26.630045Z digest=sha256:26c34e2fda6b63f396c2c1c7b9fbf9fa8b2c4ab2e5b1c5ed2dbd5d8d4ecb6cbe

Observation ead6cdc3-d165-467a-ae46-1844548f7472 · outbound

This paper cites an unresolved cited work.

Jailbreaking Large Language Diffusion Models: Revealing Hidden Safety Flaws in Diffusion-Based Text Generation Unresolved cited work

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-15T18:00:26.633525Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:00:26.633525Z digest=sha256:9c3bbcb00dffcade94dff434164572cb2005103e99fe0f8207d2e319c7564f8e

Observation 3935cd84-61ab-40a1-8366-d1fdaedccf9c · outbound

This paper cites Diffusion-LM Improves Controllable Text Generation.

Jailbreaking Large Language Diffusion Models: Revealing Hidden Safety Flaws in Diffusion-Based Text Generation Diffusion-LM Improves Controllable Text Generation

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-15T18:00:26.636775Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:00:26.636775Z digest=sha256:20f009ce1b93f26ecec2de48da78cdfb0218bfc4225ba03876c07afb47c3c938

Observation d6e24ccc-08cc-4f8e-89bd-a2154c1b7042 · outbound

This paper cites AutoDAN: Generating Stealthy Jailbreak Prompts on Aligned Large Language Models.

Jailbreaking Large Language Diffusion Models: Revealing Hidden Safety Flaws in Diffusion-Based Text Generation AutoDAN: Generating Stealthy Jailbreak Prompts on Aligned Large Language Models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-15T18:00:26.641738Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:00:26.641738Z digest=sha256:80c9f6393b8c22a999319c719f4deb8947a9a1ff9aa2ffe1de9ea5d502547674

Observation b6d68750-4282-40dc-8f9e-3f2b000626de · outbound

This paper cites dLLM-Cache: Accelerating Diffusion Large Language Models with Adaptive Caching.

Jailbreaking Large Language Diffusion Models: Revealing Hidden Safety Flaws in Diffusion-Based Text Generation dLLM-Cache: Accelerating Diffusion Large Language Models with Adaptive Caching

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-15T18:00:26.645502Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:00:26.645502Z digest=sha256:48a0bf332c62f09859a51eb19a2b6f6193124266ffc02cfec0169240276e050b

Observation 96eb5d27-4e0c-4b7a-9408-714ec46f8c3d · outbound

This paper cites The Llama 3 Herd of Models.

Jailbreaking Large Language Diffusion Models: Revealing Hidden Safety Flaws in Diffusion-Based Text Generation The Llama 3 Herd of Models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-15T18:00:26.649335Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:00:26.649335Z digest=sha256:9e272213dd9d80c1f6468ede8bf622e9906177a3a67b524d3cb8675fab5121cc

Observation 835229d0-1a82-452a-a518-75b588b54f22 · outbound

This paper cites Scaling up Masked Diffusion Models on Text.

Jailbreaking Large Language Diffusion Models: Revealing Hidden Safety Flaws in Diffusion-Based Text Generation Scaling up Masked Diffusion Models on Text

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-15T18:00:26.652662Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:00:26.652662Z digest=sha256:636ab7039951dd9e05448f6c2213ec80ca2c2de53fa12d213e7308a54c610edc

Observation e7a6e223-62cc-4ec0-8c84-7eebe301b12f · outbound

This paper cites Large Language Diffusion Models.

Jailbreaking Large Language Diffusion Models: Revealing Hidden Safety Flaws in Diffusion-Based Text Generation Large Language Diffusion Models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-15T18:00:26.657306Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:00:26.657306Z digest=sha256:a93ce9db3731d023dd0c25b0911448c3c378b6fa388bff86c271419aa9edf45b

Observation b28f704b-06c9-4de0-aebc-e21e85af9b40 · outbound

This paper cites GPT-4 Technical Report.

Jailbreaking Large Language Diffusion Models: Revealing Hidden Safety Flaws in Diffusion-Based Text Generation GPT-4 Technical Report

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-15T18:00:26.662432Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:00:26.662432Z digest=sha256:b8fd1acf465999649009870622695e573a3efaf7848c08a0946fdcf24bacc688

Observation a7d197b8-b09b-4b80-a596-c8a008952edd · outbound

This paper cites Training language models to follow instructions with human feedback.

Jailbreaking Large Language Diffusion Models: Revealing Hidden Safety Flaws in Diffusion-Based Text Generation Training language models to follow instructions with human feedback

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-15T18:00:26.666357Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:00:26.666357Z digest=sha256:51290c67809757eb84a155bc83c1714a14113871b05053abc7f3ec7dd12d50b0

Observation a640efac-a39a-4ae1-aec9-a9c4ba1411b8 · outbound

This paper cites R.; Texier, M.; and Dean, J.

Jailbreaking Large Language Diffusion Models: Revealing Hidden Safety Flaws in Diffusion-Based Text Generation R.; Texier, M.; and Dean, J

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:00:27.375130Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T18:00:26.670200Z digest=sha256:a94b46831fdd90b784c6d46ec589fe6135142fb9f37dcb811b26768bc9560c1c

Observation 1b2d30d2-7849-4bb1-912a-70e5f19d9cd0 · outbound

This paper cites toppling dominos.

Jailbreaking Large Language Diffusion Models: Revealing Hidden Safety Flaws in Diffusion-Based Text Generation toppling dominos

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:00:27.362598Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T18:00:26.674295Z digest=sha256:269e84f1a70530785a529bf72689c44047c438ac95be5179d506d2118ed979b7

Observation bce5bf8f-06df-4177-9f54-3a9e81717880 · outbound

This paper cites Qwen2.5 Technical Report.

Jailbreaking Large Language Diffusion Models: Revealing Hidden Safety Flaws in Diffusion-Based Text Generation Qwen2.5 Technical Report

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-15T18:00:26.678040Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:00:26.678040Z digest=sha256:934a2b30cb8c746b349224ff67a1b3df22bc695d40d473602524854e71005670

Observation 34685ca3-cb33-4c33-9032-e6f90d6bab95 · outbound

This paper cites Direct Preference Optimization: Your Language Model is Secretly a Reward Model.

Jailbreaking Large Language Diffusion Models: Revealing Hidden Safety Flaws in Diffusion-Based Text Generation Direct Preference Optimization: Your Language Model is Secretly a Reward Model

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-15T18:00:26.681610Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:00:26.681610Z digest=sha256:2ca40c31b8ed9b86903ae63389d10a07ceb3949385ae4612e3919fa35db1bafa

Observation 414df16e-cb0c-4a81-814b-a4ed8ad57d75 · outbound

This paper cites Rainbow Teaming: Open-Ended Generation of Diverse Adversarial Prompts.

Jailbreaking Large Language Diffusion Models: Revealing Hidden Safety Flaws in Diffusion-Based Text Generation Rainbow Teaming: Open-Ended Generation of Diverse Adversarial Prompts

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-15T18:00:26.685264Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:00:26.685264Z digest=sha256:7d7c1b27926ef66bf2a08ab67282ea1437a6a56f3c183e44b404e4ea90d2e99a

Observation 8551b4b9-21bb-4d32-97f5-1058bb05675f · outbound

This paper cites SPML: A DSL for Defending Language Models Against Prompt Attacks.

Jailbreaking Large Language Diffusion Models: Revealing Hidden Safety Flaws in Diffusion-Based Text Generation SPML: A DSL for Defending Language Models Against Prompt Attacks

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-15T18:00:26.688905Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:00:26.688905Z digest=sha256:b3ac6ae8c2639250d0cd73dc39c43f4e4e36c6246ceb2b7f0e1a4168c2f4f560

Observation cd7de84a-b390-4670-a876-47ab70565019 · outbound

This paper cites Principle-Driven Self-Alignment of Language Models from Scratch with Minimal Human Supervision.

Jailbreaking Large Language Diffusion Models: Revealing Hidden Safety Flaws in Diffusion-Based Text Generation Principle-Driven Self-Alignment of Language Models from Scratch with Minimal Human Supervision

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-15T18:00:26.693076Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:00:26.693076Z digest=sha256:f096a47ef807949bdd8627365bb0b1bf315beb03203969a495d34c1c31707b8b

Observation b2518bcc-6d4c-43cc-ae87-44c96642b359 · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

Jailbreaking Large Language Diffusion Models: Revealing Hidden Safety Flaws in Diffusion-Based Text Generation Gemini: A Family of Highly Capable Multimodal Models

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-15T18:00:26.697052Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:00:26.697052Z digest=sha256:e6edc3d88476f1e24e1f6d6129487c718b9e7d1344fe71d4703e4e453899b03b

Observation afc4f98a-59ad-4f22-a5ec-8267bb014342 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

Jailbreaking Large Language Diffusion Models: Revealing Hidden Safety Flaws in Diffusion-Based Text Generation Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-15T18:00:26.700708Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:00:26.700708Z digest=sha256:322ec1cb1be7c1f20208a7f461d12c8c07efea1eab6109a4ea7e14a8c59b5d7b

Observation 1edfaab6-6b89-41cf-932c-6cadf0feabdc · outbound

This paper cites A Comprehensive Survey in LLM(-Agent) Full Stack Safety: Data, Training and Deployment.

Jailbreaking Large Language Diffusion Models: Revealing Hidden Safety Flaws in Diffusion-Based Text Generation A Comprehensive Survey in LLM(-Agent) Full Stack Safety: Data, Training and Deployment

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-15T18:00:26.704638Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:00:26.704638Z digest=sha256:32e8fa9721a448d48d23b31996a608fead74f2b2b280d049fe04a9941b2faf46

Observation 8089f707-1dbb-4ab7-b495-bb146c394504 · outbound

This paper cites Fast-dLLM: Training-free Acceleration of Diffusion LLM by Enabling KV Cache and Parallel Decoding.

Jailbreaking Large Language Diffusion Models: Revealing Hidden Safety Flaws in Diffusion-Based Text Generation Fast-dLLM: Training-free Acceleration of Diffusion LLM by Enabling KV Cache and Parallel Decoding

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-15T18:00:26.708042Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:00:26.708042Z digest=sha256:d46fbf68c9af0a95e1d8bdedfea25c03f52fc2e6f29aef0fd78a8571715f70ad

Observation aabc0dd9-aa06-4a62-a73f-2ad7061b4116 · outbound

This paper cites Qwen2 Technical Report.

Jailbreaking Large Language Diffusion Models: Revealing Hidden Safety Flaws in Diffusion-Based Text Generation Qwen2 Technical Report

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-15T18:00:26.711709Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:00:26.711709Z digest=sha256:1f93593a9e12bbd75b5c1518b830d9d0d41ee1e9559aa47a8242cec1de982455

Observation f188c999-93a1-4a0f-8d58-3d4aad1ee6e1 · outbound

This paper cites MMaDA: Multimodal Large Diffusion Language Models.

Jailbreaking Large Language Diffusion Models: Revealing Hidden Safety Flaws in Diffusion-Based Text Generation MMaDA: Multimodal Large Diffusion Language Models

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-15T18:00:26.719717Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:00:26.719717Z digest=sha256:1732b5d1694d3ba0bd01311c7c72d17312da86d7324450835994b6ce84a93ee3

Observation d4afe3a5-a664-492c-ab6f-682388bedf7b · outbound

This paper cites Jailbreak Attacks and Defenses Against Large Language Models: A Survey.

Jailbreaking Large Language Diffusion Models: Revealing Hidden Safety Flaws in Diffusion-Based Text Generation Jailbreak Attacks and Defenses Against Large Language Models: A Survey

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-15T18:00:26.723935Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:00:26.723935Z digest=sha256:9670547ea665a5a9cbc0d9276a8bcf2d6ba2ff1d5f39614da5e3236682f223b7

Observation 397f2d54-f93c-4069-8eb6-4b8b603e1ca0 · outbound

This paper cites LLaDA-V: Large Language Diffusion Models with Visual Instruction Tuning.

Jailbreaking Large Language Diffusion Models: Revealing Hidden Safety Flaws in Diffusion-Based Text Generation LLaDA-V: Large Language Diffusion Models with Visual Instruction Tuning

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-15T18:00:26.728685Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:00:26.728685Z digest=sha256:0d86c81ec16645372558c53e86b784c894ac6d7470d7fa70b551d8b74063d178

Observation 3b26ea39-a295-475d-8cbc-ed3f3f0b5a75 · outbound

This paper cites GPT-4 Is Too Smart To Be Safe: Stealthy Chat with LLMs via Cipher.

Jailbreaking Large Language Diffusion Models: Revealing Hidden Safety Flaws in Diffusion-Based Text Generation GPT-4 Is Too Smart To Be Safe: Stealthy Chat with LLMs via Cipher

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-15T18:00:26.732311Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:00:26.732311Z digest=sha256:59d85d4653cac28ccc738dc9499b5b213a6b460eee39c58c02b7db42a41de030

Observation f1f9b991-61a3-4d92-b8f3-d4dfa3406fc2 · outbound

This paper cites How Johnny Can Persuade LLMs to Jailbreak Them: Rethinking Persuasion to Challenge AI Safety by Humanizing LLMs.

Jailbreaking Large Language Diffusion Models: Revealing Hidden Safety Flaws in Diffusion-Based Text Generation How Johnny Can Persuade LLMs to Jailbreak Them: Rethinking Persuasion to Challenge AI Safety by Humanizing LLMs

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-15T18:00:26.736255Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:00:26.736255Z digest=sha256:c6736215874af8b6dd12709e56ae9123e6e5719f852be159b2b436323ffcaaa0

Observation c75573e3-c618-4f3e-9294-c825a6be6b18 · outbound

This paper cites On Prompt-Driven Safeguarding for Large Language Models.

Jailbreaking Large Language Diffusion Models: Revealing Hidden Safety Flaws in Diffusion-Based Text Generation On Prompt-Driven Safeguarding for Large Language Models

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-15T18:00:26.739580Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:00:26.739580Z digest=sha256:7916db167e311a2068b3dc95d3ae2024e5e227db90e5156ca6bcd131e3f53ceb

Observation 8c8a5cbb-104f-41a0-8ab5-1387e7f486a9 · outbound

This paper cites Speak Out of Turn: Safety Vulnerability of Large Language Models in Multi-turn Dialogue.

Jailbreaking Large Language Diffusion Models: Revealing Hidden Safety Flaws in Diffusion-Based Text Generation Speak Out of Turn: Safety Vulnerability of Large Language Models in Multi-turn Dialogue

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-15T18:00:26.742886Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:00:26.742886Z digest=sha256:d4aa8df8bf7aab86a0f83f1d6cfe5a1e6c472d27266b6eb8c91d77bfd4656eb7

Observation de076d52-f912-4339-b58f-63eaae1a46ca · outbound

This paper cites LLaDA 1.5: Variance-Reduced Preference Optimization for Large Language Diffusion Models.

Jailbreaking Large Language Diffusion Models: Revealing Hidden Safety Flaws in Diffusion-Based Text Generation LLaDA 1.5: Variance-Reduced Preference Optimization for Large Language Diffusion Models

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-15T18:00:26.746725Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:00:26.746725Z digest=sha256:e1ca95a9118c97b53a68d2e57e558d09ad07b7ca7776cf795544315b0bbe7f62

Observation d35434cf-02a5-4e45-9782-9ea4813ba9f5 · outbound

This paper cites Universal and Transferable Adversarial Attacks on Aligned Language Models.

Jailbreaking Large Language Diffusion Models: Revealing Hidden Safety Flaws in Diffusion-Based Text Generation Universal and Transferable Adversarial Attacks on Aligned Language Models

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-15T18:00:26.754277Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:00:26.754277Z digest=sha256:f8389785af780add7dbf7d16d16e0cc08597e79a88631a495dcc21834d1d2af1

Pith citing papers

Observation 89253ad6-01f5-4fbe-89ac-e287371ac473 · inbound

Step-Wise Refusal Dynamics in Autoregressive and Diffusion Language Models cites this paper.

Step-Wise Refusal Dynamics in Autoregressive and Diffusion Language Models Jailbreaking Large Language Diffusion Models: Revealing Hidden Safety Flaws in Diffusion-Based Text Generation

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-03T05:45:43.832756Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T05:45:43.832756Z digest=sha256:9906ae1234e9f4845418ef0e72cf4d89ecc2f8a4ad613d3809d4698da3035da3

Observation 83f4722c-8391-4433-b679-c50cf5ae67ba · inbound

Re-Mask and Redirect: Exploiting Denoising Irreversibility in Diffusion Language Models cites this paper.

Re-Mask and Redirect: Exploiting Denoising Irreversibility in Diffusion Language Models Jailbreaking Large Language Diffusion Models: Revealing Hidden Safety Flaws in Diffusion-Based Text Generation

Reference 8

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T10:49:57.067735Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-15T10:47:59.055929Z digest=sha256:cdbe3fa32558cb47e21c6548bb7ee01b6c1b8a68eeb2ce2fb82a048db289cf2e

Observation 6f45eadf-bc58-411a-b0a9-71c7388ef356 · inbound

GPO-V: Jailbreak Diffusion Vision Language Model by Global Probability Optimization cites this paper.

GPO-V: Jailbreak Diffusion Vision Language Model by Global Probability Optimization Jailbreaking Large Language Diffusion Models: Revealing Hidden Safety Flaws in Diffusion-Based Text Generation

Reference 12

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T04:20:56.597587Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-11T01:48:09.652599Z digest=sha256:ff108ea131cba5cc00dd9055ffc3e01916b4c77f3d1ba09af751d1311f33f887

Observation fca03af9-3367-4685-b9b4-47bdbd0fe87f · inbound

GPO-V: Jailbreak Diffusion Vision Language Model by Global Probability Optimization cites this paper.

GPO-V: Jailbreak Diffusion Vision Language Model by Global Probability Optimization Jailbreaking Large Language Diffusion Models: Revealing Hidden Safety Flaws in Diffusion-Based Text Generation

Reference 12

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T06:31:25.745202Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-12T04:13:54.794576Z digest=sha256:2271b521820b952c5223d2f23264f644bb8f8a9faf1bff7d0b634e921f4e4082

Observation 37d59a2c-55ee-41c1-8f1b-7149b081c3a6 · inbound

BadDLM: Backdooring Diffusion Language Models with Diverse Targets cites this paper.

BadDLM: Backdooring Diffusion Language Models with Diverse Targets Jailbreaking Large Language Diffusion Models: Revealing Hidden Safety Flaws in Diffusion-Based Text Generation

Reference 64

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T06:11:25.853932Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-12T04:30:13.417357Z digest=sha256:cb0da1c39530ee95db0ae9120d07dd7bdf98c304e42b875b5c1ae341f235e674

Observation 95f1c04c-262a-46e4-9fb4-032d8ffba1aa · inbound

Machine Unlearning for Masked Diffusion Language Models cites this paper.

Machine Unlearning for Masked Diffusion Language Models Jailbreaking Large Language Diffusion Models: Revealing Hidden Safety Flaws in Diffusion-Based Text Generation

Reference 8

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T10:28:12.201872Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-20T10:25:29.989996Z digest=sha256:6b3dc9b639a651dffea7f622a9df10d3948d464001c62a7936415999d5987cc1

Observation c7f79fba-3efa-42ee-9f44-98797f550013 · inbound

SafeDiffusion-R1: Online Reward Steering for Safe Diffusion Post-Training cites this paper.

SafeDiffusion-R1: Online Reward Steering for Safe Diffusion Post-Training Jailbreaking Large Language Diffusion Models: Revealing Hidden Safety Flaws in Diffusion-Based Text Generation

Reference 87

Resolution
verified exact
arxiv_id, observed 2026-05-20T11:28:14.056238Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-05-20T11:26:55.810822Z digest=sha256:bb453db92a7828a5fd4c35b71a02386e1bc05f707c35e4f6ffb35fb9e2ff92e3

Observation d7136bfa-3c65-4751-acd5-3062ba8432db · inbound

Backdooring Masked Diffusion Language Models cites this paper.

Backdooring Masked Diffusion Language Models Jailbreaking Large Language Diffusion Models: Revealing Hidden Safety Flaws in Diffusion-Based Text Generation

Reference 14

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T07:33:07.479640Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-20T07:30:19.657903Z digest=sha256:79a5b5fa217e38ddca56f4392a40b5f6ddaeeaaf1c6b04ffcf2f9919ccd62906

Observation 10eb3ec6-b0a6-4f14-a2ea-5a5078f45087 · inbound

Backdooring Masked Diffusion Language Models cites this paper.

Backdooring Masked Diffusion Language Models Jailbreaking Large Language Diffusion Models: Revealing Hidden Safety Flaws in Diffusion-Based Text Generation

Reference 14

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T18:04:58.199130Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-06-30T18:00:05.311834Z digest=sha256:95db8f947411f5cc72ce4c5e6a08ecb506004ec960497205a6d1d80ebb5bba0b

Observation c164645a-7bf7-48ed-8a08-e6bf972c0313 · inbound

MaskForge: Structure-Aware Adaptive Attacks for Jailbreaking Diffusion Large Language Models cites this paper.

MaskForge: Structure-Aware Adaptive Attacks for Jailbreaking Diffusion Large Language Models Jailbreaking Large Language Diffusion Models: Revealing Hidden Safety Flaws in Diffusion-Based Text Generation

Reference 51

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T23:56:24.424580Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-06-28T13:43:51.171443Z digest=sha256:cdec4b6567326772b2fe12e8bd5d7f2e411c97e3ee7cb9d758218eb9123a145e

Observation 41be467e-917b-44d6-8327-95b4961d968a · inbound

Adversarial Diffusion Across Modalities: A Fusion Survey of Attacks, Defenses, and Evaluation for Text, Vision, and Vision-Language Models cites this paper.

Adversarial Diffusion Across Modalities: A Fusion Survey of Attacks, Defenses, and Evaluation for Text, Vision, and Vision-Language Models Jailbreaking Large Language Diffusion Models: Revealing Hidden Safety Flaws in Diffusion-Based Text Generation

Reference 80

Resolution
metadata mismatch
arxiv_id, observed 2026-06-26T04:38:59.178774Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-06-26T04:35:51.583460Z digest=sha256:c8fa7127c814da025163e9906559c6711ce9376ff9963cd3fd6cb7e06f9685b0

Observation 0240191a-feb7-4b2b-b5db-5cc3adf5d0d6 · inbound

Beyond the Bidirectional Promise: Re-evaluating the Robustness of Diffusion Language Models cites this paper.

Beyond the Bidirectional Promise: Re-evaluating the Robustness of Diffusion Language Models Jailbreaking Large Language Diffusion Models: Revealing Hidden Safety Flaws in Diffusion-Based Text Generation

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-01T08:22:29.809008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T08:22:29.809008Z digest=sha256:e8579a8d040e612743338bfdb26c966042359006ffeeeab9809cb15eaef55ca7