Pith. sign in

Paper Citation Record · LEDGER

Red-Teaming the Stable Diffusion Safety Filter

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 47 inbound Pith citation observations for arXiv:2210.04610.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2210.04610 v5

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 47 of 47 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 47 of 47 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-10T17:05:50.779016Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T19:30:07.408628Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 30efb5b2-38d4-4a71-be3e-c2532196c174 · inbound

SalUn: Empowering Machine Unlearning via Gradient-based Weight Saliency in Both Image Classification and Generation cites this paper.

SalUn: Empowering Machine Unlearning via Gradient-based Weight Saliency in Both Image Classification and Generation Red-Teaming the Stable Diffusion Safety Filter

Reference 137

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T17:56:23.405606Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-16T17:56:23.281678Z digest=sha256:102501ce5d5f3c86ce7ebd88b5b7677d35ff2cc08cc164afa708c961a48f1b9d

Observation dc63e949-132a-46ba-8a47-ae149d8a8cf1 · inbound

VLBiasBench: A Comprehensive Benchmark for Evaluating Bias in Large Vision-Language Model cites this paper.

VLBiasBench: A Comprehensive Benchmark for Evaluating Bias in Large Vision-Language Model Red-Teaming the Stable Diffusion Safety Filter

Reference 55

Resolution
verified exact
arxiv_id, observed 2026-05-23T23:38:37.379221Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-23T23:37:12.985780Z digest=sha256:999460056a4cfee3d1d28e7ea0a5700d48e0c3189fe1b4747949d9d4abe7f14a

Observation e2cd63f5-d91b-4cc1-ad38-050c4a62631d · inbound

T2ISafety: Benchmark for Assessing Fairness, Toxicity, and Privacy in Image Generation cites this paper.

T2ISafety: Benchmark for Assessing Fairness, Toxicity, and Privacy in Image Generation Red-Teaming the Stable Diffusion Safety Filter

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-10T17:05:50.779016Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:05:50.779016Z digest=sha256:cee1f24726ca97a7785a2872007d71fb15a923990d18c7b8929bf880e0a0882c

Observation a0c956c3-bc32-4104-a409-0336772fc11a · inbound

SAeUron: Interpretable Concept Unlearning in Diffusion Models with Sparse Autoencoders cites this paper.

SAeUron: Interpretable Concept Unlearning in Diffusion Models with Sparse Autoencoders Red-Teaming the Stable Diffusion Safety Filter

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-10T01:04:57.216274Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T01:04:57.216274Z digest=sha256:3b9a2eff8096c2acfb2155321d2bf31ec560c656aaf8dc13f3283e4c7fce5a09

Observation 62faf603-bed8-4221-871d-4a1d50cc260c · inbound

The Pitfalls of "Security by Obscurity" And What They Mean for Transparent AI cites this paper.

The Pitfalls of "Security by Obscurity" And What They Mean for Transparent AI Red-Teaming the Stable Diffusion Safety Filter

Reference 158

Resolution
unresolved
no resolver link, observed 2026-08-09T23:22:44.961099Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T23:22:44.961099Z digest=sha256:fcc9d90cfb97364630a00a7e9a57e6999f4193413f261e0ec620187d664a8e36

Observation dadb8376-2fe8-4031-baa7-492df68b2ddc · inbound

Predictive Red Teaming: Breaking Policies Without Breaking Robots cites this paper.

Predictive Red Teaming: Breaking Policies Without Breaking Robots Red-Teaming the Stable Diffusion Safety Filter

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-08T15:04:10.824381Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T15:04:10.824381Z digest=sha256:f9d795ab1667df4269339cd44581b2c3fe5037adca87e7b87a0b87b19defef4d

Observation 069feabf-56b7-48ad-bcbb-71e4bc0b117f · inbound

Red-Teaming Text-to-Image Systems by Rule-based Preference Modeling cites this paper.

Red-Teaming Text-to-Image Systems by Rule-based Preference Modeling Red-Teaming the Stable Diffusion Safety Filter

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T13:46:21.103762Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:46:21.103762Z digest=sha256:20240c1efaeff114c34d689dcf924d7714fc86232576e213c5b833a7c8a715e2

Observation 4d43f506-93da-43b2-ac54-1d6a5f4cba90 · inbound

CuRe: Cultural Gaps in the Long Tail of Text-to-Image Systems cites this paper.

CuRe: Cultural Gaps in the Long Tail of Text-to-Image Systems Red-Teaming the Stable Diffusion Safety Filter

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T05:27:16.156591Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:27:16.156591Z digest=sha256:e6068393b5fa768526e6eea3cee01a6bf2ef44ae802069c94daf0aaa49d6aa05

Observation 3dba781f-d7fc-48e6-8f27-fad422b04043 · inbound

SAGE: Exploring the Boundaries of Unsafe Concept Domain with Semantic-Augment Erasing cites this paper.

SAGE: Exploring the Boundaries of Unsafe Concept Domain with Semantic-Augment Erasing Red-Teaming the Stable Diffusion Safety Filter

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T04:55:44.680812Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:55:44.680812Z digest=sha256:2cffa0c8945778c4026565ea71e4fd8eba9c45ce458403f73aaeaa235cdd26c6

Observation 4bda3e92-afa6-4747-8e32-8713eaafad36 · inbound

GenBreak: Red Teaming Text-to-Image Generators Using Large Language Models cites this paper.

GenBreak: Red Teaming Text-to-Image Generators Using Large Language Models Red-Teaming the Stable Diffusion Safety Filter

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T04:52:19.067456Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:52:19.067456Z digest=sha256:87f7e433e3965e86a84c09dcaef090cd906e5eea7f2f8c50b18f6d47ede3979c

Observation 69baffef-f6b4-46e3-9668-eefda73c5a92 · inbound

"I Cannot Write This Because It Violates Our Content Policy": Understanding Content Moderation Policies and User Experiences in Generative AI Products cites this paper.

"I Cannot Write This Because It Violates Our Content Policy": Understanding Content Moderation Policies and User Experiences in Generative AI Products Red-Teaming the Stable Diffusion Safety Filter

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-07T00:27:19.703033Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:27:19.703033Z digest=sha256:1dcf97e5b1826d183fcb2e88dd560a9a450d482ecd854f9782e3606c1f98c273

Observation 4eeff8dc-5b69-434e-906a-28513f8aff3a · inbound

Automating Evaluation of Diffusion Model Unlearning with (Vision-) Language Model World Knowledge cites this paper.

Automating Evaluation of Diffusion Model Unlearning with (Vision-) Language Model World Knowledge Red-Teaming the Stable Diffusion Safety Filter

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T19:06:39.320386Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:06:39.320386Z digest=sha256:ec246da8cd9fa57a56b97c9377f24d466df95d9fd0df91e02ee0f7d955f7d0ac

Observation 097224c5-febd-4b76-8172-0094a64d1d89 · inbound

GIFT: Gradient-aware Immunization of diffusion models against malicious Fine-Tuning with safe concepts retention cites this paper.

GIFT: Gradient-aware Immunization of diffusion models against malicious Fine-Tuning with safe concepts retention Red-Teaming the Stable Diffusion Safety Filter

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T16:26:21.324774Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:26:21.324774Z digest=sha256:2e217f39cbba35e42e2ecffd521a59373cc8c07706cafab41cb986cc292a94cd

Observation 2fbce953-4f82-4367-8d18-e88c5bd672cc · inbound

LoReUn: Data Itself Implicitly Provides Cues to Improve Machine Unlearning cites this paper.

LoReUn: Data Itself Implicitly Provides Cues to Improve Machine Unlearning Red-Teaming the Stable Diffusion Safety Filter

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T11:42:04.433290Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:42:04.433290Z digest=sha256:50a33ec34e6d5ef3bc53c36120bc49c58fdc81338f6b5a42c504a5e716708970

Observation d30b25dc-9f4c-4de3-884c-2cbb25a12374 · inbound

PLA: Prompt Learning Attack against Text-to-Image Generative Models cites this paper.

PLA: Prompt Learning Attack against Text-to-Image Generative Models Red-Teaming the Stable Diffusion Safety Filter

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T17:42:50.883150Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:42:50.883150Z digest=sha256:952fae71026b761cd1b8beb8365a66ea8f45b8953424efa0632204ae6258e1ab

Observation 90b66d53-051e-443c-8811-51c3df87e4ca · inbound

Zero-Residual Concept Erasure via Progressive Alignment in Text-to-Image Model cites this paper.

Zero-Residual Concept Erasure via Progressive Alignment in Text-to-Image Model Red-Teaming the Stable Diffusion Safety Filter

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T00:04:32.495158Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:04:32.495158Z digest=sha256:c55ee21f61963448c09db70435232bad445449436564f8ddbb0d000a41438794

Observation f987b762-6819-4523-97cb-fd03d07f2c17 · inbound

Memory Enhanced Fractional-Order Dung Beetle Optimization for Photovoltaic Parameter Identification cites this paper.

Memory Enhanced Fractional-Order Dung Beetle Optimization for Photovoltaic Parameter Identification Red-Teaming the Stable Diffusion Safety Filter

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-05T22:32:12.208299Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T22:32:12.208299Z digest=sha256:566527a3efc5e64e6f08af7d34df90007b62ee9e61061275129668a66321896c

Observation 9a309734-ec7d-45e8-956f-e17ebd7a4b4c · inbound

Video Deepfake Abuse: How Company Choices Predictably Shape Misuse Patterns cites this paper.

Video Deepfake Abuse: How Company Choices Predictably Shape Misuse Patterns Red-Teaming the Stable Diffusion Safety Filter

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-03T19:59:09.542591Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T19:59:09.542591Z digest=sha256:2c0e108e56005a22b5d5e2d113395644099df38dfd2a80557b3e76912d455b13

Observation dceef4ca-9a89-4e9f-a4f9-a8325bacb57a · inbound

UnHype: CLIP-Guided Hypernetworks for Dynamic LoRA Unlearning cites this paper.

UnHype: CLIP-Guided Hypernetworks for Dynamic LoRA Unlearning Red-Teaming the Stable Diffusion Safety Filter

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-03T05:05:53.583424Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T05:05:53.583424Z digest=sha256:735095d3d32e2c77c81b8703e1350bee00dd52f4feea6acbfdefeb05d29d8e82

Observation 92d5cfe2-ccd0-4d97-a30c-51ef6141b5c3 · inbound

Erasure or Erosion? Evaluating Compositional Degradation in Unlearned Text-To-Image Diffusion Models cites this paper.

Erasure or Erosion? Evaluating Compositional Degradation in Unlearned Text-To-Image Diffusion Models Red-Teaming the Stable Diffusion Safety Filter

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:20:51.485800Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T19:13:19.165859Z digest=sha256:548c4cbd7ffc7b0cb9602f9da6d9bdc3e29e7d406efe8f0c6ee0a907378f42f8

Observation e1bdc07c-9e5e-4a0a-8116-98935ba9876c · inbound

Harnessing Hyperbolic Geometry for Harmful Prompt Detection and Sanitization cites this paper.

Harnessing Hyperbolic Geometry for Harmful Prompt Detection and Sanitization Red-Teaming the Stable Diffusion Safety Filter

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-10T22:25:49.139747Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T19:55:16.003348Z digest=sha256:1bf1a7f22bcf443af15193fd1d37b36f556ecfeb091b34e15f2506b876190c6f

Observation e5cfc204-e8dd-4a88-b43c-d39fea4360d7 · inbound

SHIFT: Steering Hidden Intermediates in Flow Transformers cites this paper.

SHIFT: Steering Hidden Intermediates in Flow Transformers Red-Teaming the Stable Diffusion Safety Filter

Reference 26

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T00:20:55.653350Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T18:33:59.134475Z digest=sha256:b923383865d442568a276759f2430f990722d4e52e992a716fae7d2aec73ac4b

Observation b5b97f17-2c94-4283-9dba-15eed072a9e5 · inbound

Mosaic: Multimodal Jailbreak against Closed-Source VLMs via Multi-View Ensemble Optimization cites this paper.

Mosaic: Multimodal Jailbreak against Closed-Source VLMs via Multi-View Ensemble Optimization Red-Teaming the Stable Diffusion Safety Filter

Reference 30

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T05:35:59.083195Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T18:03:36.805784Z digest=sha256:6bca3526f2bc95b67829b472500823564fc7a64cde3e7ae02ac5f61f8d778014

Observation 189cbbb5-cfcd-4f2c-91f3-49d1b99dfcec · inbound

Closed-Form Concept Erasure via Double Projections cites this paper.

Closed-Form Concept Erasure via Double Projections Red-Teaming the Stable Diffusion Safety Filter

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-05-11T08:16:00.850753Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T16:44:57.265729Z digest=sha256:c07556d9c0c33c72f0c2d403c7e269752911b2b9da0b5c7949f97995c0ad1e8e

Observation ba0a17e2-f99f-49b2-8b94-91b70afdf8fc · inbound

Evaluation without Generation: Non-Generative Assessment of Harmful Model Specialization with Applications to CSAM cites this paper.

Evaluation without Generation: Non-Generative Assessment of Harmful Model Specialization with Applications to CSAM Red-Teaming the Stable Diffusion Safety Filter

Reference 43

Resolution
verified exact
arxiv_id, observed 2026-05-11T23:31:14.149119Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-07T16:55:19.775120Z digest=sha256:563336d2d4a1bae7d712835a4cfe3f6d2ade24a30da60767eb5019e662147a5e

Observation 62d51980-131a-4b20-a2a3-6194ef5b767d · inbound

Disciplined Diffusion: Text-to-Image Diffusion Model against NSFW Generation cites this paper.

Disciplined Diffusion: Text-to-Image Diffusion Model against NSFW Generation Red-Teaming the Stable Diffusion Safety Filter

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-11T15:56:27.429603Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-09T19:03:00.734760Z digest=sha256:955f0f58778261435dd827a7da4eb4abaec8269a71b995eb5e58aef8296bd8dd

Observation edc70e0c-ed98-45da-84d0-95b651723d13 · inbound

SafeDiffusion-R1: Online Reward Steering for Safe Diffusion Post-Training cites this paper.

SafeDiffusion-R1: Online Reward Steering for Safe Diffusion Post-Training Red-Teaming the Stable Diffusion Safety Filter

Reference 20

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T11:28:14.348923Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T11:26:55.810822Z digest=sha256:3547f4006d18df854d377c6f32d54d0bfaee415c70d456d48f9bd9f1f8107a06

Observation 509a8273-35c9-4845-9910-af04415903b5 · inbound

FlowErase-RL: Rethinking Concept Erasure as Reward Optimization in Flow Matching Models cites this paper.

FlowErase-RL: Rethinking Concept Erasure as Reward Optimization in Flow Matching Models Red-Teaming the Stable Diffusion Safety Filter

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-20T05:28:04.249811Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-20T05:27:47.218349Z digest=sha256:8cc37f26f4e392e585915305cfca4e5ac584675ac09f65b3f503a64c8872a798

Observation c1b7f489-b809-418a-8ea3-e69f29e76800 · inbound

FlowErase-RL: Rethinking Concept Erasure as Reward Optimization in Flow Matching Models cites this paper.

FlowErase-RL: Rethinking Concept Erasure as Reward Optimization in Flow Matching Models Red-Teaming the Stable Diffusion Safety Filter

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-07-01T15:05:47.554863Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T18:18:15.772860Z digest=sha256:6240956f8a1c9789351221ebd6c56555f85f9cb5c7289b2d748356593fe7d8ec

Observation 09ed3b0f-e5aa-46df-a790-2325131c2d3f · inbound

Erased but Exploitable: Black-box Embedding-Aware Prompting Against Unlearned Text-to-Image Diffusion Models cites this paper.

Erased but Exploitable: Black-box Embedding-Aware Prompting Against Unlearned Text-to-Image Diffusion Models Red-Teaming the Stable Diffusion Safety Filter

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-06-29T22:34:02.827904Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-29T22:14:36.290982Z digest=sha256:b761ee15b46392db6cff5a03b35fe595d63e1e0cfafd76657fcf8f1813a45d20

Observation a6519072-f836-4a2b-960e-a315f47b80ce · inbound

Orthogonal Negative Guidance in Attention Feature Space for Text-to-Image Generation cites this paper.

Orthogonal Negative Guidance in Attention Feature Space for Text-to-Image Generation Red-Teaming the Stable Diffusion Safety Filter

Reference 44

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T08:33:15.048106Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-29T08:30:02.532832Z digest=sha256:ee24c37ebbd530f49d4a8aaacaed14ba839d2641de0638999ead3e774655047f

Observation 63f3040b-738c-4105-bb12-15918d9fe22a · inbound

Geometric Erasure by Contrastive Velocity Matching in Rectified Flows cites this paper.

Geometric Erasure by Contrastive Velocity Matching in Rectified Flows Red-Teaming the Stable Diffusion Safety Filter

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-06-28T23:52:49.372925Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-28T23:50:44.624338Z digest=sha256:4d068871b571e4bd67f07194593bc16f780e3a0771f01d83679141dd93e73f84

Observation cc55cab7-ae3a-47b2-a32c-79e6451a5842 · inbound

Benign Inputs, Harmful Outputs: Cross-Modal Jailbreaking via Distributed Semantic Recomposition cites this paper.

Benign Inputs, Harmful Outputs: Cross-Modal Jailbreaking via Distributed Semantic Recomposition Red-Teaming the Stable Diffusion Safety Filter

Reference 19

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T23:26:23.287355Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-28T14:17:52.902183Z digest=sha256:013c0106eeff9fdcce4bb5754e99161e2b3496f670319e447b758429d647a311

Observation eec0650a-35c6-4258-bffd-8661e9867565 · inbound

Initialization is Half the Battle: Generating Diverse Images from a Guidance Potential Posterior cites this paper.

Initialization is Half the Battle: Generating Diverse Images from a Guidance Potential Posterior Red-Teaming the Stable Diffusion Safety Filter

Reference 30

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T22:36:16.679465Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-28T15:19:32.238094Z digest=sha256:ed18b0ace144f913b456d43ca8e2a1400abb7b5451b290c2876c3fad2107455f

Observation 771d36e9-0670-42fd-936b-1f580833b4b8 · inbound

RedEdit: Agentic Red-Teaming of Image Safety Classifiers via MCTS-Guided Photo-Editing cites this paper.

RedEdit: Agentic Red-Teaming of Image Safety Classifiers via MCTS-Guided Photo-Editing Red-Teaming the Stable Diffusion Safety Filter

Reference 38

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T14:07:02.811855Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-28T00:42:33.088388Z digest=sha256:859dae147075b5618373bff6c75bef372bf7f55a32cabfd04f0a883d3acca415

Observation 3831dd77-d44e-4b0f-8849-4381e9cb55a6 · inbound

Where Should Knowledge Enter? A Layered Framework for Knowledge Infusion in Multimodal Iterative Generative Model cites this paper.

Where Should Knowledge Enter? A Layered Framework for Knowledge Infusion in Multimodal Iterative Generative Model Red-Teaming the Stable Diffusion Safety Filter

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-07-02T13:16:58.543137Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-28T01:28:55.698724Z digest=sha256:d6eeb8fce5d5f53e7b961d45127349b5a0793486a1ed3c3dcf7fac36f49d0282

Observation ce4804bc-5012-46ae-86b3-58ab0b04dd2b · inbound

Unified Safe In-context Image Generation in Multimodal Diffusion Transformers via Restricting Unsafe Information Flows cites this paper.

Unified Safe In-context Image Generation in Multimodal Diffusion Transformers via Restricting Unsafe Information Flows Red-Teaming the Stable Diffusion Safety Filter

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-07-02T16:07:09.105956Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-27T22:58:44.478077Z digest=sha256:c86b42d7bb7352faa590847dabad2ae600ccb861978f4ccb89c85c4f2f2a4a42

Observation 508e50a8-f7bf-454a-a0cd-32d27c4895c5 · inbound

Pulling The REINS: Training-Free Safety Alignment of Video Diffusion Models via Representation Steering cites this paper.

Pulling The REINS: Training-Free Safety Alignment of Video Diffusion Models via Representation Steering Red-Teaming the Stable Diffusion Safety Filter

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-07-03T17:48:45.925768Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-27T03:44:36.071488Z digest=sha256:7218dbfa745618adb8f6f427ce233c17363840ce0718f53dd4b3e661042ab71e

Observation 9edef38f-2782-43e0-97bf-b8ffdb098c6a · inbound

Co-occurring associated retained concepts in Diffusion Unlearning cites this paper.

Co-occurring associated retained concepts in Diffusion Unlearning Red-Teaming the Stable Diffusion Safety Filter

Reference 12

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T16:09:57.486746Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-26T00:50:58.689718Z digest=sha256:318ac7f89bb8dbf8a621f110bd450382da5817326e8ee0c1634dc093f443b46e

Observation 11244526-f4b7-414c-80b4-95699b93d437 · inbound

Concept Removal for Frontier Image Generative Models cites this paper.

Concept Removal for Frontier Image Generative Models Red-Teaming the Stable Diffusion Safety Filter

Reference 100

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T19:30:07.410007Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-25T21:18:32.620951Z digest=sha256:76240f484ce4cfb47a997931364beddcdac76a66418ec0015af59e8421fed867

Observation 9d19c59f-add2-4392-a18a-58a68fe889b0 · inbound

Safe Autoregressive Image Generation with Iterative Self-Improving Codebooks cites this paper.

Safe Autoregressive Image Generation with Iterative Self-Improving Codebooks Red-Teaming the Stable Diffusion Safety Filter

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-07-04T12:49:52.907553Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-26T05:51:01.446139Z digest=sha256:b9a6d30ab294a532fee1a1f4f92f1b42e3e7b24b058d72e17ba95b78995cd1e3

Observation b4211f0d-2159-4da2-837f-294e9363b48e · inbound

Evaluating Intellectual Property Guardrails of Generative Image Models: A Technical Report cites this paper.

Evaluating Intellectual Property Guardrails of Generative Image Models: A Technical Report Red-Teaming the Stable Diffusion Safety Filter

Reference 18

Resolution
unresolved
no resolver link, observed 2026-07-12T09:37:15.997513Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:37:15.997513Z digest=sha256:c287cb5afbd7f6757c80988e7e0b187676f0c966289104fc63efe96a02cfdb44

Observation 5bc1920f-8788-4ea1-adde-2a285b409554 · inbound

Introspective Attention Modulation for Safe Text-to-Image Generation cites this paper.

Introspective Attention Modulation for Safe Text-to-Image Generation Red-Teaming the Stable Diffusion Safety Filter

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:54.327706Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:54.327706Z digest=sha256:af75efa28df4e813e9eb828f173fd15ac0d2cf78c1a7cbb3ec34c38abab846c3

Observation 87ceca04-b934-45cf-a12e-bcc1bdbcd289 · inbound

Dynamic Defense Profiling Enables Cognitive Jailbreak of Text-to-Image Models cites this paper.

Dynamic Defense Profiling Enables Cognitive Jailbreak of Text-to-Image Models Red-Teaming the Stable Diffusion Safety Filter

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-01T17:07:54.486767Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T17:07:54.486767Z digest=sha256:f38d393d80552d36decc7a5c30276af400d034b312d8184323b1f0eb9462fb9e

Observation dab85d5f-c012-4c50-b048-84e37d048f83 · inbound

DECAF: De-Clustering for Adaptive Representational Unlearning cites this paper.

DECAF: De-Clustering for Adaptive Representational Unlearning Red-Teaming the Stable Diffusion Safety Filter

Reference 93

Resolution
unresolved
no resolver link, observed 2026-07-31T23:33:57.639498Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-31T23:33:57.639498Z digest=sha256:765ff023852addd13765535e56f357874ef96f77505b3a8150e4b9781d29524b

Observation d1e860ea-16ee-43f7-bec3-4b0c4c30aa4a · inbound

TYPO: Instruction-Dense Visual Jailbreaks against Commercial Closed-Source Image-Generation Models cites this paper.

TYPO: Instruction-Dense Visual Jailbreaks against Commercial Closed-Source Image-Generation Models Red-Teaming the Stable Diffusion Safety Filter

Reference 62

Resolution
unresolved
no resolver link, observed 2026-07-31T09:23:52.492793Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-31T09:23:52.492793Z digest=sha256:e397f60e8d963109c7612a307dd709c25460863ea809cd3712d25d6a0c8f0991

Observation 01080458-b9aa-43e6-a40b-854760dd5075 · inbound

A Multimodal Automatic Redteaming Evaluation based on Atomic Jailbreak Strategy Decoupling and Combination cites this paper.

A Multimodal Automatic Redteaming Evaluation based on Atomic Jailbreak Strategy Decoupling and Combination Red-Teaming the Stable Diffusion Safety Filter

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-07T00:14:43.164166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:14:43.164166Z digest=sha256:b6f1c460bea7956f107b0b32b903a347c62667fca37287e5f568c674bf5299ed