REVIEW 14 cited by
Mist: Towards Improved Adversarial Examples for Diffusion Models
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Diffusion Models (DMs) have empowered great success in artificial-intelligence-generated content, especially in artwork creation, yet raising new concerns in intellectual properties and copyright. For example, infringers can make profits by imitating non-authorized human-created paintings with DMs. Recent researches suggest that various adversarial examples for diffusion models can be effective tools against these copyright infringements. However, current adversarial examples show weakness in transferability over different painting-imitating methods and robustness under straightforward adversarial defense, for example, noise purification. We surprisingly find that the transferability of adversarial examples can be significantly enhanced by exploiting a fused and modified adversarial loss term under consistent parameters. In this work, we comprehensively evaluate the cross-method transferability of adversarial examples. The experimental observation shows that our method generates more transferable adversarial examples with even stronger robustness against the simple adversarial defense.
Forward citations
Cited by 14 Pith papers
-
Cross-Branch Conflict as a Shield: Safeguarding Facial Identities in Unified Multimodal Image Editing
CCS jointly perturbs the ViT and VAE pathways and reduces their CKA agreement, causing unified multimodal image editors to lose facial identity.
-
Beyond Invisibility: Learning Robust Visible Watermarks for Stronger Copyright Protection
HARVIM learns watermark placement to maximize reconstruction error under an inpainting-based removal model, showing modest gains over random watermarks.
-
SRAP: SVD-Refined Adversarial Perturbations for Imperceptible Face-Swap Defense
SRAP combines per-channel truncated SVD and an identity-importance mask to make PGD perturbations for face-swap defense more imperceptible while retaining competitive identity disruption.
-
Delving into the Temporal Challenges of Unified Video Protection Against Image-to-Video and Fine-Tuning-based Customization
TC-UAP learns a shared multi-frame adversarial perturbation that protects videos of the same identity from both fine-tuning-based and reference-based video customization, remaining effective on unseen clips and under ...
-
Defending from GeoLocalization through Adversarial Road Trips
RoadTrip Attack uses beam search over adaptive geographic intermediate targets to produce stronger, more transferable, lower-visibility adversarial examples against retrieval-based image geolocalizers than PGD, FGSM, ...
-
SyncBreaker:Stage-Aware Multimodal Adversarial Attacks on Audio-Driven Talking Head Generation
SyncBreaker jointly attacks image and audio streams with Multi-Interval Sampling and Cross-Attention Fooling to degrade speech-driven talking head generation more than single-modality baselines.
-
LoReUn: Data Itself Implicitly Provides Cues to Improve Machine Unlearning
LoReUn, a plug-in loss-based reweighting strategy, improves approximate machine unlearning by focusing updates on hard-to-forget low-loss data points.
-
Silence is Golden: Leveraging Adversarial Examples to Nullify Audio Control in LDM-based Talking-Head Generation
Silencer adds a nearly invisible disturbance to portraits that makes LDM-based talking-head models keep the mouth silent, and it survives several image-purification countermeasures.
-
Immunizing Images from Text to Image Editing via Adversarial Cross-Attention
An imperceptible adversarial noise, computed with a LLaVA caption as a stand-in for the unknown edit prompt, disrupts cross-attention in Stable Diffusion-based editors and makes text-guided edits fail.
-
Evaluating Adversarial Protections for Diffusion Personalization: A Comprehensive Study
A unified benchmark of eight perturbation-based protections shows budget-dependent trade-offs between stealth and disruption, with no method winning across all metrics.
-
What is Adversarial Training for Diffusion Models?
Diffusion models trained with an equivariant adversarial-smoothing regularizer tolerate heavy training-data corruption but lose image quality on clean data.
-
Structure Disruption: Subverting Malicious Diffusion-Based Inpainting via Self-Attention Query Perturbation
SDA adds a small perturbation that disrupts self-attention queries in the first denoising step, stopping Stable Diffusion inpainting from producing coherent edits.
-
Dac-Fake: A Divide and Conquer Framework for Detecting Fake News on Social Media
The abstract claims a new fake news detector with 97.88%, 96.05%, and 97.32% accuracy, but the manuscript body is a different paper on image inpainting protection.
-
Is Perturbation-Based Image Protection Disruptive to Image Editing?
Perturbation-based protections do not reliably block diffusion editing, and in many cases they increase the edited image's alignment with the guidance prompt.
Discussion (0). Continue with ORCID to comment.