REVIEW 4 cited by
Custom-Edit: Text-Guided Image Editing with Customized Diffusion Models
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Text-to-image diffusion models can generate diverse, high-fidelity images based on user-provided text prompts. Recent research has extended these models to support text-guided image editing. While text guidance is an intuitive editing interface for users, it often fails to ensure the precise concept conveyed by users. To address this issue, we propose Custom-Edit, in which we (i) customize a diffusion model with a few reference images and then (ii) perform text-guided editing. Our key discovery is that customizing only language-relevant parameters with augmented prompts improves reference similarity significantly while maintaining source similarity. Moreover, we provide our recipe for each customization and editing process. We compare popular customization methods and validate our findings on two editing methods using various datasets.
Forward citations
Cited by 4 Pith papers
-
Beyond Invisibility: Learning Robust Visible Watermarks for Stronger Copyright Protection
HARVIM learns watermark placement to maximize reconstruction error under an inpainting-based removal model, showing modest gains over random watermarks.
-
AtomDiffuser: Time-Aware Degradation Modeling for Drift and Beam Damage in STEM Imaging
AtomDiffuser predicts affine drift and spatially varying beam damage between STEM image frames, trained on synthetic degradation and shown qualitatively on real cryo-STEM data.
-
Per-Query Visual Concept Learning
A prompt- and seed-specific, attention-based loss step improves both identity preservation and prompt adherence for six personalization methods across SD, SDXL, and FLUX backbones.
-
In-Context Brush: Zero-shot Customized Subject Insertion with Context-Aware Latent Space Manipulation
In-Context Brush performs zero-shot customized subject insertion by amplifying prompt and reference attention and reweighting attention heads in a pre-trained Flux-Fill diffusion transformer.
Discussion (0). Continue with ORCID to comment.