REVIEW 4 cited by
Improved StyleGAN Embedding: Where are the Good Latents?
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
StyleGAN is able to produce photorealistic images that are almost indistinguishable from real photos. The reverse problem of finding an embedding for a given image poses a challenge. Embeddings that reconstruct an image well are not always robust to editing operations. In this paper, we address the problem of finding an embedding that both reconstructs images and also supports image editing tasks. First, we introduce a new normalized space to analyze the diversity and the quality of the reconstructed latent codes. This space can help answer the question of where good latent codes are located in latent space. Second, we propose an improved embedding algorithm using a novel regularization method based on our analysis. Finally, we analyze the quality of different embedding algorithms. We compare our results with the current state-of-the-art methods and achieve a better trade-off between reconstruction quality and editing quality.
Forward citations
Cited by 4 Pith papers
-
DIVE: Inverting Conditional Diffusion Models for Discriminative Tasks
A frozen conditional diffusion model can be inverted via gradient-based discrete optimization, plus a learned layout prior, to perform object detection and faster classification without training a discriminative head.
-
Pose and Facial Expression Transfer by using StyleGAN
A self-supervised StyleGAN2 method transfers pose and expression from a source face onto a target identity with near-real-time inference.
-
Stable Flow: Vital Layers for Training-Free Image Editing
An automatic vital-layer selection for FLUX enables training-free, stable text-driven image editing via selective attention injection.
-
MambaStyle: Efficient StyleGAN Inversion for Real Image Editing with State-Space Models
A Mamba state-space model encoder, MambaStyle, inverts real images into StyleGAN's latent space with fewer parameters and faster inference than prior encoders while keeping competitive reconstruction and editing quality.
Discussion (0). Continue with ORCID to comment.