REVIEW 2 cited by
InstaFace: Identity-Preserving Facial Editing with Single Image Inference
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Facial appearance editing is crucial for digital avatars, AR/VR, and personalized content creation, driving realistic user experiences. However, preserving identity with generative models is challenging, especially in scenarios with limited data availability. Traditional methods often require multiple images and still struggle with unnatural face shifts, inconsistent hair alignment, or excessive smoothing effects. To overcome these challenges, we introduce a novel diffusion-based framework, InstaFace, to generate realistic images while preserving identity using only a single image. Central to InstaFace, we introduce an efficient guidance network that harnesses 3D perspectives by integrating multiple 3DMM-based conditionals without introducing additional trainable parameters. Moreover, to ensure maximum identity retention as well as preservation of background, hair, and other contextual features like accessories, we introduce a novel module that utilizes feature embeddings from a facial recognition model and a pre-trained vision-language model. Quantitative evaluations demonstrate that our method outperforms several state-of-the-art approaches in terms of identity preservation, photorealism, and effective control of pose, expression, and lighting.
Forward citations
Cited by 2 Pith papers
-
AvatarBack: Back-Head Generation for Complete 3D Avatars from Front-View Images
AvatarBack adds a generative back-head prior and a learned spatial alignment to Gaussian-splatting head avatars, improving rear geometry and texture while keeping frontal quality.
-
AvatarMakeup: Realistic Makeup Transfer for 3D Animatable Head Avatars
A coarse-to-fine pipeline transfers makeup from one reference image to an animatable 3D Gaussian avatar, using UV-map averaging for cross-view consistency and diffusion refinement for detail.
Discussion (0). Sign in to comment.