Pith. sign in

REVIEW 9 cited by

Arc2Face: A Foundation Model for ID-Consistent Human Faces

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2403.11641 v2 pith:HTZQO4XL submitted 2024-03-18 cs.CV

classification cs.CV
keywords facemodelarc2facefeaturesimagesmodelsachievearcface
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

This paper presents Arc2Face, an identity-conditioned face foundation model, which, given the ArcFace embedding of a person, can generate diverse photo-realistic images with an unparalleled degree of face similarity than existing models. Despite previous attempts to decode face recognition features into detailed images, we find that common high-resolution datasets (e.g. FFHQ) lack sufficient identities to reconstruct any subject. To that end, we meticulously upsample a significant portion of the WebFace42M database, the largest public dataset for face recognition (FR). Arc2Face builds upon a pretrained Stable Diffusion model, yet adapts it to the task of ID-to-face generation, conditioned solely on ID vectors. Deviating from recent works that combine ID with text embeddings for zero-shot personalization of text-to-image models, we emphasize on the compactness of FR features, which can fully capture the essence of the human face, as opposed to hand-crafted prompts. Crucially, text-augmented models struggle to decouple identity and text, usually necessitating some description of the given face to achieve satisfactory similarity. Arc2Face, however, only needs the discriminative features of ArcFace to guide the generation, offering a robust prior for a plethora of tasks where ID consistency is of paramount importance. As an example, we train a FR model on synthetic images from our model and achieve superior performance to existing synthetic datasets.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 9 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Investigation of Accuracy and Bias in Face Recognition Trained with Synthetic Data

    cs.CV 2025-07 conditional novelty 6.0 of 10

    Balanced synthetic face data from Stable Diffusion v3.5 reduces racial bias in face recognition models but does not yet match real-data accuracy on hard benchmarks.

  2. Omni-ID: Holistic Identity Representation Designed for Generative Tasks

    cs.CV 2024-12 conditional novelty 6.0 of 10

    Omni-ID is a fixed-size, multi-view face representation trained with few-to-many reconstruction that reports higher identity preservation than ArcFace and CLIP in face generation and personalized text-to-image tasks.

  3. VariFace: Fair and Diverse Synthetic Dataset Generation for Face Recognition

    cs.CV 2024-12 conditional novelty 6.0 of 10

    VariFace synthetic face datasets, generated with CLIP-based labels, Vendi-score diversity guidance, and divergence-score conditioning, achieve state-of-the-art face verification accuracy and, when scaled to 6M images,...

  4. ControlFace: Harnessing Facial Parametric Control for Face Rigging

    cs.CV 2024-12 conditional novelty 6.0 of 10

    ControlFace performs zero-shot face rigging from 3DMM renderings using a dual-branch U-Net, a control mixer module, and reference control guidance, and reports the best average DECA re-inference error on FFHQ baselines.

  5. HyperFace: Generating Synthetic Face Recognition Datasets by Exploring Face Embedding Hypersphere

    cs.CV 2024-11 conditional novelty 6.0 of 10

    A face-recognition dataset generator places synthetic identities as an optimized packing on the face embedding hypersphere, then renders them with a diffusion generator, training models to competitive accuracy on real...

  6. Vec2Face+ for Face Dataset Generation

    cs.CV 2025-07 conditional novelty 5.0 of 10

    A synthetic face dataset with 4M to 12M images trains a matcher whose average accuracy on five benchmarks is 0.09 to 0.14 points higher than CASIA-WebFace, while twin verification and bias remain unsolved.

  7. Inverting Black-Box Face Recognition Systems via Zero-Order Optimization in Eigenface Space

    cs.CV 2025-06 conditional novelty 5.0 of 10

    DarkerBB reconstructs recognizable color faces from a face recognition system using only similarity scores, via zero-order optimization in a PCA eigenface space.

  8. KAN See Your Face

    cs.CV 2024-11 conditional novelty 5.0 of 10

    A KAN-based mapping network can translate embeddings from privacy-preserving face recognition systems into the input space of a face diffusion model, enabling face reconstruction attacks with high attack success rates.

  9. LoRA Diffusion: Zero-Shot LoRA Synthesis for Diffusion Model Personalization

    cs.LG 2024-12 reject novelty 3.0 of 10

    A VAE plus diffusion hypernetwork synthesizes Stable Diffusion LoRAs for faces from ArcFace embeddings, aiming for zero-shot personalization without per-user fine-tuning.

Pith tools