Skip to content
Preprint

Identity-Conditioned Latent Consistency Distillation for Face Synthesis

Aug 2026 · 0 citations · 23 references
Computer Science

TL;DR

This work shows that identity-conditioned face synthesis can be performed at a substantially lower computational cost by a latent Consistency Model with few iterations, without compromising image quality for large-scale synthetic face generation.

Abstract

Diffusion models have achieved strong results in high-fidelity image synthesis, but their iterative sampling process makes large-scale generation computationally expensive. This limitation is especially relevant when generating synthetic face datasets for face recognition, where a large number of subjects with many samples in different poses, expressions, ages, etc., are required. In this work, we show that identity-conditioned face synthesis can be performed at a substantially lower computational cost by a latent Consistency Model with few iterations, without compromising image quality. For training, we distill knowledge from the foundation Diffusion Model Arc2Face (teacher) by adapting its original text-to-image pipeline to an embedding-to-face setting, replacing textual prompts with ArcFace identity embeddings. Our distilled model (student) generates identity-conditioned face images with an average inference time of 0.4819 seconds per image, compared with 2.102 seconds for Arc2Face, resulting in a 4.36$\times$ speed-up. Quantitative results, based on FID scores, show that the distilled model remains competitive with Arc2Face across all evaluation protocols. On 100k generated images, it achieves near-parity on CelebA (13.921 vs. 12.928) and outperforms the teacher on WebFace42M (9.317 vs. 9.802). Further evaluations on Synth-500 and AgeDB show a moderate performance gap for the former but comparable results for the latter. These results indicate that Arc2Face can be accelerated through task-specific latent consistency distillation while preserving high image quality for large-scale synthetic face generation. Our proposal is publicly available at https://github.com/UFPR-IPASP-PR/FaceRec-IdentityConsistency.

View source

Similar papers

Preprint Aug 2026

When Diffusion Models Forget Who You Are: Identity Preservation in Face Inpainting under Large Occlusions

Face inpainting with diffusion models has recently achieved impressive visual quality, yet preserving identity fidelity under significant occlusion and conflicting text guidance remains a major challenge. To address this issue, we present Reference Semantic Inpainting for Face (ReSem-Face), a cascaded diffusion framewo...

Feng Ding, Shuhuai Xie, Yue Zhou et al. · 0 citations
Open access 2026

Thermal Face Recognition From Synthetic Data: FLUX.1 Kontext Image-to-Image Adaptation and Cross-Sensor Evaluation

Thermal face recognition (FR) is well suited to surveillance under nighttime and adverse illumination, but its progress is constrained by the scarcity of large, annotated thermal face datasets. This work investigates whether a generative foundation model can be adapted with very few samples to produce identity-preservi...

Gabriel Hermosilla, Ashley Jara, Gabriel Olmos et al. · 0 citations
Preprint Aug 2026

XYZFlow:Scaling Multi dimensional Shortcut Flows for Efficient Generative Modeling

High-fidelity image generation faces a trade-off between speed and quality. Diffusion models produce strong visuals but require costly iterative sampling. Existing efficient methods mainly distill pretrained models into few-step samplers, a challenging process that depends heavily on teacher-model quality. In this pape...

Jin-Xiu Liu, Xuan Liu, Kang-Fu Mei et al. · 0 citations
Preprint Sep 2026

Persistent Identity Preservation in Generative Image Models: A Benchmark and Evaluation System

Generative image models can now produce high-quality images, follow complex instructions, and support precise edits, but they still struggle to preserve who or what is being depicted. When generating or editing images of a specific subject, identity may drift as the pose, expression, appearance, viewpoint, or surroundi...

Meng-Wei Ren, Xuan-Er Zhang, Zhi-Hao Xia · 0 citations
Jul 2026

Sketch-text-driven rectified flow for identity-preserving 4D face generation

Sketch-text-driven rectified flow (STDRF), a conditional rectified-flow framework for identity-preserving and semantically controllable 4D face generation, and a sketch encoder enhanced by Geometric Contour and Texture Detail preprocessing and MixStyle domain adaptation are proposed.

Baodong Wang, Fang Liu, Wei Cao et al. · 0 citations
Preprint Aug 2026

Unmasking Face Embeddings: Reading, Rendering and Naming with Foundation Models

This work uses simple pre-computed linear transformations, estimated from paired embeddings alone, to connect existing FR models with off-the-shelf foundation models, exposing face embeddings as semantically and visually rich biometric representations for web-scale foundation models.

Fizza Rubab, Yi-Ying Tong, Arun Ross · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.