Search papers, labs, and topics across Lattice.
This paper summarizes the AFMFR competition, which focused on adapting the CLIP ViT-L/14 foundation model for face recognition using synthetic training data generated by IDPERTURB. Participants competed in two tracks, with the Full Data Track showcasing significant improvements in model performance through full fine-tuning, while the Limited Data Track highlighted effective adaptation strategies under resource constraints. The results indicate that synthetic data can enhance face recognition capabilities, with the leading solutions outperforming baseline models across multiple benchmarks and demonstrating fairness across demographic groups.
Adapting foundation models with synthetic data can dramatically enhance face recognition accuracy, with some methods even surpassing traditional baselines.
This paper presents a summary of the Competition on Adapting Foundation Models for Face Recognition Using Synthetic Training Data (AFMFR), held at the 2026 International Joint Conference on Biometrics (IJCB 2026). The competition received a total of eight valid submissions from four distinct teams across two complementary tracks: a Full Data Track, in which participants adapt the CLIP ViT-L/14 foundation model using large-scale synthetic identity data, and a Limited Data Track, designed to reflect more resource-constrained adaptation regimes. All training data was generated exclusively using IDPERTURB. Submitted solutions are ranked based on verification and identification performance across a diverse suite of benchmarks, including LFW, CFP-FP, AgeDB-30, CALFW, CPLFW, IJB-B, IJB-C, and TinyFace, using the Borda count method. Fairness evaluation is additionally conducted on the RFW dataset across four demographic groups. The results demonstrate that adaptation of the CLIP foundation model with synthetic training data substantially improves over the off-the-shelf model and, in several cases, surpasses the baseline. Notably, full fine-tuning with Sub-Center ArcFace (DMSTI-Neurotechnology) leads the Full Data Track, while rank-stabilized LoRA adaptation (Idiap-BSP) proves most effective under limited-data conditions.