Pith. sign in

REVIEW 2 cited by

Training Deep Face Recognition Systems with Synthetic Data

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 1802.05891 v1 pith:KNW77TWE submitted 2018-02-16 cs.CV

classification cs.CV
keywords datafacesyntheticrecognitiondeepimagesperformancesystems
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Recent advances in deep learning have significantly increased the performance of face recognition systems. The performance and reliability of these models depend heavily on the amount and quality of the training data. However, the collection of annotated large datasets does not scale well and the control over the quality of the data decreases with the size of the dataset. In this work, we explore how synthetically generated data can be used to decrease the number of real-world images needed for training deep face recognition systems. In particular, we make use of a 3D morphable face model for the generation of images with arbitrary amounts of facial identities and with full control over image variations, such as pose, illumination, and background. In our experiments with an off-the-shelf face recognition software we observe the following phenomena: 1) The amount of real training data needed to train competitive deep face recognition systems can be reduced significantly. 2) Combining large-scale real-world data with synthetic data leads to an increased performance. 3) Models trained only on synthetic data with strong variations in pose, illumination, and background perform very well across different datasets even without dataset adaptation. 4) The real-to-virtual performance gap can be closed when using synthetic data for pre-training, followed by fine-tuning with real-world images. 5) There are no observable negative effects of pre-training with synthetic data. Thus, any face recognition system in our experiments benefits from using synthetic face images. The synthetic data generator, as well as all experiments, are publicly available.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Benchmarking Face Recognition without Real Faces

    cs.CV 2026-07 conditional novelty 6.0 of 10

    MorphFace and Vec2Face can replace real photo benchmarks for ranking face recognition models.

  2. Bringing Balance to Hand Shape Classification: Mitigating Data Imbalance Through Generative Models

    cs.CV 2025-07 conditional novelty 6.0 of 10

    Pre-training an EfficientNet classifier on GAN-generated balanced hand images, then fine-tuning on real data, raises accuracy on the imbalanced RWTH handshape benchmark from 80.6% to 85.3%.

Pith tools